初始化项目,由ModelHub XC社区提供模型
Model: omerasim/fehm-8b-v1-GGUF Source: Original Platform
This commit is contained in:
37
.gitattributes
vendored
Normal file
37
.gitattributes
vendored
Normal file
@@ -0,0 +1,37 @@
|
|||||||
|
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.model filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||||
|
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
fehm-8b-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
fehm-8b-v1-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
71
README.md
Normal file
71
README.md
Normal file
@@ -0,0 +1,71 @@
|
|||||||
|
---
|
||||||
|
license: apache-2.0
|
||||||
|
language:
|
||||||
|
- tr
|
||||||
|
- en
|
||||||
|
base_model: Qwen/Qwen3-8B
|
||||||
|
tags:
|
||||||
|
- text-generation
|
||||||
|
- conversational
|
||||||
|
- turkish
|
||||||
|
- nage
|
||||||
|
- conflux
|
||||||
|
- gguf
|
||||||
|
model_name: Fehm-8B
|
||||||
|
pipeline_tag: text-generation
|
||||||
|
---
|
||||||
|
|
||||||
|
# Fehm-8B GGUF — Nage AI
|
||||||
|
|
||||||
|
**Fehm** (Arabic: understanding, deep comprehension) is Nage AI's general assistant model.
|
||||||
|
Fine-tuned from Qwen3-8B using [CONFLUX](https://github.com/NageAI/conflux) cross-architecture knowledge transfer.
|
||||||
|
|
||||||
|
## Downloads
|
||||||
|
|
||||||
|
| File | Quant | Size | Use case |
|
||||||
|
|------|-------|------|----------|
|
||||||
|
| fehm-8b-v1-Q4_K_M.gguf | Q4_K_M | 5.0 GB | Ollama, LM Studio, mobile |
|
||||||
|
| fehm-8b-v1-Q8_0.gguf | Q8_0 | 8.7 GB | High quality inference |
|
||||||
|
|
||||||
|
## Benchmarks
|
||||||
|
|
||||||
|
| Benchmark | Score |
|
||||||
|
|-----------|-------|
|
||||||
|
| MMLU (5-shot) | **74.95%** |
|
||||||
|
| HumanEval (pass@1) | **71.34%** |
|
||||||
|
| SFT val loss | 0.671 (CONFLUX) vs 0.717 (baseline) |
|
||||||
|
|
||||||
|
## CONFLUX Framework
|
||||||
|
|
||||||
|
Trained with [CONFLUX v0.3.0](https://github.com/NageAI/conflux) — cross-architecture knowledge transfer (Llama-3.1-8B source).
|
||||||
|
|
||||||
|
- **6.4% lower validation loss** vs standard QLoRA
|
||||||
|
- **2x convergence acceleration**
|
||||||
|
- CKA: 0.817 | EVR: 0.9946
|
||||||
|
|
||||||
|
## Training
|
||||||
|
|
||||||
|
- **Base:** Qwen3-8B (Apache 2.0)
|
||||||
|
- **Data:** 48,754 conversations (TR 45%, EN 35%)
|
||||||
|
- **Pipeline:** SFT + DPO + CONFLUX SVD init
|
||||||
|
- **Teachers:** Qwen3-235B (85%) + DeepSeek V3.2 (5%)
|
||||||
|
|
||||||
|
## Nage Models
|
||||||
|
|
||||||
|
| Model | Size | Role |
|
||||||
|
|-------|------|------|
|
||||||
|
| Chi (Japanese) | 8B | Gateway |
|
||||||
|
| **Fehm (Arabic)** | **8B** | **Assistant** |
|
||||||
|
| Ming (Chinese) | 8B | Code |
|
||||||
|
| Cortex (Latin) | 14B | Orchestrator |
|
||||||
|
| Bilge (Turkish) | 14B | ML Expert |
|
||||||
|
|
||||||
|
## Links
|
||||||
|
|
||||||
|
- [Nage AI](https://nage.ai)
|
||||||
|
- [CONFLUX](https://github.com/NageAI/conflux)
|
||||||
|
- [FP16 Model](https://huggingface.co/omerasim/fehm-8b-conflux-v1)
|
||||||
|
|
||||||
|
## License
|
||||||
|
|
||||||
|
Apache 2.0
|
||||||
3
fehm-8b-v1-Q4_K_M.gguf
Normal file
3
fehm-8b-v1-Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:614b73c421591cf6a9f95cdfda8f630ccef720ed8270842d253156569744d1ce
|
||||||
|
size 5027783872
|
||||||
3
fehm-8b-v1-Q8_0.gguf
Normal file
3
fehm-8b-v1-Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:5491bfbbfb23261530b6036a6f2091bc73c120d82b1fa80c8db0481c646c01a4
|
||||||
|
size 8709519008
|
||||||
Reference in New Issue
Block a user