初始化项目,由ModelHub XC社区提供模型
Model: QuantFactory/Llama3.1-BestMix-Chem-Einstein-8B-GGUF Source: Original Platform
This commit is contained in:
49
.gitattributes
vendored
Normal file
49
.gitattributes
vendored
Normal file
@@ -0,0 +1,49 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:1f21be2159d8c5a58838339dd707052c1f67042abf0c80bcd1569ada06c83a7e
|
||||
size 3179131744
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a828db071e09858b0283466893e946cb698bf8106b04ae82d359442f3a5b2538
|
||||
size 4321956704
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a374c6f56e16738ef7d22e8faf79072d454ff5d4c7b666389bd03bc026237c44
|
||||
size 4018918240
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:cf5fb91b9caae72fb08fdf22d3fee5d3d4f19cc649e7eb139c537d2d4cf34af3
|
||||
size 3664499552
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:83cbe71722c8fa06c08df860e0b80866ee92e26d106cf2a64278e2568accc51a
|
||||
size 4661212000
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:eee8ccf25b05993cdd6e0886403f156d879ff85ed6043bbb0c38f4d122d8d976
|
||||
size 5130253152
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:1a53aa7124c731f33b0b616d7c66a6f78c6a133240acd9e3227f1188f743c1ee
|
||||
size 4920734560
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:817288fc8fcab1b84af0124075a33d4aa1ea392d00f36a2b0c8b2892a835e929
|
||||
size 4692669280
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:016468724556f7d9ab0a1f5dd1bfb0d44d378178145068e4bee0114a9bf1a94e
|
||||
size 5599294304
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:7577fb7b0b0956c349d53111f46e361919358a5da93c7b10299c25778c1bca65
|
||||
size 6068335456
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:3a0e5cabb269c2574bdda9396a42fcbb77c0f9a384af7cb27b59ab9776b96839
|
||||
size 5732987744
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:274ad916e6bd3fb96fe9325bcc8b10333e17f7bfa9fccf155bb48f816cc7329e
|
||||
size 5599294304
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:da69affbd6c0479c368d1a044787c5b8684d07651c1112b45d8df2528b86d08a
|
||||
size 6596006752
|
||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:86f7c8ffc23d4f34d16813667c1cbc9af3cc756e94a37dfbacad6fe35bc04c2e
|
||||
size 8540771168
|
||||
150
README.md
Normal file
150
README.md
Normal file
@@ -0,0 +1,150 @@
|
||||
|
||||
---
|
||||
|
||||
license: apache-2.0
|
||||
tags:
|
||||
- merge
|
||||
- TIES
|
||||
- Llama3
|
||||
- BestMix
|
||||
- Chemistry
|
||||
- Einstein
|
||||
- instruction-following
|
||||
- conversational
|
||||
- long-form-generation
|
||||
- scientific
|
||||
base_model:
|
||||
- bunnycore/Best-Mix-Llama-3.1-8B
|
||||
|
||||
---
|
||||
|
||||
[](https://hf.co/QuantFactory)
|
||||
|
||||
|
||||
# QuantFactory/Llama3.1-BestMix-Chem-Einstein-8B-GGUF
|
||||
This is quantized version of [ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B](https://huggingface.co/ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B) created using llama.cpp
|
||||
|
||||
# Original Model Card
|
||||
|
||||
|
||||
|
||||
# **ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B**
|
||||
|
||||
**Llama3.1-BestMix-Chem-Einstein-8B** is an innovative, meticulously blended model designed to excel in **instruction-following**, **chemistry-focused tasks**, and **long-form conversational generation**. This model fuses the **best qualities** of multiple Llama3-based architectures, making it highly versatile for both general and specialized tasks. 💻🧠✨
|
||||
|
||||
## 🌟 **Family Tree**
|
||||
|
||||
This model is the result of merging the following:
|
||||
|
||||
- [**bunnycore/Best-Mix-Llama-3.1-8B**](https://huggingface.co/bunnycore/Best-Mix-Llama-3.1-8B): A balanced blend of top Llama models, optimized for general performance across reasoning, instruction-following, and math.
|
||||
- [**USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B**](https://huggingface.co/USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B): A model specialized in **scientific knowledge** and **chemistry**, excelling in chemistry benchmarks.
|
||||
- [**Weyaxi/Einstein-v6.1-Llama3-8B**](https://huggingface.co/Weyaxi/Einstein-v6.1-Llama3-8B): Fine-tuned for **long-form generation**, **conversation-heavy tasks**, and optimized with cutting-edge techniques for efficient memory usage and fast performance.
|
||||
|
||||
---
|
||||
|
||||
## 🧬 **Model Lineage**
|
||||
|
||||
### **A: bunnycore/Best-Mix-Llama-3.1-8B**
|
||||
|
||||
- A masterful **blend** of several Llama3 models like **Aurora_faustus**, **TitanFusion**, and **OpenMath2**.
|
||||
- Provides a **balanced performance** in a variety of tasks such as reasoning, math, and instruction-following.
|
||||
- Key contributor to the **overall versatility** of the merged model.
|
||||
|
||||
### **B: USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B**
|
||||
|
||||
- Specializes in **chemistry** and **scientific knowledge**, outperforming many larger models in **chemistry benchmarks**.
|
||||
- Adds **scientific rigor** and domain-specific expertise to the merged model, making it perfect for scientific and academic tasks.
|
||||
|
||||
### **C: Weyaxi/Einstein-v6.1-Llama3-8B**
|
||||
|
||||
- Fine-tuned on a wide range of **instructive** and **conversational datasets** like **WizardLM**, **Alpaca**, and **ShareGPT**.
|
||||
- Optimized for **long-form text generation** and enhanced with **xformers attention** and **flash attention** techniques for better performance.
|
||||
- Key player in **dialogue-based tasks** and **long conversation generation**.
|
||||
|
||||
---
|
||||
|
||||
## 🛠️ **Merge Details**
|
||||
|
||||
This model was merged using the **TIES merge method**, ensuring a smooth integration of the key strengths from each contributing model. Here's the configuration used:
|
||||
|
||||
```yaml
|
||||
yaml
|
||||
Copy code
|
||||
models:
|
||||
- model: bunnycore/Best-Mix-Llama-3.1-8B
|
||||
parameters:
|
||||
density: [1, 0.7, 0.5]
|
||||
weight: 1.0
|
||||
|
||||
- model: USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B
|
||||
parameters:
|
||||
density: 0.6
|
||||
weight: [0.3, 0.7, 1.0]
|
||||
|
||||
- model: Weyaxi/Einstein-v6.1-Llama3-8B
|
||||
parameters:
|
||||
density: 0.4
|
||||
weight:
|
||||
- filter: mlp
|
||||
value: 0.5
|
||||
- filter: self_attn
|
||||
value: 0.7
|
||||
- value: 0.5
|
||||
|
||||
merge_method: ties
|
||||
base_model: bunnycore/Best-Mix-Llama-3.1-8B
|
||||
parameters:
|
||||
normalize: true
|
||||
int8_mask: true
|
||||
dtype: float16
|
||||
|
||||
```
|
||||
|
||||
---
|
||||
|
||||
## 🎯 **Key Features & Capabilities**
|
||||
|
||||
### **1. Instruction Following & General Reasoning**:
|
||||
|
||||
With the foundation of **Best-Mix**, this model excels in **general-purpose reasoning**, instruction-following, and tasks that require high adaptability.
|
||||
|
||||
### **2. Scientific & Chemistry Expertise**:
|
||||
|
||||
Thanks to the contribution from **KALE-LM-Chem**, this model shines in **scientific research**, particularly **chemistry-focused tasks**, making it ideal for academic and research purposes.
|
||||
|
||||
### **3. Long-Form & Conversational Mastery**:
|
||||
|
||||
With **Einstein-v6.1**, the model handles **long-form generation** effortlessly, excelling in extended conversations and structured dialogue applications.
|
||||
|
||||
---
|
||||
|
||||
## 🚀 **Performance Benchmarks**
|
||||
|
||||
While still in its early stages, **Llama3.1-BestMix-Chem-Einstein-8B** is expected to perform well across a variety of benchmarks, including:
|
||||
|
||||
- **Chemistry-focused benchmarks** (KALE-LM-Chem)
|
||||
- **Instruction-following tasks** (Best-Mix)
|
||||
- **Conversational AI** and **long-form text generation** (Einstein-v6.1)
|
||||
|
||||
Further testing and evaluation will continue to refine this model's capabilities.
|
||||
|
||||
---
|
||||
|
||||
## 📜 **License**
|
||||
|
||||
This model is open-sourced under the **Apache-2.0 License**, allowing free use and modification with proper attribution.
|
||||
|
||||
---
|
||||
|
||||
## 💡 **Tags**
|
||||
|
||||
- `merge`
|
||||
- `TIES`
|
||||
- `BestMix`
|
||||
- `Chemistry`
|
||||
- `Einstein`
|
||||
- `instruction-following`
|
||||
- `long-form-generation`
|
||||
- `conversational`
|
||||
|
||||
---
|
||||
1
configuration.json
Normal file
1
configuration.json
Normal file
@@ -0,0 +1 @@
|
||||
{"framework": "pytorch", "task": "others", "allow_remote": true}
|
||||
Reference in New Issue
Block a user