初始化项目,由ModelHub XC社区提供模型
Model: QuantFactory/Llama3.1-BestMix-Chem-Einstein-8B-GGUF Source: Original Platform
This commit is contained in:
49
.gitattributes
vendored
Normal file
49
.gitattributes
vendored
Normal file
@@ -0,0 +1,49 @@
|
|||||||
|
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.model filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||||
|
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q2_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:1f21be2159d8c5a58838339dd707052c1f67042abf0c80bcd1569ada06c83a7e
|
||||||
|
size 3179131744
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_L.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:a828db071e09858b0283466893e946cb698bf8106b04ae82d359442f3a5b2538
|
||||||
|
size 4321956704
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:a374c6f56e16738ef7d22e8faf79072d454ff5d4c7b666389bd03bc026237c44
|
||||||
|
size 4018918240
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q3_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:cf5fb91b9caae72fb08fdf22d3fee5d3d4f19cc649e7eb139c537d2d4cf34af3
|
||||||
|
size 3664499552
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:83cbe71722c8fa06c08df860e0b80866ee92e26d106cf2a64278e2568accc51a
|
||||||
|
size 4661212000
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_1.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:eee8ccf25b05993cdd6e0886403f156d879ff85ed6043bbb0c38f4d122d8d976
|
||||||
|
size 5130253152
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:1a53aa7124c731f33b0b616d7c66a6f78c6a133240acd9e3227f1188f743c1ee
|
||||||
|
size 4920734560
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q4_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:817288fc8fcab1b84af0124075a33d4aa1ea392d00f36a2b0c8b2892a835e929
|
||||||
|
size 4692669280
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:016468724556f7d9ab0a1f5dd1bfb0d44d378178145068e4bee0114a9bf1a94e
|
||||||
|
size 5599294304
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_1.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:7577fb7b0b0956c349d53111f46e361919358a5da93c7b10299c25778c1bca65
|
||||||
|
size 6068335456
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:3a0e5cabb269c2574bdda9396a42fcbb77c0f9a384af7cb27b59ab9776b96839
|
||||||
|
size 5732987744
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q5_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:274ad916e6bd3fb96fe9325bcc8b10333e17f7bfa9fccf155bb48f816cc7329e
|
||||||
|
size 5599294304
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q6_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:da69affbd6c0479c368d1a044787c5b8684d07651c1112b45d8df2528b86d08a
|
||||||
|
size 6596006752
|
||||||
3
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf
Normal file
3
Llama3.1-BestMix-Chem-Einstein-8B.Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:86f7c8ffc23d4f34d16813667c1cbc9af3cc756e94a37dfbacad6fe35bc04c2e
|
||||||
|
size 8540771168
|
||||||
150
README.md
Normal file
150
README.md
Normal file
@@ -0,0 +1,150 @@
|
|||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
license: apache-2.0
|
||||||
|
tags:
|
||||||
|
- merge
|
||||||
|
- TIES
|
||||||
|
- Llama3
|
||||||
|
- BestMix
|
||||||
|
- Chemistry
|
||||||
|
- Einstein
|
||||||
|
- instruction-following
|
||||||
|
- conversational
|
||||||
|
- long-form-generation
|
||||||
|
- scientific
|
||||||
|
base_model:
|
||||||
|
- bunnycore/Best-Mix-Llama-3.1-8B
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
[](https://hf.co/QuantFactory)
|
||||||
|
|
||||||
|
|
||||||
|
# QuantFactory/Llama3.1-BestMix-Chem-Einstein-8B-GGUF
|
||||||
|
This is quantized version of [ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B](https://huggingface.co/ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B) created using llama.cpp
|
||||||
|
|
||||||
|
# Original Model Card
|
||||||
|
|
||||||
|
|
||||||
|
|
||||||
|
# **ZeroXClem/Llama3.1-BestMix-Chem-Einstein-8B**
|
||||||
|
|
||||||
|
**Llama3.1-BestMix-Chem-Einstein-8B** is an innovative, meticulously blended model designed to excel in **instruction-following**, **chemistry-focused tasks**, and **long-form conversational generation**. This model fuses the **best qualities** of multiple Llama3-based architectures, making it highly versatile for both general and specialized tasks. 💻🧠✨
|
||||||
|
|
||||||
|
## 🌟 **Family Tree**
|
||||||
|
|
||||||
|
This model is the result of merging the following:
|
||||||
|
|
||||||
|
- [**bunnycore/Best-Mix-Llama-3.1-8B**](https://huggingface.co/bunnycore/Best-Mix-Llama-3.1-8B): A balanced blend of top Llama models, optimized for general performance across reasoning, instruction-following, and math.
|
||||||
|
- [**USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B**](https://huggingface.co/USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B): A model specialized in **scientific knowledge** and **chemistry**, excelling in chemistry benchmarks.
|
||||||
|
- [**Weyaxi/Einstein-v6.1-Llama3-8B**](https://huggingface.co/Weyaxi/Einstein-v6.1-Llama3-8B): Fine-tuned for **long-form generation**, **conversation-heavy tasks**, and optimized with cutting-edge techniques for efficient memory usage and fast performance.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 🧬 **Model Lineage**
|
||||||
|
|
||||||
|
### **A: bunnycore/Best-Mix-Llama-3.1-8B**
|
||||||
|
|
||||||
|
- A masterful **blend** of several Llama3 models like **Aurora_faustus**, **TitanFusion**, and **OpenMath2**.
|
||||||
|
- Provides a **balanced performance** in a variety of tasks such as reasoning, math, and instruction-following.
|
||||||
|
- Key contributor to the **overall versatility** of the merged model.
|
||||||
|
|
||||||
|
### **B: USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B**
|
||||||
|
|
||||||
|
- Specializes in **chemistry** and **scientific knowledge**, outperforming many larger models in **chemistry benchmarks**.
|
||||||
|
- Adds **scientific rigor** and domain-specific expertise to the merged model, making it perfect for scientific and academic tasks.
|
||||||
|
|
||||||
|
### **C: Weyaxi/Einstein-v6.1-Llama3-8B**
|
||||||
|
|
||||||
|
- Fine-tuned on a wide range of **instructive** and **conversational datasets** like **WizardLM**, **Alpaca**, and **ShareGPT**.
|
||||||
|
- Optimized for **long-form text generation** and enhanced with **xformers attention** and **flash attention** techniques for better performance.
|
||||||
|
- Key player in **dialogue-based tasks** and **long conversation generation**.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 🛠️ **Merge Details**
|
||||||
|
|
||||||
|
This model was merged using the **TIES merge method**, ensuring a smooth integration of the key strengths from each contributing model. Here's the configuration used:
|
||||||
|
|
||||||
|
```yaml
|
||||||
|
yaml
|
||||||
|
Copy code
|
||||||
|
models:
|
||||||
|
- model: bunnycore/Best-Mix-Llama-3.1-8B
|
||||||
|
parameters:
|
||||||
|
density: [1, 0.7, 0.5]
|
||||||
|
weight: 1.0
|
||||||
|
|
||||||
|
- model: USTC-KnowledgeComputingLab/Llama3-KALE-LM-Chem-1.5-8B
|
||||||
|
parameters:
|
||||||
|
density: 0.6
|
||||||
|
weight: [0.3, 0.7, 1.0]
|
||||||
|
|
||||||
|
- model: Weyaxi/Einstein-v6.1-Llama3-8B
|
||||||
|
parameters:
|
||||||
|
density: 0.4
|
||||||
|
weight:
|
||||||
|
- filter: mlp
|
||||||
|
value: 0.5
|
||||||
|
- filter: self_attn
|
||||||
|
value: 0.7
|
||||||
|
- value: 0.5
|
||||||
|
|
||||||
|
merge_method: ties
|
||||||
|
base_model: bunnycore/Best-Mix-Llama-3.1-8B
|
||||||
|
parameters:
|
||||||
|
normalize: true
|
||||||
|
int8_mask: true
|
||||||
|
dtype: float16
|
||||||
|
|
||||||
|
```
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 🎯 **Key Features & Capabilities**
|
||||||
|
|
||||||
|
### **1. Instruction Following & General Reasoning**:
|
||||||
|
|
||||||
|
With the foundation of **Best-Mix**, this model excels in **general-purpose reasoning**, instruction-following, and tasks that require high adaptability.
|
||||||
|
|
||||||
|
### **2. Scientific & Chemistry Expertise**:
|
||||||
|
|
||||||
|
Thanks to the contribution from **KALE-LM-Chem**, this model shines in **scientific research**, particularly **chemistry-focused tasks**, making it ideal for academic and research purposes.
|
||||||
|
|
||||||
|
### **3. Long-Form & Conversational Mastery**:
|
||||||
|
|
||||||
|
With **Einstein-v6.1**, the model handles **long-form generation** effortlessly, excelling in extended conversations and structured dialogue applications.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 🚀 **Performance Benchmarks**
|
||||||
|
|
||||||
|
While still in its early stages, **Llama3.1-BestMix-Chem-Einstein-8B** is expected to perform well across a variety of benchmarks, including:
|
||||||
|
|
||||||
|
- **Chemistry-focused benchmarks** (KALE-LM-Chem)
|
||||||
|
- **Instruction-following tasks** (Best-Mix)
|
||||||
|
- **Conversational AI** and **long-form text generation** (Einstein-v6.1)
|
||||||
|
|
||||||
|
Further testing and evaluation will continue to refine this model's capabilities.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 📜 **License**
|
||||||
|
|
||||||
|
This model is open-sourced under the **Apache-2.0 License**, allowing free use and modification with proper attribution.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 💡 **Tags**
|
||||||
|
|
||||||
|
- `merge`
|
||||||
|
- `TIES`
|
||||||
|
- `BestMix`
|
||||||
|
- `Chemistry`
|
||||||
|
- `Einstein`
|
||||||
|
- `instruction-following`
|
||||||
|
- `long-form-generation`
|
||||||
|
- `conversational`
|
||||||
|
|
||||||
|
---
|
||||||
1
configuration.json
Normal file
1
configuration.json
Normal file
@@ -0,0 +1 @@
|
|||||||
|
{"framework": "pytorch", "task": "others", "allow_remote": true}
|
||||||
Reference in New Issue
Block a user