初始化项目,由ModelHub XC社区提供模型
Model: liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF Source: Original Platform
This commit is contained in:
42
.gitattributes
vendored
Normal file
42
.gitattributes
vendored
Normal file
@@ -0,0 +1,42 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Qwen2.5-Coder-7B-Instruct-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
3
Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:49b697287016eb0ffcb3268cab336322028427f3fbab0f623d2979b62417b7d4
|
||||
size 2780343104
|
||||
3
Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:c562a3bdfb052f410ba993ea426866ef04e09ee9dd0485cb0586258dbb8bc2ca
|
||||
size 3574012736
|
||||
3
Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:0a5e1a2238fb65dc13b21aeba5bb2b207446591f28775d4b11a0f0327c1c1a13
|
||||
size 4218473280
|
||||
3
Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:30a7a77286ed27f1a2702853acf636c85bc91f0ec2844919c9516f0e579e75b6
|
||||
size 4683074368
|
||||
3
Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:8f2d74ed7c32ed36378bd1084ab282e54e2dfa22b11f99d44a4853f74fb2074d
|
||||
size 5444832064
|
||||
3
Qwen2.5-Coder-7B-Instruct-Q6_K.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-Q6_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:f86e5216c6d76c598092c016782be24fabba2dd662a530423734006a404ed258
|
||||
size 6254199616
|
||||
3
Qwen2.5-Coder-7B-Instruct-Q8_0.gguf
Normal file
3
Qwen2.5-Coder-7B-Instruct-Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:690422ec614f1b21cddebde2fd68f38678c7d3686b1817e7e131bc1f18fcaeef
|
||||
size 8098526016
|
||||
69
README.md
Normal file
69
README.md
Normal file
@@ -0,0 +1,69 @@
|
||||
---
|
||||
license: other
|
||||
base_model: Qwen/Qwen2.5-Coder-7B-Instruct
|
||||
base_model_relation: quantized
|
||||
pipeline_tag: text-generation
|
||||
library_name: gguf
|
||||
tags:
|
||||
- gguf
|
||||
- ollama
|
||||
- local-llm
|
||||
- llama.cpp
|
||||
- lm-studio
|
||||
- quantized
|
||||
- imatrix
|
||||
- sub-4-bit
|
||||
- qwen2
|
||||
- codeqwen
|
||||
quantized_by: liodon-ai
|
||||
---
|
||||
|
||||
# Qwen2.5-Coder-7B-Instruct — iMatrix GGUF
|
||||
|
||||
GGUF quantizations of [Qwen/Qwen2.5-Coder-7B-Instruct](https://huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct), published by [Liodon AI](https://huggingface.co/liodon-ai).
|
||||
|
||||
## Quick Start
|
||||
|
||||
**llama.cpp**
|
||||
```bash
|
||||
llama-cli -hf liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF:Q4_K_M
|
||||
```
|
||||
|
||||
**Ollama**
|
||||
```bash
|
||||
ollama run hf.co/liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF:Q4_K_M
|
||||
```
|
||||
|
||||
**LM Studio / Jan** — search `liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF` and pick your quant.
|
||||
|
||||
## Quants
|
||||
|
||||
| Quant | Size | VRAM est. | Notes |
|
||||
|-------|------|-----------|-------|
|
||||
| `IQ2_M` | 2.78 GB | ~3 GB | 2-bit, iMatrix — smallest usable |
|
||||
| `IQ3_M` | 3.57 GB | ~4 GB | 3-bit, iMatrix — great quality/size tradeoff |
|
||||
| `IQ4_XS` | 4.22 GB | ~5 GB | 4-bit extra-small, iMatrix |
|
||||
| `Q4_K_M` | 4.68 GB | ~5 GB | 4-bit, iMatrix-calibrated (recommended) |
|
||||
| `Q5_K_M` | 5.44 GB | ~6 GB | 5-bit, iMatrix-calibrated |
|
||||
| `Q6_K` | 6.25 GB | ~7 GB | 6-bit, iMatrix-calibrated, near-lossless |
|
||||
| `Q8_0` | 8.10 GB | ~9 GB | 8-bit, essentially lossless |
|
||||
|
||||
|
||||
## What is iMatrix?
|
||||
|
||||
Standard quantization treats all weights equally. iMatrix runs 128 calibration chunks through
|
||||
the full-precision model to find which weights matter most, then allocates more precision where
|
||||
it counts. At Q2/Q3/Q4 this means noticeably better coherence and instruction-following —
|
||||
**same file size, better output**.
|
||||
|
||||
Calibration: 2M tokens of [WikiText-103](https://huggingface.co/datasets/wikitext).
|
||||
|
||||
> Also see plain (non-iMatrix) quants: `liodon-ai/Qwen2.5-Coder-7B-Instruct-GGUF`
|
||||
|
||||
## Source
|
||||
|
||||
- **Model**: [Qwen/Qwen2.5-Coder-7B-Instruct](https://huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct)
|
||||
- **License**: other
|
||||
|
||||
---
|
||||
*Quantized by [Liodon AI](https://huggingface.co/liodon-ai)*
|
||||
Reference in New Issue
Block a user