commit 23a1c646288c287cf6cabf5f987a5c0a716be319 Author: ModelHub XC Date: Fri Sep 4 19:50:16 2026 +0800 初始化项目,由ModelHub XC社区提供模型 Model: liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF Source: Original Platform diff --git a/.gitattributes b/.gitattributes new file mode 100644 index 0000000..daf6388 --- /dev/null +++ b/.gitattributes @@ -0,0 +1,42 @@ +*.7z filter=lfs diff=lfs merge=lfs -text +*.arrow filter=lfs diff=lfs merge=lfs -text +*.bin filter=lfs diff=lfs merge=lfs -text +*.bz2 filter=lfs diff=lfs merge=lfs -text +*.ckpt filter=lfs diff=lfs merge=lfs -text +*.ftz filter=lfs diff=lfs merge=lfs -text +*.gz filter=lfs diff=lfs merge=lfs -text +*.h5 filter=lfs diff=lfs merge=lfs -text +*.joblib filter=lfs diff=lfs merge=lfs -text +*.lfs.* filter=lfs diff=lfs merge=lfs -text +*.mlmodel filter=lfs diff=lfs merge=lfs -text +*.model filter=lfs diff=lfs merge=lfs -text +*.msgpack filter=lfs diff=lfs merge=lfs -text +*.npy filter=lfs diff=lfs merge=lfs -text +*.npz filter=lfs diff=lfs merge=lfs -text +*.onnx filter=lfs diff=lfs merge=lfs -text +*.ot filter=lfs diff=lfs merge=lfs -text +*.parquet filter=lfs diff=lfs merge=lfs -text +*.pb filter=lfs diff=lfs merge=lfs -text +*.pickle filter=lfs diff=lfs merge=lfs -text +*.pkl filter=lfs diff=lfs merge=lfs -text +*.pt filter=lfs diff=lfs merge=lfs -text +*.pth filter=lfs diff=lfs merge=lfs -text +*.rar filter=lfs diff=lfs merge=lfs -text +*.safetensors filter=lfs diff=lfs merge=lfs -text +saved_model/**/* filter=lfs diff=lfs merge=lfs -text +*.tar.* filter=lfs diff=lfs merge=lfs -text +*.tar filter=lfs diff=lfs merge=lfs -text +*.tflite filter=lfs diff=lfs merge=lfs -text +*.tgz filter=lfs diff=lfs merge=lfs -text +*.wasm filter=lfs diff=lfs merge=lfs -text +*.xz filter=lfs diff=lfs merge=lfs -text +*.zip filter=lfs diff=lfs merge=lfs -text +*.zst filter=lfs diff=lfs merge=lfs -text +*tfevents* filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text +Qwen2.5-Coder-7B-Instruct-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text diff --git a/Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf b/Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf new file mode 100644 index 0000000..2be7215 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-IQ2_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:49b697287016eb0ffcb3268cab336322028427f3fbab0f623d2979b62417b7d4 +size 2780343104 diff --git a/Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf b/Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf new file mode 100644 index 0000000..7ab8b55 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-IQ3_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:c562a3bdfb052f410ba993ea426866ef04e09ee9dd0485cb0586258dbb8bc2ca +size 3574012736 diff --git a/Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf b/Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf new file mode 100644 index 0000000..1986e11 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-IQ4_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:0a5e1a2238fb65dc13b21aeba5bb2b207446591f28775d4b11a0f0327c1c1a13 +size 4218473280 diff --git a/Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf b/Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf new file mode 100644 index 0000000..786c2a0 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-Q4_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:30a7a77286ed27f1a2702853acf636c85bc91f0ec2844919c9516f0e579e75b6 +size 4683074368 diff --git a/Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf b/Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf new file mode 100644 index 0000000..6d82a68 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-Q5_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:8f2d74ed7c32ed36378bd1084ab282e54e2dfa22b11f99d44a4853f74fb2074d +size 5444832064 diff --git a/Qwen2.5-Coder-7B-Instruct-Q6_K.gguf b/Qwen2.5-Coder-7B-Instruct-Q6_K.gguf new file mode 100644 index 0000000..99d1e31 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-Q6_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:f86e5216c6d76c598092c016782be24fabba2dd662a530423734006a404ed258 +size 6254199616 diff --git a/Qwen2.5-Coder-7B-Instruct-Q8_0.gguf b/Qwen2.5-Coder-7B-Instruct-Q8_0.gguf new file mode 100644 index 0000000..25bbc70 --- /dev/null +++ b/Qwen2.5-Coder-7B-Instruct-Q8_0.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:690422ec614f1b21cddebde2fd68f38678c7d3686b1817e7e131bc1f18fcaeef +size 8098526016 diff --git a/README.md b/README.md new file mode 100644 index 0000000..a35a321 --- /dev/null +++ b/README.md @@ -0,0 +1,69 @@ +--- +license: other +base_model: Qwen/Qwen2.5-Coder-7B-Instruct +base_model_relation: quantized +pipeline_tag: text-generation +library_name: gguf +tags: +- gguf +- ollama +- local-llm +- llama.cpp +- lm-studio +- quantized +- imatrix +- sub-4-bit +- qwen2 +- codeqwen +quantized_by: liodon-ai +--- + +# Qwen2.5-Coder-7B-Instruct — iMatrix GGUF + +GGUF quantizations of [Qwen/Qwen2.5-Coder-7B-Instruct](https://huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct), published by [Liodon AI](https://huggingface.co/liodon-ai). + +## Quick Start + +**llama.cpp** +```bash +llama-cli -hf liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF:Q4_K_M +``` + +**Ollama** +```bash +ollama run hf.co/liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF:Q4_K_M +``` + +**LM Studio / Jan** — search `liodon-ai/Qwen2.5-Coder-7B-Instruct-imatrix-GGUF` and pick your quant. + +## Quants + +| Quant | Size | VRAM est. | Notes | +|-------|------|-----------|-------| +| `IQ2_M` | 2.78 GB | ~3 GB | 2-bit, iMatrix — smallest usable | +| `IQ3_M` | 3.57 GB | ~4 GB | 3-bit, iMatrix — great quality/size tradeoff | +| `IQ4_XS` | 4.22 GB | ~5 GB | 4-bit extra-small, iMatrix | +| `Q4_K_M` | 4.68 GB | ~5 GB | 4-bit, iMatrix-calibrated (recommended) | +| `Q5_K_M` | 5.44 GB | ~6 GB | 5-bit, iMatrix-calibrated | +| `Q6_K` | 6.25 GB | ~7 GB | 6-bit, iMatrix-calibrated, near-lossless | +| `Q8_0` | 8.10 GB | ~9 GB | 8-bit, essentially lossless | + + +## What is iMatrix? + +Standard quantization treats all weights equally. iMatrix runs 128 calibration chunks through +the full-precision model to find which weights matter most, then allocates more precision where +it counts. At Q2/Q3/Q4 this means noticeably better coherence and instruction-following — +**same file size, better output**. + +Calibration: 2M tokens of [WikiText-103](https://huggingface.co/datasets/wikitext). + +> Also see plain (non-iMatrix) quants: `liodon-ai/Qwen2.5-Coder-7B-Instruct-GGUF` + +## Source + +- **Model**: [Qwen/Qwen2.5-Coder-7B-Instruct](https://huggingface.co/Qwen/Qwen2.5-Coder-7B-Instruct) +- **License**: other + +--- +*Quantized by [Liodon AI](https://huggingface.co/liodon-ai)*