commit 0d2535a4bf53d3bbfec14be04e95a8f59f5b65a0 Author: ModelHub XC Date: Wed May 27 09:36:17 2026 +0800 初始化项目,由ModelHub XC社区提供模型 Model: mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF Source: Original Platform diff --git a/.gitattributes b/.gitattributes new file mode 100644 index 0000000..cd9c52e --- /dev/null +++ b/.gitattributes @@ -0,0 +1,60 @@ +*.7z filter=lfs diff=lfs merge=lfs -text +*.arrow filter=lfs diff=lfs merge=lfs -text +*.bin filter=lfs diff=lfs merge=lfs -text +*.bz2 filter=lfs diff=lfs merge=lfs -text +*.ckpt filter=lfs diff=lfs merge=lfs -text +*.ftz filter=lfs diff=lfs merge=lfs -text +*.gz filter=lfs diff=lfs merge=lfs -text +*.h5 filter=lfs diff=lfs merge=lfs -text +*.joblib filter=lfs diff=lfs merge=lfs -text +*.lfs.* filter=lfs diff=lfs merge=lfs -text +*.mlmodel filter=lfs diff=lfs merge=lfs -text +*.model filter=lfs diff=lfs merge=lfs -text +*.msgpack filter=lfs diff=lfs merge=lfs -text +*.npy filter=lfs diff=lfs merge=lfs -text +*.npz filter=lfs diff=lfs merge=lfs -text +*.onnx filter=lfs diff=lfs merge=lfs -text +*.ot filter=lfs diff=lfs merge=lfs -text +*.parquet filter=lfs diff=lfs merge=lfs -text +*.pb filter=lfs diff=lfs merge=lfs -text +*.pickle filter=lfs diff=lfs merge=lfs -text +*.pkl filter=lfs diff=lfs merge=lfs -text +*.pt filter=lfs diff=lfs merge=lfs -text +*.pth filter=lfs diff=lfs merge=lfs -text +*.rar filter=lfs diff=lfs merge=lfs -text +*.safetensors filter=lfs diff=lfs merge=lfs -text +saved_model/**/* filter=lfs diff=lfs merge=lfs -text +*.tar.* filter=lfs diff=lfs merge=lfs -text +*.tar filter=lfs diff=lfs merge=lfs -text +*.tflite filter=lfs diff=lfs merge=lfs -text +*.tgz filter=lfs diff=lfs merge=lfs -text +*.wasm filter=lfs diff=lfs merge=lfs -text +*.xz filter=lfs diff=lfs merge=lfs -text +*.zip filter=lfs diff=lfs merge=lfs -text +*.zst filter=lfs diff=lfs merge=lfs -text +*tfevents* filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.imatrix.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_1.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_M.gguf new file mode 100644 index 0000000..30b9d61 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:c770562e776eb1fe40c2fec66a2fe58cb66b770e3d7f454678c0229d56e91189 +size 1127019456 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_S.gguf new file mode 100644 index 0000000..014da6b --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:2b25c933eae2f3b449cb45e87249d4db2bab4ce001ce4f73a9b4a17e10024ca4 +size 1055257536 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_M.gguf new file mode 100644 index 0000000..19cee77 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b5b090d66aac89f1e7c3b4f46c1b5097ad30e646309adbe107ae8ea0f9aa76af +size 1512985536 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_S.gguf new file mode 100644 index 0000000..a698003 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:6c7ae582ca0edd6b1dccc3271ba1685fadbca1a67a1980702580f2aba074a302 +size 1417302976 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XS.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XS.gguf new file mode 100644 index 0000000..e8d2e6a --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:3190d7d10d0e8a60a78828e5e48f5425c164c801b50af03edb7067ff91ffaca5 +size 1354101696 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XXS.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XXS.gguf new file mode 100644 index 0000000..861579c --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XXS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:080e751286c6cf1369a969db2ae93aabb6f454c939af4201ad5c2cf4c67c117c +size 1246622656 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_M.gguf new file mode 100644 index 0000000..eecb401 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:2a987f718277c38a286cce38e732235436cf90e6df255dfa8afe62d393689250 +size 1962897856 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_S.gguf new file mode 100644 index 0000000..3f90070 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:383793846a92b960658272e4c382a20ed46c7086ca51478dbf06d421dcdaf662 +size 1899532736 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XS.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XS.gguf new file mode 100644 index 0000000..e1fe6a7 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:98e59ee7be9411c86e4149ef775e98f3c22191b1b19e8edc9932121c390daf53 +size 1814376896 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XXS.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XXS.gguf new file mode 100644 index 0000000..2cd3c48 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XXS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:1399e461436ab9a4f73701b807ab4ab6397320aa5b48815104929a522e42432f +size 1670190016 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_NL.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_NL.gguf new file mode 100644 index 0000000..ca5f9e6 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_NL.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:fbcdd84c3f43058719f04f38c42d9c3edba255aa05bde71cd4325578a2d1fdf2 +size 2381345216 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_XS.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_XS.gguf new file mode 100644 index 0000000..4d06ded --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:e2feec0461fe87d73976358318d5434cfabf0ff7223f4d7bdee6b203617101ff +size 2270753216 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K.gguf new file mode 100644 index 0000000..4b1592d --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:444ebbea3035519542a6f2cc24233e9f5db54ee7090f354b50cc3166010d723c +size 1669501376 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K_S.gguf new file mode 100644 index 0000000..b3fbe8f --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:9955ed8ed86f96f8aafceb78f0de631f1e0df364814b8c8d7ae752b66c299b63 +size 1563455936 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_L.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_L.gguf new file mode 100644 index 0000000..ccd5e04 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_L.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:20e43b39d9cde020daacfcc849ff777079c85ddcade03f397ba279ba728a1fe4 +size 2239787456 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_M.gguf new file mode 100644 index 0000000..56ef2ad --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:d8e777d10bd3612198fe8df9563418febcbac960a67d4867ec6ae0f7a05a682f +size 2075619776 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_S.gguf new file mode 100644 index 0000000..6bb27e7 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:027de066f357f6abc90526c1be3b1191cf78da52dae683248ed2d22ec8be5b9a +size 1886998976 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_0.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_0.gguf new file mode 100644 index 0000000..6c1f316 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_0.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:6960bbde4a536755e3e8996522c9eae7ad7dadad679b5ebf700232c55abdebd1 +size 2375774656 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_1.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_1.gguf new file mode 100644 index 0000000..94d607d --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_1.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:cff50f694727a22ca9773a0eb39a108307c849f21b44f26496b14be003fc4a38 +size 2596630976 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_M.gguf new file mode 100644 index 0000000..9d6bf86 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:abbf40c9eefc04244a7d23b5f9f6e37a3dc3b97f76cd080994c83a034e4e86e4 +size 2497282496 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_S.gguf new file mode 100644 index 0000000..1ec6a2d --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:7f403df9331f3ee8e9f53d673519d8ced9afa729bb890a9e1c8716f4db8574d7 +size 2383311296 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_M.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_M.gguf new file mode 100644 index 0000000..32f4e44 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:56f4f4b0b72047ae1c36c5e0340858a92e12b9e6b66495ba52ef9dd5792c666e +size 2889515456 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_S.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_S.gguf new file mode 100644 index 0000000..3045a6a --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:797034dd6efee44c4530ad81bed7a43ce2f5f0f0903926ad0f5f0921f8a261ff +size 2823713216 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q6_K.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q6_K.gguf new file mode 100644 index 0000000..e4ee555 --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q6_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b2963cd197861d2c3e5692f842561aa1438e5b711314502bd1f852d3b8555741 +size 3306262976 diff --git a/Qwen3-4B-Instruct-2507-NanoWriter-v2.imatrix.gguf b/Qwen3-4B-Instruct-2507-NanoWriter-v2.imatrix.gguf new file mode 100644 index 0000000..2dd369c --- /dev/null +++ b/Qwen3-4B-Instruct-2507-NanoWriter-v2.imatrix.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:9fd22d2aa47aa911f9d1d53cf3c46c093eea229fa358116f7f0aca7c17245e29 +size 3872640 diff --git a/README.md b/README.md new file mode 100644 index 0000000..e8bddf7 --- /dev/null +++ b/README.md @@ -0,0 +1,101 @@ +--- +base_model: 0xA50C1A1/Qwen3-4B-Instruct-2507-NanoWriter-v2 +datasets: +- Gryphe/Opus-WritingPrompts +- Gryphe/ChatGPT-4o-Writing-Prompts +- 0xA50C1A1/Storytelling-DPO +language: +- en +library_name: transformers +license: apache-2.0 +mradermacher: + readme_rev: 1 +quantized_by: mradermacher +tags: +- text-generation-inference +- transformers +- unsloth +- qwen3 +- heretic +- uncensored +- decensored +- abliterated +- creative-writing +- storytelling +- writing +--- +## About + + + + + + + + + +weighted/imatrix quants of https://huggingface.co/0xA50C1A1/Qwen3-4B-Instruct-2507-NanoWriter-v2 + + + +***For a convenient overview and download list, visit our [model page for this model](https://hf.tst.eu/model#Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF).*** + +static quants are available at https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-GGUF +## Usage + +If you are unsure how to use GGUF files, refer to one of [TheBloke's +READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for +more details, including on how to concatenate multi-part files. + +## Provided Quants + +(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants) + +| Link | Type | Size/GB | Notes | +|:-----|:-----|--------:|:------| +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.imatrix.gguf) | imatrix | 0.1 | imatrix file (for creating your own quants) | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_S.gguf) | i1-IQ1_S | 1.2 | for the desperate | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ1_M.gguf) | i1-IQ1_M | 1.2 | mostly desperate | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 1.3 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_XS.gguf) | i1-IQ2_XS | 1.5 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_S.gguf) | i1-IQ2_S | 1.5 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ2_M.gguf) | i1-IQ2_M | 1.6 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K_S.gguf) | i1-Q2_K_S | 1.7 | very low quality | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q2_K.gguf) | i1-Q2_K | 1.8 | IQ3_XXS probably better | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XXS.gguf) | i1-IQ3_XXS | 1.8 | lower quality | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_XS.gguf) | i1-IQ3_XS | 1.9 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_S.gguf) | i1-Q3_K_S | 2.0 | IQ3_XS probably better | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_S.gguf) | i1-IQ3_S | 2.0 | beats Q3_K* | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ3_M.gguf) | i1-IQ3_M | 2.1 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_M.gguf) | i1-Q3_K_M | 2.2 | IQ3_S probably better | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q3_K_L.gguf) | i1-Q3_K_L | 2.3 | IQ3_M probably better | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_XS.gguf) | i1-IQ4_XS | 2.4 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_0.gguf) | i1-Q4_0 | 2.5 | fast, low quality | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-IQ4_NL.gguf) | i1-IQ4_NL | 2.5 | prefer IQ4_XS | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_S.gguf) | i1-Q4_K_S | 2.5 | optimal size/speed/quality | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_K_M.gguf) | i1-Q4_K_M | 2.6 | fast, recommended | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q4_1.gguf) | i1-Q4_1 | 2.7 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_S.gguf) | i1-Q5_K_S | 2.9 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q5_K_M.gguf) | i1-Q5_K_M | 3.0 | | +| [GGUF](https://huggingface.co/mradermacher/Qwen3-4B-Instruct-2507-NanoWriter-v2-i1-GGUF/resolve/main/Qwen3-4B-Instruct-2507-NanoWriter-v2.i1-Q6_K.gguf) | i1-Q6_K | 3.4 | practically like static Q6_K | + +Here is a handy graph by ikawrakow comparing some lower-quality quant +types (lower is better): + +![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png) + +And here are Artefact2's thoughts on the matter: +https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9 + +## FAQ / Model Request + +See https://huggingface.co/mradermacher/model_requests for some answers to +questions you might have and/or if you want some other model quantized. + +## Thanks + +I thank my company, [nethype GmbH](https://www.nethype.de/), for letting +me use its servers and providing upgrades to my workstation to enable +this work in my free time. Additional thanks to [@nicoboss](https://huggingface.co/nicoboss) for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to. + +