commit 17ab12d3a2030850e87a46c3563e5853e3a8e636 Author: ModelHub XC Date: Sat Jun 13 18:19:19 2026 +0800 初始化项目,由ModelHub XC社区提供模型 Model: mradermacher/Llama-3.2-3B-Renoia-i1-GGUF Source: Original Platform diff --git a/.gitattributes b/.gitattributes new file mode 100644 index 0000000..2ec43a4 --- /dev/null +++ b/.gitattributes @@ -0,0 +1,60 @@ +*.7z filter=lfs diff=lfs merge=lfs -text +*.arrow filter=lfs diff=lfs merge=lfs -text +*.bin filter=lfs diff=lfs merge=lfs -text +*.bz2 filter=lfs diff=lfs merge=lfs -text +*.ckpt filter=lfs diff=lfs merge=lfs -text +*.ftz filter=lfs diff=lfs merge=lfs -text +*.gz filter=lfs diff=lfs merge=lfs -text +*.h5 filter=lfs diff=lfs merge=lfs -text +*.joblib filter=lfs diff=lfs merge=lfs -text +*.lfs.* filter=lfs diff=lfs merge=lfs -text +*.mlmodel filter=lfs diff=lfs merge=lfs -text +*.model filter=lfs diff=lfs merge=lfs -text +*.msgpack filter=lfs diff=lfs merge=lfs -text +*.npy filter=lfs diff=lfs merge=lfs -text +*.npz filter=lfs diff=lfs merge=lfs -text +*.onnx filter=lfs diff=lfs merge=lfs -text +*.ot filter=lfs diff=lfs merge=lfs -text +*.parquet filter=lfs diff=lfs merge=lfs -text +*.pb filter=lfs diff=lfs merge=lfs -text +*.pickle filter=lfs diff=lfs merge=lfs -text +*.pkl filter=lfs diff=lfs merge=lfs -text +*.pt filter=lfs diff=lfs merge=lfs -text +*.pth filter=lfs diff=lfs merge=lfs -text +*.rar filter=lfs diff=lfs merge=lfs -text +*.safetensors filter=lfs diff=lfs merge=lfs -text +saved_model/**/* filter=lfs diff=lfs merge=lfs -text +*.tar.* filter=lfs diff=lfs merge=lfs -text +*.tar filter=lfs diff=lfs merge=lfs -text +*.tflite filter=lfs diff=lfs merge=lfs -text +*.tgz filter=lfs diff=lfs merge=lfs -text +*.wasm filter=lfs diff=lfs merge=lfs -text +*.xz filter=lfs diff=lfs merge=lfs -text +*.zip filter=lfs diff=lfs merge=lfs -text +*.zst filter=lfs diff=lfs merge=lfs -text +*tfevents* filter=lfs diff=lfs merge=lfs -text +imatrix.dat filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q2_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q4_1.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama-3.2-3B-Renoia.i1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text diff --git a/Llama-3.2-3B-Renoia.i1-IQ1_M.gguf b/Llama-3.2-3B-Renoia.i1-IQ1_M.gguf new file mode 100644 index 0000000..b79a7cf --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ1_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b62c69473032f4571f24a20aada225f7e69354d6e2f5db49b2d14fba4d355c57 +size 924194880 diff --git a/Llama-3.2-3B-Renoia.i1-IQ1_S.gguf b/Llama-3.2-3B-Renoia.i1-IQ1_S.gguf new file mode 100644 index 0000000..fbeaebd --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ1_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:89d608b1a2283984ec7fecbbea08d873f6d9e565e9792369836f308fcd0c0e3d +size 868161600 diff --git a/Llama-3.2-3B-Renoia.i1-IQ2_M.gguf b/Llama-3.2-3B-Renoia.i1-IQ2_M.gguf new file mode 100644 index 0000000..e6bbb3b --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ2_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:c88236f981468470e61b1a18c4059f19d7b30cdb03e2ac11b23f70891c648ff1 +size 1229035584 diff --git a/Llama-3.2-3B-Renoia.i1-IQ2_S.gguf b/Llama-3.2-3B-Renoia.i1-IQ2_S.gguf new file mode 100644 index 0000000..fe97a72 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ2_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:ceff5bde0f829e86e088447c9596269dfd018a090c137f01e6d9b613be964339 +size 1154324544 diff --git a/Llama-3.2-3B-Renoia.i1-IQ2_XS.gguf b/Llama-3.2-3B-Renoia.i1-IQ2_XS.gguf new file mode 100644 index 0000000..be8d5bf --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ2_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:d720bac8d87c9c2981b7e07fb9407d121263d837f44dab94bc0782ac6abf1345 +size 1100552256 diff --git a/Llama-3.2-3B-Renoia.i1-IQ2_XXS.gguf b/Llama-3.2-3B-Renoia.i1-IQ2_XXS.gguf new file mode 100644 index 0000000..32026fc --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ2_XXS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:885ce47feef5cf805abb14ccc7c4a24216f1a19a92c8a41ef7e7a620d9a8846a +size 1017583680 diff --git a/Llama-3.2-3B-Renoia.i1-IQ3_M.gguf b/Llama-3.2-3B-Renoia.i1-IQ3_M.gguf new file mode 100644 index 0000000..2c897f0 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ3_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:c0d274165a3e33a79679dc0814c09ad0192be238cb6b271fbf4e30f0ea895c34 +size 1599672384 diff --git a/Llama-3.2-3B-Renoia.i1-IQ3_S.gguf b/Llama-3.2-3B-Renoia.i1-IQ3_S.gguf new file mode 100644 index 0000000..dd03abe --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ3_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:011f5f61bd64ea66380c8da35ba21890be61521a0249b70bcdbc052cad2f5dc7 +size 1542852672 diff --git a/Llama-3.2-3B-Renoia.i1-IQ3_XS.gguf b/Llama-3.2-3B-Renoia.i1-IQ3_XS.gguf new file mode 100644 index 0000000..658f70f --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ3_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b66f1bf36740f5660265ea09e05410d847f5051d824efdf84aa3aa9c6ada9dad +size 1476792384 diff --git a/Llama-3.2-3B-Renoia.i1-IQ3_XXS.gguf b/Llama-3.2-3B-Renoia.i1-IQ3_XXS.gguf new file mode 100644 index 0000000..9f288ef --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ3_XXS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:d81fbafc8f38614e1b6658ed0be6be00eb131492c59122e2b7a89b4551ec112e +size 1348769856 diff --git a/Llama-3.2-3B-Renoia.i1-IQ4_NL.gguf b/Llama-3.2-3B-Renoia.i1-IQ4_NL.gguf new file mode 100644 index 0000000..4f01cce --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ4_NL.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:8bdb0c2458262b4cacd7b22890596be2f68f6232c8fc183eda0592aa50811b7c +size 1917194304 diff --git a/Llama-3.2-3B-Renoia.i1-IQ4_XS.gguf b/Llama-3.2-3B-Renoia.i1-IQ4_XS.gguf new file mode 100644 index 0000000..8cf92d4 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-IQ4_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:12e3cf978e38b02a4773c52f7221cc7ce0de5a3ae6c76d02897e355931a32ea4 +size 1829113920 diff --git a/Llama-3.2-3B-Renoia.i1-Q2_K.gguf b/Llama-3.2-3B-Renoia.i1-Q2_K.gguf new file mode 100644 index 0000000..83e1574 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q2_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:92376b68bf56ebd11dbbb48e8c566c68c513a5fd2b6d5c92fec23a55963e37bc +size 1363939392 diff --git a/Llama-3.2-3B-Renoia.i1-Q2_K_S.gguf b/Llama-3.2-3B-Renoia.i1-Q2_K_S.gguf new file mode 100644 index 0000000..29a0b58 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q2_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b60a8f5c7c64c1b9f4a8234c406011e7b5c60d34080095c057251989d56896c7 +size 1274286144 diff --git a/Llama-3.2-3B-Renoia.i1-Q3_K_L.gguf b/Llama-3.2-3B-Renoia.i1-Q3_K_L.gguf new file mode 100644 index 0000000..db03942 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q3_K_L.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:a5199bcb6f461f714dcecf4ab15291627806ca9cf49ff0672c3dee8a78a19a48 +size 1815351360 diff --git a/Llama-3.2-3B-Renoia.i1-Q3_K_M.gguf b/Llama-3.2-3B-Renoia.i1-Q3_K_M.gguf new file mode 100644 index 0000000..2084eba --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q3_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:5b6a3372eae1e76a4faaf602964c175e172259662df6ced3b1ade78772eb558e +size 1687162944 diff --git a/Llama-3.2-3B-Renoia.i1-Q3_K_S.gguf b/Llama-3.2-3B-Renoia.i1-Q3_K_S.gguf new file mode 100644 index 0000000..672ce61 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q3_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:7db5f82b6fd559c954e1dc9963c1503e89f8c0751dcbcf25cba5a5b478d73bb1 +size 1542852672 diff --git a/Llama-3.2-3B-Renoia.i1-Q4_0.gguf b/Llama-3.2-3B-Renoia.i1-Q4_0.gguf new file mode 100644 index 0000000..195f9b2 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q4_0.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:4cfb2000777df5be35cb0e6d9ee5c303a1dc6933fe446571632462cd6f492acd +size 1921912896 diff --git a/Llama-3.2-3B-Renoia.i1-Q4_1.gguf b/Llama-3.2-3B-Renoia.i1-Q4_1.gguf new file mode 100644 index 0000000..6d55a52 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q4_1.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:f9f4d1cf55a17104d0bc5875a44ea35bd2790352b2a17139c69f9a01eb86f7c2 +size 2093355072 diff --git a/Llama-3.2-3B-Renoia.i1-Q4_K_M.gguf b/Llama-3.2-3B-Renoia.i1-Q4_K_M.gguf new file mode 100644 index 0000000..95cd6fb --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q4_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:e175a29471f3ec0dcf26853fb2a96a88f90a095a730f7c8a1313cca30335d8ad +size 2019381312 diff --git a/Llama-3.2-3B-Renoia.i1-Q4_K_S.gguf b/Llama-3.2-3B-Renoia.i1-Q4_K_S.gguf new file mode 100644 index 0000000..e00dfa1 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q4_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:22f587f409f0ea2532e5823756ac05be111dc2029ed53a223c4dc96ebafee741 +size 1928204352 diff --git a/Llama-3.2-3B-Renoia.i1-Q5_K_M.gguf b/Llama-3.2-3B-Renoia.i1-Q5_K_M.gguf new file mode 100644 index 0000000..2cc7061 --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q5_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:eb16e5cfdcbedcacf53dda862d67cb06a6740e4e21552fa8892f7097110c49fc +size 2322157632 diff --git a/Llama-3.2-3B-Renoia.i1-Q5_K_S.gguf b/Llama-3.2-3B-Renoia.i1-Q5_K_S.gguf new file mode 100644 index 0000000..4d510fe --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q5_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:2ee9c37351fc69a64ae4de564f7e72393bc44a7089604ed03f60692fe21a14f9 +size 2269515840 diff --git a/Llama-3.2-3B-Renoia.i1-Q6_K.gguf b/Llama-3.2-3B-Renoia.i1-Q6_K.gguf new file mode 100644 index 0000000..7f89a0d --- /dev/null +++ b/Llama-3.2-3B-Renoia.i1-Q6_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:5c87d3d5d5b95510e997463676bea6454314948c52ecf622e9e6158c2b641fdb +size 2643857472 diff --git a/README.md b/README.md new file mode 100644 index 0000000..c6b6b80 --- /dev/null +++ b/README.md @@ -0,0 +1,89 @@ +--- +base_model: trollek/Llama-3.2-3B-Renoia +datasets: +- trollek/Danoia-v03 +- trollek/Danoia-v02 +- trollek/ProbingPanoia-v01 +- WhiteRabbitNeo/WRN-Chapter-1 +- WhiteRabbitNeo/WRN-Chapter-2 +- migtissera/Trinity-2-v0.2-10K +- trollek/Panoia-v02 +- trollek/Danoia-v01 +language: +- da +- en +library_name: transformers +license: llama3.2 +quantized_by: mradermacher +tags: +- mergekit +- merge +--- +## About + + + + + + +weighted/imatrix quants of https://huggingface.co/trollek/Llama-3.2-3B-Renoia + + +static quants are available at https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-GGUF +## Usage + +If you are unsure how to use GGUF files, refer to one of [TheBloke's +READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for +more details, including on how to concatenate multi-part files. + +## Provided Quants + +(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants) + +| Link | Type | Size/GB | Notes | +|:-----|:-----|--------:|:------| +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ1_S.gguf) | i1-IQ1_S | 1.0 | for the desperate | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ1_M.gguf) | i1-IQ1_M | 1.0 | mostly desperate | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 1.1 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ2_XS.gguf) | i1-IQ2_XS | 1.2 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ2_S.gguf) | i1-IQ2_S | 1.3 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ2_M.gguf) | i1-IQ2_M | 1.3 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q2_K_S.gguf) | i1-Q2_K_S | 1.4 | very low quality | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ3_XXS.gguf) | i1-IQ3_XXS | 1.4 | lower quality | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q2_K.gguf) | i1-Q2_K | 1.5 | IQ3_XXS probably better | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ3_XS.gguf) | i1-IQ3_XS | 1.6 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ3_S.gguf) | i1-IQ3_S | 1.6 | beats Q3_K* | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q3_K_S.gguf) | i1-Q3_K_S | 1.6 | IQ3_XS probably better | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ3_M.gguf) | i1-IQ3_M | 1.7 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q3_K_M.gguf) | i1-Q3_K_M | 1.8 | IQ3_S probably better | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q3_K_L.gguf) | i1-Q3_K_L | 1.9 | IQ3_M probably better | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ4_XS.gguf) | i1-IQ4_XS | 1.9 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-IQ4_NL.gguf) | i1-IQ4_NL | 2.0 | prefer IQ4_XS | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q4_0.gguf) | i1-Q4_0 | 2.0 | fast, low quality | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q4_K_S.gguf) | i1-Q4_K_S | 2.0 | optimal size/speed/quality | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q4_K_M.gguf) | i1-Q4_K_M | 2.1 | fast, recommended | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q4_1.gguf) | i1-Q4_1 | 2.2 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q5_K_S.gguf) | i1-Q5_K_S | 2.4 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q5_K_M.gguf) | i1-Q5_K_M | 2.4 | | +| [GGUF](https://huggingface.co/mradermacher/Llama-3.2-3B-Renoia-i1-GGUF/resolve/main/Llama-3.2-3B-Renoia.i1-Q6_K.gguf) | i1-Q6_K | 2.7 | practically like static Q6_K | + +Here is a handy graph by ikawrakow comparing some lower-quality quant +types (lower is better): + +![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png) + +And here are Artefact2's thoughts on the matter: +https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9 + +## FAQ / Model Request + +See https://huggingface.co/mradermacher/model_requests for some answers to +questions you might have and/or if you want some other model quantized. + +## Thanks + +I thank my company, [nethype GmbH](https://www.nethype.de/), for letting +me use its servers and providing upgrades to my workstation to enable +this work in my free time. Additional thanks to [@nicoboss](https://huggingface.co/nicoboss) for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to. + + diff --git a/imatrix.dat b/imatrix.dat new file mode 100644 index 0000000..1c93470 --- /dev/null +++ b/imatrix.dat @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:214c87aabc3da623feaffb3ffe9de81445c35947622b7b3c6707b767e94e0086 +size 2988377