初始化项目,由ModelHub XC社区提供模型

Model: mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-20 12:17:10 +08:00
commit 234af217c2
26 changed files with 256 additions and 0 deletions

59
.gitattributes vendored Normal file
View File

@@ -0,0 +1,59 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.imatrix.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q2_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f206ee91a33bc2b96bf90f5960d7da9af63a6ce2a2a073bf0f160ae322b8ab72
size 7959871776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d4f9b665de3441674d2270b984020e67355cb5b4c0ca9d6f0e28f61ea63eaf7a
size 7323844896

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:20cd1d0a40f8edace1f8d5f6a1127d44d9904501636fb210a0c5e7fb7134489b
size 11362864416

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c09acfe1840a391fd5edb664c05f957af0031a9af64649906efcd37b3c11782d
size 10514828576

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:03c367e59e155fa52879cb38758c15ad41159d7a8f8244a2eaef25de66406d26
size 9951838496

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7cbcfb382ddb371206824ca8549233b7c1451e8c7e685a637f98a6b26d2f4cbe
size 9019916576

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0e742840a97feb54127c26dd92736239f35b113f0e3f639e6993eec1f70fe208
size 14930086176

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4626b0565376b389f4409f9bad41713649b237d75f62d95669c29480a82e23e2
size 14434306336

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:eac5f535216f97daa737224df59a62487c33da5201d284e1eb003dc1bc29822c
size 13702924576

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3899d736a16f8a8ab619ec89b41d06a594e21f76c52d9c9d503feb5ad791959f
size 12821040416

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4b6e15230151b1c84a53861929580950a165bf456692cbe3967cf60a125c4f57
size 17690498336

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:16894ab093fc4ef26bf3b23cdfa6dadbc2a312e0076198ef046f16b0e59a54f1
size 12344655136

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ca8bf8adbd1103e7ea132c3060c9e9f064dd6b11a0c5a51fae65a01a2d790c43
size 11465817376

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c957da97dec015bf6d28f6fa2b3e58abe5527e48230d6bf93b7912b5847b4be3
size 17330997536

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:855d3e51ceb1984078c86e7610457d9cb7ed18bf070cbffbac93544208e5fa5c
size 15971780896

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:abd1fabfe391563e0a28b89e26af005f897f0b5c6a01b6eecdc717cfe6b884ef
size 14389741856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:18ec9b3f5ca48139f87bf48e1e537e1e93aa97f4e95193f70c1ca08fb8cd8685
size 18703090976

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:157ed8037277d6dbddf4f098154ef8e9e184cdd20eb9170991ee3aff237ba5ae
size 20636525856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:294662d14675cf5ca5d948e8d9ddba03e0ccf2dbd6710533a0ecec18bf85d606
size 19762152736

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5c8a38ad41a48d2e488fae5589686a78d1e8a9e543ca30e02ae7c4733e9e65b1
size 18771248416

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3bbffbb0e38ff1d46de61e6f771110fed561d3c550280af90f982cd4bbe0f371
size 23214834976

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b9652bb07816af2c1dc7c4ef308496dd22d9dd70d0ac2b3f05bc194184944802
size 22635496736

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f0e7cf9297520df10502776cf186ffd343842306b1ec3ec048902c9ee5e23d7f
size 26883309856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:171d1d4b530aa813e92f47b0a3f6076251333036404b63b4966c44935b62156b
size 15273216

125
README.md Normal file
View File

@@ -0,0 +1,125 @@
---
base_model: DavidAU/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking
datasets:
- TeichAI/glm-4.7-2000x
language:
- en
- zh
library_name: transformers
license: apache-2.0
mradermacher:
readme_rev: 1
quantized_by: mradermacher
tags:
- GLM 4.7 Flash distill
- unsloth
- thinking
- reasoning
- heretic
- uncensored
- abliterated
- thinking
- reasoning
- deep reasoning
- fine tune
- creative
- creative writing
- fiction writing
- plot generation
- sub-plot generation
- fiction writing
- story generation
- scene continue
- storytelling
- fiction story
- science fiction
- romance
- all genres
- story
- writing
- vivid prosing
- vivid writing
- fiction
- roleplaying
- bfloat16
- swearing
- rp
- horror
- r rated
- x rated
- all use cases
---
## About
<!-- ### quantize_version: 2 -->
<!-- ### output_tensor_quantised: 1 -->
<!-- ### convert_type: hf -->
<!-- ### vocab_type: -->
<!-- ### tags: nicoboss -->
<!-- ### quants: Q2_K IQ3_M Q4_K_S IQ3_XXS Q3_K_M small-IQ4_NL Q4_K_M IQ2_M Q6_K IQ4_XS Q2_K_S IQ1_M Q3_K_S IQ2_XXS Q3_K_L IQ2_XS Q5_K_S IQ2_S IQ1_S Q5_K_M Q4_0 IQ3_XS Q4_1 IQ3_S -->
<!-- ### quants_skip: -->
<!-- ### skip_mmproj: -->
weighted/imatrix quants of https://huggingface.co/DavidAU/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking
<!-- provided-files -->
***For a convenient overview and download list, visit our [model page for this model](https://hf.tst.eu/model#Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF).***
static quants are available at https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-GGUF
## Usage
If you are unsure how to use GGUF files, refer to one of [TheBloke's
READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for
more details, including on how to concatenate multi-part files.
## Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|:-----|:-----|--------:|:------|
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.imatrix.gguf) | imatrix | 0.1 | imatrix file (for creating your own quants) |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ1_S.gguf) | i1-IQ1_S | 7.4 | for the desperate |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ1_M.gguf) | i1-IQ1_M | 8.1 | mostly desperate |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 9.1 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_XS.gguf) | i1-IQ2_XS | 10.1 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_S.gguf) | i1-IQ2_S | 10.6 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ2_M.gguf) | i1-IQ2_M | 11.5 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q2_K_S.gguf) | i1-Q2_K_S | 11.6 | very low quality |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q2_K.gguf) | i1-Q2_K | 12.4 | IQ3_XXS probably better |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_XXS.gguf) | i1-IQ3_XXS | 12.9 | lower quality |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_XS.gguf) | i1-IQ3_XS | 13.8 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_S.gguf) | i1-Q3_K_S | 14.5 | IQ3_XS probably better |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_S.gguf) | i1-IQ3_S | 14.5 | beats Q3_K* |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ3_M.gguf) | i1-IQ3_M | 15.0 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_M.gguf) | i1-Q3_K_M | 16.1 | IQ3_S probably better |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q3_K_L.gguf) | i1-Q3_K_L | 17.4 | IQ3_M probably better |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-IQ4_XS.gguf) | i1-IQ4_XS | 17.8 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_0.gguf) | i1-Q4_0 | 18.8 | fast, low quality |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_K_S.gguf) | i1-Q4_K_S | 18.9 | optimal size/speed/quality |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_K_M.gguf) | i1-Q4_K_M | 19.9 | fast, recommended |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q4_1.gguf) | i1-Q4_1 | 20.7 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q5_K_S.gguf) | i1-Q5_K_S | 22.7 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q5_K_M.gguf) | i1-Q5_K_M | 23.3 | |
| [GGUF](https://huggingface.co/mradermacher/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking-i1-GGUF/resolve/main/Qwen3-32B-VL-GLM-4.7-Flash-HI16-Heretic-Uncensored-Thinking.i1-Q6_K.gguf) | i1-Q6_K | 27.0 | practically like static Q6_K |
Here is a handy graph by ikawrakow comparing some lower-quality quant
types (lower is better):
![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png)
And here are Artefact2's thoughts on the matter:
https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9
## FAQ / Model Request
See https://huggingface.co/mradermacher/model_requests for some answers to
questions you might have and/or if you want some other model quantized.
## Thanks
I thank my company, [nethype GmbH](https://www.nethype.de/), for letting
me use its servers and providing upgrades to my workstation to enable
this work in my free time. Additional thanks to [@nicoboss](https://huggingface.co/nicoboss) for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.
<!-- end -->