初始化项目,由ModelHub XC社区提供模型
Model: mradermacher/Khaosara-7B-Instruct-i1-GGUF Source: Original Platform
This commit is contained in:
60
.gitattributes
vendored
Normal file
60
.gitattributes
vendored
Normal file
@@ -0,0 +1,60 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.imatrix.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q2_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Khaosara-7B-Instruct.i1-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
3
Khaosara-7B-Instruct.i1-IQ1_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ1_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:ab3fdcdf589557dccc1ad63c0efacd95858bd6f5b240067cff0d0935ad7d4597
|
||||
size 1757664800
|
||||
3
Khaosara-7B-Instruct.i1-IQ1_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ1_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:327ef3312996d24c128065938d5d7bfa68004577ff131d56a09f6a00baf1f40c
|
||||
size 1615320608
|
||||
3
Khaosara-7B-Instruct.i1-IQ2_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ2_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:643daabb2233cbcf1bd680bad763e9edadec835d73f6212ebab223dd20acd500
|
||||
size 2504250912
|
||||
3
Khaosara-7B-Instruct.i1-IQ2_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ2_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:0b8081e3845ce4691f69b1b0dec49b5911395ef936a3b4613724f263742e453b
|
||||
size 2314458656
|
||||
3
Khaosara-7B-Instruct.i1-IQ2_XS.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ2_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:29caf53b8a0978e572eae9427e31986a84ec92f25ae2ccf025b199750fff1f56
|
||||
size 2201474592
|
||||
3
Khaosara-7B-Instruct.i1-IQ2_XXS.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ2_XXS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:6665e5ef2e699a200eb7154b6bd1ed16e7a9140c25325fac5295e175c68fe01c
|
||||
size 1994905120
|
||||
3
Khaosara-7B-Instruct.i1-IQ3_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ3_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:ff5f198bebbb3701670daa01b3464879126faba365a75a9f193ea9508c88f6b6
|
||||
size 3288847904
|
||||
3
Khaosara-7B-Instruct.i1-IQ3_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ3_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:dc8755fac47b0f2fbdb98cda51359455245dc10137ca1cd1b96ada652d8ee8b3
|
||||
size 3186349600
|
||||
3
Khaosara-7B-Instruct.i1-IQ3_XS.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ3_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:e64e88373fba97065f5796e5f17d7752d24b1cec745a1fae488f53ba67472231
|
||||
size 3022771744
|
||||
3
Khaosara-7B-Instruct.i1-IQ3_XXS.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ3_XXS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:267af36613123214961266bb1009f8cae84acda1aa96b1bc9cd3fe93711a548d
|
||||
size 2830882336
|
||||
3
Khaosara-7B-Instruct.i1-IQ4_NL.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ4_NL.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:65d5625afa8c61b9052bbd743c036caf52db15d32d74ff10035fc0926d53655d
|
||||
size 4130068000
|
||||
3
Khaosara-7B-Instruct.i1-IQ4_XS.gguf
Normal file
3
Khaosara-7B-Instruct.i1-IQ4_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:7661bddbada18a82329aebc78af64f1fc6f3c3e2745e48e5b5be34d3c889cb1e
|
||||
size 3911964192
|
||||
3
Khaosara-7B-Instruct.i1-Q2_K.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q2_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:01121970e3365d423650e0c418bd006dc0179e604c4c6d5b59a4c446e8e0c8ce
|
||||
size 2722879008
|
||||
3
Khaosara-7B-Instruct.i1-Q2_K_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q2_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:9ce09c81cd8348ca87bdc01ab62ba7df57c62ca7ab46c3e41cbd66d442269358
|
||||
size 2532562464
|
||||
3
Khaosara-7B-Instruct.i1-Q3_K_L.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q3_K_L.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:742e52af205a208500da8bac183de06a19d3e0a890fc1c065cb378561f0dced7
|
||||
size 3825980960
|
||||
3
Khaosara-7B-Instruct.i1-Q3_K_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q3_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:0dcc23409891dd7df7905bca90a0092bc01540e726e155796abe023c04e108ee
|
||||
size 3522942496
|
||||
3
Khaosara-7B-Instruct.i1-Q3_K_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q3_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:bc91e3b4145cd4df90f2fb1a946c79e4e437c98bf32a822d3f3f81bc539fe4be
|
||||
size 3168523808
|
||||
3
Khaosara-7B-Instruct.i1-Q4_0.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q4_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:7a7fff97289d2ad2abe000ed342312dd7cfcfea1a92dfb193240c6f11796579f
|
||||
size 4127970848
|
||||
3
Khaosara-7B-Instruct.i1-Q4_1.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q4_1.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:3869845e431854eca46d72278352be713906b705add9d02898a00049872b8bd7
|
||||
size 4557887008
|
||||
3
Khaosara-7B-Instruct.i1-Q4_K_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:4baa8f8641e2c63494d471428d0de4c649f52a27e0c1611446404831d99ac8f7
|
||||
size 4372813344
|
||||
3
Khaosara-7B-Instruct.i1-Q4_K_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q4_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:f85b7acf5197a9de6e25ba196aa431220b4eca686c3cb9e414d501053e80fffd
|
||||
size 4144748064
|
||||
3
Khaosara-7B-Instruct.i1-Q5_K_M.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q5_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:b28ad7a0676b9f926711cf592b8ac8b58667b24ef71825ee9631fc3e4257b77f
|
||||
size 5136176672
|
||||
3
Khaosara-7B-Instruct.i1-Q5_K_S.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q5_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:38f25befc6ff56d3a7d2b7d42c9e981e566c5379bb9c5b322ec1bf81b9f79f17
|
||||
size 5002483232
|
||||
3
Khaosara-7B-Instruct.i1-Q6_K.gguf
Normal file
3
Khaosara-7B-Instruct.i1-Q6_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:cb5a68db26e8eafca5a23686a0c49f8cea419857d946d1e01ecec101611b9feb
|
||||
size 5947250208
|
||||
3
Khaosara-7B-Instruct.imatrix.gguf
Normal file
3
Khaosara-7B-Instruct.imatrix.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:b1009f9a807a6a507b1b9f821bb6e2eef610a3cc8d179163d2854eba93a0a7b6
|
||||
size 5015200
|
||||
102
README.md
Normal file
102
README.md
Normal file
@@ -0,0 +1,102 @@
|
||||
---
|
||||
base_model: c4tdr0ut/Khaosara-7B-Instruct
|
||||
language:
|
||||
- en
|
||||
library_name: transformers
|
||||
license: other
|
||||
mradermacher:
|
||||
readme_rev: 1
|
||||
quantized_by: mradermacher
|
||||
tags:
|
||||
- text-generation-inference
|
||||
- transformers
|
||||
- mistral
|
||||
- mistral-7b
|
||||
- khaosara
|
||||
- axolotl
|
||||
- unhinged
|
||||
- x.ai
|
||||
- sota
|
||||
- liger
|
||||
- sft
|
||||
- continual-pretraining
|
||||
- factual
|
||||
- domain-expert
|
||||
- conversational
|
||||
- research
|
||||
---
|
||||
## About
|
||||
|
||||
<!-- ### quantize_version: 2 -->
|
||||
<!-- ### output_tensor_quantised: 1 -->
|
||||
<!-- ### convert_type: hf -->
|
||||
<!-- ### vocab_type: -->
|
||||
<!-- ### tags: nicoboss -->
|
||||
<!-- ### quants: Q2_K IQ3_M Q4_K_S IQ3_XXS Q3_K_M small-IQ4_NL Q4_K_M IQ2_M Q6_K IQ4_XS Q2_K_S IQ1_M Q3_K_S IQ2_XXS Q3_K_L IQ2_XS Q5_K_S IQ2_S IQ1_S Q5_K_M Q4_0 IQ3_XS Q4_1 IQ3_S -->
|
||||
<!-- ### quants_skip: -->
|
||||
<!-- ### skip_mmproj: -->
|
||||
weighted/imatrix quants of https://huggingface.co/c4tdr0ut/Khaosara-7B-Instruct
|
||||
|
||||
<!-- provided-files -->
|
||||
|
||||
***For a convenient overview and download list, visit our [model page for this model](https://hf.tst.eu/model#Khaosara-7B-Instruct-i1-GGUF).***
|
||||
|
||||
static quants are available at https://huggingface.co/mradermacher/Khaosara-7B-Instruct-GGUF
|
||||
## Usage
|
||||
|
||||
If you are unsure how to use GGUF files, refer to one of [TheBloke's
|
||||
READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for
|
||||
more details, including on how to concatenate multi-part files.
|
||||
|
||||
## Provided Quants
|
||||
|
||||
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
|
||||
|
||||
| Link | Type | Size/GB | Notes |
|
||||
|:-----|:-----|--------:|:------|
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.imatrix.gguf) | imatrix | 0.1 | imatrix file (for creating your own quants) |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ1_S.gguf) | i1-IQ1_S | 1.7 | for the desperate |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ1_M.gguf) | i1-IQ1_M | 1.9 | mostly desperate |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 2.1 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ2_XS.gguf) | i1-IQ2_XS | 2.3 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ2_S.gguf) | i1-IQ2_S | 2.4 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ2_M.gguf) | i1-IQ2_M | 2.6 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q2_K_S.gguf) | i1-Q2_K_S | 2.6 | very low quality |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q2_K.gguf) | i1-Q2_K | 2.8 | IQ3_XXS probably better |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ3_XXS.gguf) | i1-IQ3_XXS | 2.9 | lower quality |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ3_XS.gguf) | i1-IQ3_XS | 3.1 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q3_K_S.gguf) | i1-Q3_K_S | 3.3 | IQ3_XS probably better |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ3_S.gguf) | i1-IQ3_S | 3.3 | beats Q3_K* |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ3_M.gguf) | i1-IQ3_M | 3.4 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q3_K_M.gguf) | i1-Q3_K_M | 3.6 | IQ3_S probably better |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q3_K_L.gguf) | i1-Q3_K_L | 3.9 | IQ3_M probably better |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ4_XS.gguf) | i1-IQ4_XS | 4.0 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q4_0.gguf) | i1-Q4_0 | 4.2 | fast, low quality |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-IQ4_NL.gguf) | i1-IQ4_NL | 4.2 | prefer IQ4_XS |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q4_K_S.gguf) | i1-Q4_K_S | 4.2 | optimal size/speed/quality |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q4_K_M.gguf) | i1-Q4_K_M | 4.5 | fast, recommended |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q4_1.gguf) | i1-Q4_1 | 4.7 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q5_K_S.gguf) | i1-Q5_K_S | 5.1 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q5_K_M.gguf) | i1-Q5_K_M | 5.2 | |
|
||||
| [GGUF](https://huggingface.co/mradermacher/Khaosara-7B-Instruct-i1-GGUF/resolve/main/Khaosara-7B-Instruct.i1-Q6_K.gguf) | i1-Q6_K | 6.0 | practically like static Q6_K |
|
||||
|
||||
Here is a handy graph by ikawrakow comparing some lower-quality quant
|
||||
types (lower is better):
|
||||
|
||||

|
||||
|
||||
And here are Artefact2's thoughts on the matter:
|
||||
https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9
|
||||
|
||||
## FAQ / Model Request
|
||||
|
||||
See https://huggingface.co/mradermacher/model_requests for some answers to
|
||||
questions you might have and/or if you want some other model quantized.
|
||||
|
||||
## Thanks
|
||||
|
||||
I thank my company, [nethype GmbH](https://www.nethype.de/), for letting
|
||||
me use its servers and providing upgrades to my workstation to enable
|
||||
this work in my free time. Additional thanks to [@nicoboss](https://huggingface.co/nicoboss) for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.
|
||||
|
||||
<!-- end -->
|
||||
Reference in New Issue
Block a user