初始化项目,由ModelHub XC社区提供模型

Model: bartowski/Bielik-11B-v2.2-Instruct-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-08-03 22:09:13 +08:00
commit 6390ada1af
28 changed files with 268 additions and 0 deletions

60
.gitattributes vendored Normal file
View File

@@ -0,0 +1,60 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q6_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q5_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_0_8_8.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_0_4_8.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_0_4_4.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q3_K_XL.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q2_K_L.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct-f16.gguf filter=lfs diff=lfs merge=lfs -text
Bielik-11B-v2.2-Instruct.imatrix filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c7eac89aefebfd598d85f97ea10f18cc630373507adef46fba10ebe2464b6f84
size 3823567136

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6428afdfa740eef58e26728b02ea0f720f5d76c3c127dc1619697d83fe78096c
size 5038780704

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:41c4fc25dc278ded70b58015b20a952f885830bd7c33e0ed265a1a535db562e6
size 4627738912

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1ae1ea45896a960d88c7f2dc43b18937579182811d00ec92571ca329499bc467
size 6006415648

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5520d85e7706c5a846b8a4de141481faced220c4800b5aa01fa3a79414c01ffe
size 4164337952

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b54e173b2bb08c9fc10d8a6ea75654aee68e5b905ff195cc85edaab9eb1f54a5
size 4292849952

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:908334717b8512f8c742f4005160757ffbb5fe3b3bc648b54fc83c8e647161ef
size 5880000800

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1c907540de4cb2c32ee019a9d2589d49eb0c97aaa8baa5113cb6fa63b3a49ecf
size 5404995872

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0643afd9579d50e648052bfc3dc76b7c676ac32f6cbb57f5b1aec8527d55cc7f
size 4852724000

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:90d620bbe568d8305306feee13f47331dcadf866e2a9728e2fbb2f3d13f21f91
size 5995147552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:941d8c2ffe24504aa507759177243573724b936121adf069f246faa73de02c2b
size 6340567328

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1d0f27fe75ef57e04534af8e4255b4bd01bd58c78a1d4ceb617fbea23e3ae231
size 6318547232

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9cf2a74b78c50e5fb472fdb917b36aecfa5f462a23798b7c50ca9cdf2933c029
size 6318547232

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:35458f89ac594fe13439d25bb4c2664a063dccd4563e70353d231f3ba33aba2d
size 6318547232

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ad49b2199ecbff60a3790b5cb1f44b64d1d37be95f641c4dce07afff4c832ace
size 6821720352

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8b551119b9308172aaa7c1c56bc5f1f83915e6b5911ee6ee3e496d095b3bfd5d
size 6724051232

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3821f2017c83c318e8abcaf0b852e334a8abf5e7695e4b6521c1836e577f1023
size 6364684576

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0e4e025ae137f2b1dd001915b4b3925456313ccea3a9f5a6e1575f9f16887da5
size 7988261152

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0760dadfeae46bf73b1365ba1e858c06b014d3e5c62690ae2b2e90482e135547
size 7907041568

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0ab2b0d295ddd0d39a82d79ebce072340bb6c7fc93b0e20bc6e3f95357e0f6ff
size 7698145568

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:41578163677a14c784ab490252c3d09f0b584f74f227ab3b559c6761f95051c3
size 9163968800

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e2cdd4a5f4f783eaf6ad80bc3f246c608a9f2e35946a1030ade0b52a0c3b48c7
size 9227710752

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2fb26fd0690f5bcdea8ebf8bd364e6bdd54bf2963a1a56d9a6db12b5ddf4ece1
size 11868811552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2ba78bf3c276efd32e2d05210bc811db85c9bc9235e1c61b28c17a1736dbad6d
size 22339170304

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3ba169aba64841c8e1c537f9e41a9cf9175b5efccdabe790eebd9fcb12949567
size 7794028

132
README.md Normal file
View File

@@ -0,0 +1,132 @@
---
base_model: speakleash/Bielik-11B-v2.2-Instruct
language:
- pl
library_name: transformers
license: apache-2.0
pipeline_tag: text-generation
tags:
- finetuned
quantized_by: bartowski
inference:
parameters:
temperature: 0.2
widget:
- messages:
- role: user
content: Co przedstawia polskie godło?
extra_gated_description: If you want to learn more about how you can use the model,
please refer to our <a href="https://bielik.ai/terms/">Terms of Use</a>.
---
## Llamacpp imatrix Quantizations of Bielik-11B-v2.2-Instruct
Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b3634">b3634</a> for quantization.
Original model: https://huggingface.co/speakleash/Bielik-11B-v2.2-Instruct
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8)
Run them in [LM Studio](https://lmstudio.ai/)
## Prompt format
```
<s><|im_start|> system
{system_prompt}<|im_end|>
<|im_start|> user
{prompt}<|im_end|>
<|im_start|> assistant
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [Bielik-11B-v2.2-Instruct-f16.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-f16.gguf) | f16 | 22.34GB | false | Full F16 weights. |
| [Bielik-11B-v2.2-Instruct-Q8_0.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q8_0.gguf) | Q8_0 | 11.87GB | false | Extremely high quality, generally unneeded but max available quant. |
| [Bielik-11B-v2.2-Instruct-Q6_K_L.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q6_K_L.gguf) | Q6_K_L | 9.23GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q6_K.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q6_K.gguf) | Q6_K | 9.16GB | false | Very high quality, near perfect, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q5_K_L.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q5_K_L.gguf) | Q5_K_L | 7.99GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q5_K_M.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q5_K_M.gguf) | Q5_K_M | 7.91GB | false | High quality, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q5_K_S.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q5_K_S.gguf) | Q5_K_S | 7.70GB | false | High quality, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q4_K_L.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_K_L.gguf) | Q4_K_L | 6.82GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q4_K_M.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_K_M.gguf) | Q4_K_M | 6.72GB | false | Good quality, default size for must use cases, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q4_K_S.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_K_S.gguf) | Q4_K_S | 6.36GB | false | Slightly lower quality with more space savings, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q4_0.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_0.gguf) | Q4_0 | 6.34GB | false | Legacy format, generally not worth using over similarly sized formats |
| [Bielik-11B-v2.2-Instruct-Q4_0_8_8.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_0_8_8.gguf) | Q4_0_8_8 | 6.32GB | false | Optimized for ARM and CPU inference, much faster than Q4_0 at similar quality. |
| [Bielik-11B-v2.2-Instruct-Q4_0_4_8.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_0_4_8.gguf) | Q4_0_4_8 | 6.32GB | false | Optimized for ARM and CPU inference, much faster than Q4_0 at similar quality. |
| [Bielik-11B-v2.2-Instruct-Q4_0_4_4.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q4_0_4_4.gguf) | Q4_0_4_4 | 6.32GB | false | Optimized for ARM and CPU inference, much faster than Q4_0 at similar quality. |
| [Bielik-11B-v2.2-Instruct-IQ4_XS.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-IQ4_XS.gguf) | IQ4_XS | 6.01GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [Bielik-11B-v2.2-Instruct-Q3_K_XL.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q3_K_XL.gguf) | Q3_K_XL | 6.00GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [Bielik-11B-v2.2-Instruct-Q3_K_L.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q3_K_L.gguf) | Q3_K_L | 5.88GB | false | Lower quality but usable, good for low RAM availability. |
| [Bielik-11B-v2.2-Instruct-Q3_K_M.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q3_K_M.gguf) | Q3_K_M | 5.40GB | false | Low quality. |
| [Bielik-11B-v2.2-Instruct-IQ3_M.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-IQ3_M.gguf) | IQ3_M | 5.04GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [Bielik-11B-v2.2-Instruct-Q3_K_S.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q3_K_S.gguf) | Q3_K_S | 4.85GB | false | Low quality, not recommended. |
| [Bielik-11B-v2.2-Instruct-IQ3_XS.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-IQ3_XS.gguf) | IQ3_XS | 4.63GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [Bielik-11B-v2.2-Instruct-Q2_K_L.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q2_K_L.gguf) | Q2_K_L | 4.29GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [Bielik-11B-v2.2-Instruct-Q2_K.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-Q2_K.gguf) | Q2_K | 4.16GB | false | Very low quality but surprisingly usable. |
| [Bielik-11B-v2.2-Instruct-IQ2_M.gguf](https://huggingface.co/bartowski/Bielik-11B-v2.2-Instruct-GGUF/blob/main/Bielik-11B-v2.2-Instruct-IQ2_M.gguf) | IQ2_M | 3.82GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
Some say that this improves the quality, others don't notice any difference. If you use these models PLEASE COMMENT with your findings. I would like feedback that these are actually used and useful so I don't keep uploading quants no one is using.
Thanks!
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset
Thank you ZeroWw for the inspiration to experiment with embed/output
## Downloading using huggingface-cli
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/Bielik-11B-v2.2-Instruct-GGUF --include "Bielik-11B-v2.2-Instruct-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/Bielik-11B-v2.2-Instruct-GGUF --include "Bielik-11B-v2.2-Instruct-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (Bielik-11B-v2.2-Instruct-Q8_0) or download them all in place (./)
## Which file should I choose?
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggerganov/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU and Apple Metal, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
The I-quants are *not* compatible with Vulcan, which is also AMD, so if you have an AMD card double check if you're using the rocBLAS build or the Vulcan build. At the time of writing this, LM Studio has a preview with ROCm support, and other inference engines have specific builds for ROCm.
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}