初始化项目,由ModelHub XC社区提供模型

Model: bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-08-27 09:56:13 +08:00
commit 096b2974ba
29 changed files with 300 additions and 0 deletions

49
.gitattributes vendored Normal file
View File

@@ -0,0 +1,49 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bin.* filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zstandard filter=lfs diff=lfs merge=lfs -text
*.tfevents* filter=lfs diff=lfs merge=lfs -text
*.db* filter=lfs diff=lfs merge=lfs -text
*.ark* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*data* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.meta filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.index filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.gguf* filter=lfs diff=lfs merge=lfs -text
*.ggml filter=lfs diff=lfs merge=lfs -text
*.llamafile* filter=lfs diff=lfs merge=lfs -text
*.pt2 filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
TheDrummer_Snowpiercer-15B-v1.imatrix filter=lfs diff=lfs merge=lfs -text

172
README.md Normal file
View File

@@ -0,0 +1,172 @@
---
quantized_by: bartowski
pipeline_tag: text-generation
base_model: TheDrummer/Snowpiercer-15B-v1
license: mit
base_model_relation: quantized
---
## Llamacpp imatrix Quantizations of Snowpiercer-15B-v1 by TheDrummer
Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b5338">b5338</a> for quantization.
Original model: https://huggingface.co/TheDrummer/Snowpiercer-15B-v1
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8)
Run them in [LM Studio](https://lmstudio.ai/)
Run them directly with [llama.cpp](https://github.com/ggerganov/llama.cpp), or any other llama.cpp based project
## Prompt format
```
<|im_start|>system
{system_prompt}<|im_end|>
<|im_start|>user
{prompt}<|im_end|>
<|im_start|>assistant
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [Snowpiercer-15B-v1-bf16.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-bf16.gguf) | bf16 | 29.96GB | false | Full BF16 weights. |
| [Snowpiercer-15B-v1-Q8_0.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q8_0.gguf) | Q8_0 | 15.92GB | false | Extremely high quality, generally unneeded but max available quant. |
| [Snowpiercer-15B-v1-Q6_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q6_K_L.gguf) | Q6_K_L | 12.62GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [Snowpiercer-15B-v1-Q6_K.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q6_K.gguf) | Q6_K | 12.29GB | false | Very high quality, near perfect, *recommended*. |
| [Snowpiercer-15B-v1-Q5_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q5_K_L.gguf) | Q5_K_L | 11.07GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [Snowpiercer-15B-v1-Q5_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q5_K_M.gguf) | Q5_K_M | 10.65GB | false | High quality, *recommended*. |
| [Snowpiercer-15B-v1-Q5_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q5_K_S.gguf) | Q5_K_S | 10.39GB | false | High quality, *recommended*. |
| [Snowpiercer-15B-v1-Q4_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q4_K_L.gguf) | Q4_K_L | 9.61GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [Snowpiercer-15B-v1-Q4_1.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q4_1.gguf) | Q4_1 | 9.50GB | false | Legacy format, similar performance to Q4_K_S but with improved tokens/watt on Apple silicon. |
| [Snowpiercer-15B-v1-Q4_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q4_K_M.gguf) | Q4_K_M | 9.11GB | false | Good quality, default size for most use cases, *recommended*. |
| [Snowpiercer-15B-v1-Q4_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q4_K_S.gguf) | Q4_K_S | 8.66GB | false | Slightly lower quality with more space savings, *recommended*. |
| [Snowpiercer-15B-v1-IQ4_NL.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ4_NL.gguf) | IQ4_NL | 8.64GB | false | Similar to IQ4_XS, but slightly larger. Offers online repacking for ARM CPU inference. |
| [Snowpiercer-15B-v1-Q4_0.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q4_0.gguf) | Q4_0 | 8.63GB | false | Legacy format, offers online repacking for ARM and AVX CPU inference. |
| [Snowpiercer-15B-v1-Q3_K_XL.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q3_K_XL.gguf) | Q3_K_XL | 8.58GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [Snowpiercer-15B-v1-IQ4_XS.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ4_XS.gguf) | IQ4_XS | 8.20GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [Snowpiercer-15B-v1-Q3_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q3_K_L.gguf) | Q3_K_L | 7.99GB | false | Lower quality but usable, good for low RAM availability. |
| [Snowpiercer-15B-v1-Q3_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q3_K_M.gguf) | Q3_K_M | 7.40GB | false | Low quality. |
| [Snowpiercer-15B-v1-IQ3_M.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ3_M.gguf) | IQ3_M | 6.94GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [Snowpiercer-15B-v1-Q3_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q3_K_S.gguf) | Q3_K_S | 6.71GB | false | Low quality, not recommended. |
| [Snowpiercer-15B-v1-Q2_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q2_K_L.gguf) | Q2_K_L | 6.45GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [Snowpiercer-15B-v1-IQ3_XS.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ3_XS.gguf) | IQ3_XS | 6.42GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [Snowpiercer-15B-v1-IQ3_XXS.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ3_XXS.gguf) | IQ3_XXS | 5.99GB | false | Lower quality, new method with decent performance, comparable to Q3 quants. |
| [Snowpiercer-15B-v1-Q2_K.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-Q2_K.gguf) | Q2_K | 5.79GB | false | Very low quality but surprisingly usable. |
| [Snowpiercer-15B-v1-IQ2_M.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ2_M.gguf) | IQ2_M | 5.35GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
| [Snowpiercer-15B-v1-IQ2_S.gguf](https://huggingface.co/bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF/blob/main/TheDrummer_Snowpiercer-15B-v1-IQ2_S.gguf) | IQ2_S | 4.98GB | false | Low quality, uses SOTA techniques to be usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
## Downloading using huggingface-cli
<details>
<summary>Click to view download instructions</summary>
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF --include "TheDrummer_Snowpiercer-15B-v1-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/TheDrummer_Snowpiercer-15B-v1-GGUF --include "TheDrummer_Snowpiercer-15B-v1-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (TheDrummer_Snowpiercer-15B-v1-Q8_0) or download them all in place (./)
</details>
## ARM/AVX information
Previously, you would download Q4_0_4_4/4_8/8_8, and these would have their weights interleaved in memory in order to improve performance on ARM and AVX machines by loading up more data in one pass.
Now, however, there is something called "online repacking" for weights. details in [this PR](https://github.com/ggerganov/llama.cpp/pull/9921). If you use Q4_0 and your hardware would benefit from repacking weights, it will do it automatically on the fly.
As of llama.cpp build [b4282](https://github.com/ggerganov/llama.cpp/releases/tag/b4282) you will not be able to run the Q4_0_X_X files and will instead need to use Q4_0.
Additionally, if you want to get slightly better quality for , you can use IQ4_NL thanks to [this PR](https://github.com/ggerganov/llama.cpp/pull/10541) which will also repack the weights for ARM, though only the 4_4 for now. The loading time may be slower but it will result in an overall speed incrase.
<details>
<summary>Click to view Q4_0_X_X information (deprecated</summary>
I'm keeping this section to show the potential theoretical uplift in performance from using the Q4_0 with online repacking.
<details>
<summary>Click to view benchmarks on an AVX2 system (EPYC7702)</summary>
| model | size | params | backend | threads | test | t/s | % (vs Q4_0) |
| ------------------------------ | ---------: | ---------: | ---------- | ------: | ------------: | -------------------: |-------------: |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp512 | 204.03 ± 1.03 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp1024 | 282.92 ± 0.19 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp2048 | 259.49 ± 0.44 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg128 | 39.12 ± 0.27 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg256 | 39.31 ± 0.69 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg512 | 40.52 ± 0.03 | 100% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp512 | 301.02 ± 1.74 | 147% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp1024 | 287.23 ± 0.20 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp2048 | 262.77 ± 1.81 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg128 | 18.80 ± 0.99 | 48% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg256 | 24.46 ± 3.04 | 83% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg512 | 36.32 ± 3.59 | 90% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp512 | 271.71 ± 3.53 | 133% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp1024 | 279.86 ± 45.63 | 100% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp2048 | 320.77 ± 5.00 | 124% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg128 | 43.51 ± 0.05 | 111% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg256 | 43.35 ± 0.09 | 110% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg512 | 42.60 ± 0.31 | 105% |
Q4_0_8_8 offers a nice bump to prompt processing and a small bump to text generation
</details>
</details>
## Which file should I choose?
<details>
<summary>Click here for details</summary>
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggerganov/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
</details>
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset.
Thank you ZeroWw for the inspiration to experiment with embed/output.
Thank you to LM Studio for sponsoring my work.
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e8de0fb43b1b231b0833cd171539cc1332ee68ec23febb618a4402614bc021b3
size 5352369216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:306a263bf5d64ff8159716c1e90379526e7ab14ce0ecc344ac1db8697a870e1b
size 4981107776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1bc1abe517dbe90706ec3b35d8510f658194c20a61a8d116b51a9f5fadace08a
size 6938668096

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b7e9edd87f3b98b174f50bd43fc75f3a00430574751e6705c65a47932d361eee
size 6424865856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:946acbb21980dc0aa67caf0b542c8ed2143cd2c7a7be1cd92187949f39282a5b
size 5992328256

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5f2f752821310f324307b92de1197d05b735019efd4cb61b9ce4ea079aa5d36f
size 8638426176

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:83a6fa2fcccdb21e07c8a7e635c53eb8d37117a69cfeba87669b39f998105e89
size 8199662656

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:77e4afcf4a38b88b6803571b9e89cc646217aa358ea9b05b1764f733f1b99233
size 5794163776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bb79784498b964d3c0787791985f305d24350626edb0cf6b6d30004ceff81a43
size 6449523776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:288efb6e8c7bdd8e9d4cda9c13b3c4c8084f8fcdff51fe76f88a91a8add5e194
size 7990193216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:32406f946e6a510f11650499bdac70f4b1b0b7989f618e3bbc99eee156258c51
size 7396437056

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:726a657ae9a30aff95980588b7709faaf6b2908cdccfa2dc9b1aa3d4a43fa44d
size 6706097216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c4bcb530973d48a3e1027047034153905e978558126898213f71105e99888d31
size 8577395776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6dd8703c45d08bc3e0ae955b98692d1ca619df26b0b692c18735bec032cda469
size 8633183296

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d5a7ef5e88c1d46b06339bcaad7fa35fa8264c7f1fdece0822292b458c6021ed
size 9499569216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:53bf32e20d1d7fa5a0a6cea34d30649032e49ad46c96912440573a876145fd56
size 9610611776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:89a8996236399e2bd70f106c6aa31c2880d8de3638105c9e1fc192783b422352
size 9112538176

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b2c159918a862ecc17dbf093d721c7b1f82452c94294ca3b22f8b0dc5a6a9e6d
size 8663329856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6defc3da88f4333da717f04d0821359a3f8f616659cc907e7b857329f8e84997
size 11068787776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1ae81d861af480c1463c6e17fb6914af0daf6566b600187b9bcaf2f0f6960e07
size 10654600256

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bf1f3016cd1c2883f94f214187757dece4464173467feb90c6ec903a08dbed3a
size 10393480256

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2f3deb421348b98834cda1729cc3b7d4533b07656d6d315ff2bfdf7938b0dbe4
size 12293041216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5ba30ae688f71e9be19377cb71298d1c41e89a023fd0155d3125bc7bfb65f2dd
size 12618099776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a1380873094d319fd73e01a6ab9cb1297d8ff50c379ab5eb99fc5abf1f7385b4
size 15919475776

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:178fb4f593250f43afa1cf3a581ecc836ef249546ac243928ab48ae3714e9883
size 29957286688

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d481a75018bf5b2048210aee7fead30a7eecfb19e1eb2fad79aadaf8bf261c90
size 8818028

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}