初始化项目,由ModelHub XC社区提供模型

Model: bartowski/burtenshaw_GemmaCoder3-12B-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-29 15:36:13 +08:00
commit 1c9c56679b
29 changed files with 308 additions and 0 deletions

49
.gitattributes vendored Normal file
View File

@@ -0,0 +1,49 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bin.* filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zstandard filter=lfs diff=lfs merge=lfs -text
*.tfevents* filter=lfs diff=lfs merge=lfs -text
*.db* filter=lfs diff=lfs merge=lfs -text
*.ark* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*data* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.meta filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.index filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.gguf* filter=lfs diff=lfs merge=lfs -text
*.ggml filter=lfs diff=lfs merge=lfs -text
*.llamafile* filter=lfs diff=lfs merge=lfs -text
*.pt2 filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
burtenshaw_GemmaCoder3-12B.imatrix filter=lfs diff=lfs merge=lfs -text

180
README.md Normal file
View File

@@ -0,0 +1,180 @@
---
quantized_by: bartowski
pipeline_tag: text-generation
base_model_relation: quantized
datasets: open-r1/codeforces-cots
licence: license
base_model: burtenshaw/GemmaCoder3-12B
tags:
- generated_from_trainer
- trl
- sft
model_name: gemma-3-12b-it-codeforces-SFT
---
## Llamacpp imatrix Quantizations of GemmaCoder3-12B by burtenshaw
Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b5010">b5010</a> for quantization.
Original model: https://huggingface.co/burtenshaw/GemmaCoder3-12B
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8)
Run them in [LM Studio](https://lmstudio.ai/)
Run them directly with [llama.cpp](https://github.com/ggerganov/llama.cpp), or any other llama.cpp based project
## Prompt format
```
<bos><start_of_turn>user
{system_prompt}
{prompt}<end_of_turn>
<start_of_turn>model
<end_of_turn>
<start_of_turn>model
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [GemmaCoder3-12B-bf16.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-bf16.gguf) | bf16 | 23.54GB | false | Full BF16 weights. |
| [GemmaCoder3-12B-Q8_0.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q8_0.gguf) | Q8_0 | 12.51GB | false | Extremely high quality, generally unneeded but max available quant. |
| [GemmaCoder3-12B-Q6_K_L.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q6_K_L.gguf) | Q6_K_L | 9.90GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [GemmaCoder3-12B-Q6_K.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q6_K.gguf) | Q6_K | 9.66GB | false | Very high quality, near perfect, *recommended*. |
| [GemmaCoder3-12B-Q5_K_L.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q5_K_L.gguf) | Q5_K_L | 8.69GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [GemmaCoder3-12B-Q5_K_M.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q5_K_M.gguf) | Q5_K_M | 8.45GB | false | High quality, *recommended*. |
| [GemmaCoder3-12B-Q5_K_S.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q5_K_S.gguf) | Q5_K_S | 8.23GB | false | High quality, *recommended*. |
| [GemmaCoder3-12B-Q4_1.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q4_1.gguf) | Q4_1 | 7.56GB | false | Legacy format, similar performance to Q4_K_S but with improved tokens/watt on Apple silicon. |
| [GemmaCoder3-12B-Q4_K_L.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q4_K_L.gguf) | Q4_K_L | 7.54GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [GemmaCoder3-12B-Q4_K_M.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q4_K_M.gguf) | Q4_K_M | 7.30GB | false | Good quality, default size for most use cases, *recommended*. |
| [GemmaCoder3-12B-Q4_K_S.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q4_K_S.gguf) | Q4_K_S | 6.94GB | false | Slightly lower quality with more space savings, *recommended*. |
| [GemmaCoder3-12B-Q4_0.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q4_0.gguf) | Q4_0 | 6.91GB | false | Legacy format, offers online repacking for ARM and AVX CPU inference. |
| [GemmaCoder3-12B-IQ4_NL.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ4_NL.gguf) | IQ4_NL | 6.89GB | false | Similar to IQ4_XS, but slightly larger. Offers online repacking for ARM CPU inference. |
| [GemmaCoder3-12B-Q3_K_XL.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q3_K_XL.gguf) | Q3_K_XL | 6.72GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [GemmaCoder3-12B-IQ4_XS.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ4_XS.gguf) | IQ4_XS | 6.55GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [GemmaCoder3-12B-Q3_K_L.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q3_K_L.gguf) | Q3_K_L | 6.48GB | false | Lower quality but usable, good for low RAM availability. |
| [GemmaCoder3-12B-Q3_K_M.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q3_K_M.gguf) | Q3_K_M | 6.01GB | false | Low quality. |
| [GemmaCoder3-12B-IQ3_M.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ3_M.gguf) | IQ3_M | 5.66GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [GemmaCoder3-12B-Q3_K_S.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q3_K_S.gguf) | Q3_K_S | 5.46GB | false | Low quality, not recommended. |
| [GemmaCoder3-12B-IQ3_XS.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ3_XS.gguf) | IQ3_XS | 5.21GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [GemmaCoder3-12B-Q2_K_L.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q2_K_L.gguf) | Q2_K_L | 5.01GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [GemmaCoder3-12B-IQ3_XXS.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ3_XXS.gguf) | IQ3_XXS | 4.78GB | false | Lower quality, new method with decent performance, comparable to Q3 quants. |
| [GemmaCoder3-12B-Q2_K.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-Q2_K.gguf) | Q2_K | 4.77GB | false | Very low quality but surprisingly usable. |
| [GemmaCoder3-12B-IQ2_M.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ2_M.gguf) | IQ2_M | 4.31GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
| [GemmaCoder3-12B-IQ2_S.gguf](https://huggingface.co/bartowski/burtenshaw_GemmaCoder3-12B-GGUF/blob/main/burtenshaw_GemmaCoder3-12B-IQ2_S.gguf) | IQ2_S | 4.02GB | false | Low quality, uses SOTA techniques to be usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
## Downloading using huggingface-cli
<details>
<summary>Click to view download instructions</summary>
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/burtenshaw_GemmaCoder3-12B-GGUF --include "burtenshaw_GemmaCoder3-12B-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/burtenshaw_GemmaCoder3-12B-GGUF --include "burtenshaw_GemmaCoder3-12B-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (burtenshaw_GemmaCoder3-12B-Q8_0) or download them all in place (./)
</details>
## ARM/AVX information
Previously, you would download Q4_0_4_4/4_8/8_8, and these would have their weights interleaved in memory in order to improve performance on ARM and AVX machines by loading up more data in one pass.
Now, however, there is something called "online repacking" for weights. details in [this PR](https://github.com/ggerganov/llama.cpp/pull/9921). If you use Q4_0 and your hardware would benefit from repacking weights, it will do it automatically on the fly.
As of llama.cpp build [b4282](https://github.com/ggerganov/llama.cpp/releases/tag/b4282) you will not be able to run the Q4_0_X_X files and will instead need to use Q4_0.
Additionally, if you want to get slightly better quality for , you can use IQ4_NL thanks to [this PR](https://github.com/ggerganov/llama.cpp/pull/10541) which will also repack the weights for ARM, though only the 4_4 for now. The loading time may be slower but it will result in an overall speed incrase.
<details>
<summary>Click to view Q4_0_X_X information (deprecated</summary>
I'm keeping this section to show the potential theoretical uplift in performance from using the Q4_0 with online repacking.
<details>
<summary>Click to view benchmarks on an AVX2 system (EPYC7702)</summary>
| model | size | params | backend | threads | test | t/s | % (vs Q4_0) |
| ------------------------------ | ---------: | ---------: | ---------- | ------: | ------------: | -------------------: |-------------: |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp512 | 204.03 ± 1.03 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp1024 | 282.92 ± 0.19 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp2048 | 259.49 ± 0.44 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg128 | 39.12 ± 0.27 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg256 | 39.31 ± 0.69 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg512 | 40.52 ± 0.03 | 100% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp512 | 301.02 ± 1.74 | 147% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp1024 | 287.23 ± 0.20 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp2048 | 262.77 ± 1.81 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg128 | 18.80 ± 0.99 | 48% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg256 | 24.46 ± 3.04 | 83% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg512 | 36.32 ± 3.59 | 90% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp512 | 271.71 ± 3.53 | 133% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp1024 | 279.86 ± 45.63 | 100% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp2048 | 320.77 ± 5.00 | 124% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg128 | 43.51 ± 0.05 | 111% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg256 | 43.35 ± 0.09 | 110% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg512 | 42.60 ± 0.31 | 105% |
Q4_0_8_8 offers a nice bump to prompt processing and a small bump to text generation
</details>
</details>
## Which file should I choose?
<details>
<summary>Click here for details</summary>
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggerganov/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
</details>
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset.
Thank you ZeroWw for the inspiration to experiment with embed/output.
Thank you to LM Studio for sponsoring my work.
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0885a2ebc512bf8ca8a91c6c0e15070730836f8322faf9585c94d54c658165d9
size 4310461376

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3d15b5126639697095e140d23f64a731d3d12763bb9f14d625065501666221ee
size 4020710336

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d2447950548a9828838126c01fe7f0bb979b0510f550d822f78cb5f34dc414e2
size 5655722816

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:130225feb6840d1e661ce9576603ef904b878132cecdd994f89603d67a7cb8c1
size 5206166336

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b95578ab9f008f81fff3ed7db8f19238988f93c0561f3b3afd946926d674eacf
size 4784901056

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:cb6eb051b2a1419d954b0740ed2bc40f17615acbe7329f1c3aa1e5f0c6226e83
size 6887164736

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d4fb23cc46e5d0d5f3d8bd578b6e11d2b7621d6fe9606afcb64af4c0c942fb95
size 6550965056

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8b846622da11857a5cc275a377c666254e8944add208edb026a8919ea7532aa2
size 4768222016

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c96932c18b80bfb4c325323dc556413c5ab5befa0003ca7887c2ecb24823dd31
size 5012075456

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e57564bc3be07866237eb69cb4055f84c2a8ecad3f6790df3a18e19d66ba205a
size 6480186176

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:075da4176fb1e156cf947ffc028a55dc7288e0faaa06cb792f9fe3c18f3d0038
size 6008818496

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3c489423701b62312db80749607220cc5785320bb2843d25f9defb0631264bbe
size 5458316096

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:cb7ac803c340b2203f4765d7faa8ecee7aec6c921597b7cb330e13df7bad8ad9
size 6724039616

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:32cbad286f802fd4ccc6c6949c3949b481b03d4f83435fb17f36e8d46b82b0c0
size 6909283136

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f89cf70d4b8a58546598e975344e7a2decb1d3abc4880f4c4fa1662c436acc3e
size 7559564096

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5d3403d01526307138dc2c4624b0039a1048c2b8753e256d847790ea89b1391a
size 7544632256

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:47f0a2848eeed783cb03336afd8cc69f6ee0e088e3cec11ab6d9fe16457dc3d4
size 7300778816

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:390e5689138a54660e30d48ed445e6dfcd8fc33b230d84b15764d1ce38360ad3
size 6935333696

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3b62ece292c7990b5d43159961f02db01629c5d92005739e4b5047fa509cf0a6
size 8688890816

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:38e4ad9a3a6ed179270c0a01a07117591874efd27598735a8b0dc07ca4aefe16
size 8445037376

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:30afc5e95735274b8e5b13bb0c444ad4339d7c6635417f0ba24f0ccbfa434a0b
size 8231963456

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:70aa1dc4a916a4240a89866d2cdb0560f0ddf22f61cab7d1f7b2eafdc7ee3bbd
size 9660812096

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:982d9d6e63728b078eb8e9151bc6f3fd6c18d11455f35630ffb12c742f57bc89
size 9904665536

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:cae558035b50ab41b691b7454828830111b3d9c15b04b4ab8d00f0ebfc952d55
size 12510213056

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c3bb9fd86b454e53e957e12b42e35370dcade6de8f4c4139c6660dc023ded139
size 23540151968

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:907d3a2ed1caf19e43ee5b69deed83aa64f96d6598edf1aae257fd164a7481a5
size 7433114

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}