初始化项目,由ModelHub XC社区提供模型

Model: bartowski/TQ2.5-14B-Sugarquill-v1-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-09-01 14:06:12 +08:00
commit 553462f86b
30 changed files with 276 additions and 0 deletions

62
.gitattributes vendored Normal file
View File

@@ -0,0 +1,62 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q6_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q5_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_0_8_8.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_0_4_8.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_0_4_4.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q3_K_XL.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q2_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1-f16.gguf filter=lfs diff=lfs merge=lfs -text
TQ2.5-14B-Sugarquill-v1.imatrix filter=lfs diff=lfs merge=lfs -text

132
README.md Normal file
View File

@@ -0,0 +1,132 @@
---
quantized_by: bartowski
pipeline_tag: text-generation
language:
- en
license: apache-2.0
base_model: allura-org/TQ2.5-14B-Sugarquill-v1
datasets:
- Mielikki/Erebus-87k
- allura-org/r_shortstories_24k
---
## Llamacpp imatrix Quantizations of TQ2.5-14B-Sugarquill-v1
Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b4014">b4014</a> for quantization.
Original model: https://huggingface.co/allura-org/TQ2.5-14B-Sugarquill-v1
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8)
Run them in [LM Studio](https://lmstudio.ai/)
## Prompt format
```
<|im_start|>system
{system_prompt}<|im_end|>
<|im_start|>user
{prompt}<|im_end|>
<|im_start|>assistant
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [TQ2.5-14B-Sugarquill-v1-f16.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-f16.gguf) | f16 | 29.55GB | false | Full F16 weights. |
| [TQ2.5-14B-Sugarquill-v1-Q8_0.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q8_0.gguf) | Q8_0 | 15.70GB | false | Extremely high quality, generally unneeded but max available quant. |
| [TQ2.5-14B-Sugarquill-v1-Q6_K_L.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q6_K_L.gguf) | Q6_K_L | 12.50GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q6_K.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q6_K.gguf) | Q6_K | 12.12GB | false | Very high quality, near perfect, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q5_K_L.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q5_K_L.gguf) | Q5_K_L | 10.99GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q5_K_M.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q5_K_M.gguf) | Q5_K_M | 10.51GB | false | High quality, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q5_K_S.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q5_K_S.gguf) | Q5_K_S | 10.27GB | false | High quality, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q4_K_L.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_K_L.gguf) | Q4_K_L | 9.57GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q4_K_M.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_K_M.gguf) | Q4_K_M | 8.99GB | false | Good quality, default size for must use cases, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q3_K_XL.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q3_K_XL.gguf) | Q3_K_XL | 8.61GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [TQ2.5-14B-Sugarquill-v1-Q4_K_S.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_K_S.gguf) | Q4_K_S | 8.57GB | false | Slightly lower quality with more space savings, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q4_0.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_0.gguf) | Q4_0 | 8.54GB | false | Legacy format, generally not worth using over similarly sized formats |
| [TQ2.5-14B-Sugarquill-v1-Q4_0_8_8.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_0_8_8.gguf) | Q4_0_8_8 | 8.52GB | false | Optimized for ARM inference. Requires 'sve' support (see link below). *Don't use on Mac or Windows*. |
| [TQ2.5-14B-Sugarquill-v1-Q4_0_4_8.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_0_4_8.gguf) | Q4_0_4_8 | 8.52GB | false | Optimized for ARM inference. Requires 'i8mm' support (see link below). *Don't use on Mac or Windows*. |
| [TQ2.5-14B-Sugarquill-v1-Q4_0_4_4.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q4_0_4_4.gguf) | Q4_0_4_4 | 8.52GB | false | Optimized for ARM inference. Should work well on all ARM chips, pick this if you're unsure. *Don't use on Mac or Windows*. |
| [TQ2.5-14B-Sugarquill-v1-IQ4_XS.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ4_XS.gguf) | IQ4_XS | 8.12GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [TQ2.5-14B-Sugarquill-v1-Q3_K_L.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q3_K_L.gguf) | Q3_K_L | 7.92GB | false | Lower quality but usable, good for low RAM availability. |
| [TQ2.5-14B-Sugarquill-v1-Q3_K_M.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q3_K_M.gguf) | Q3_K_M | 7.34GB | false | Low quality. |
| [TQ2.5-14B-Sugarquill-v1-IQ3_M.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ3_M.gguf) | IQ3_M | 6.92GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [TQ2.5-14B-Sugarquill-v1-Q3_K_S.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q3_K_S.gguf) | Q3_K_S | 6.66GB | false | Low quality, not recommended. |
| [TQ2.5-14B-Sugarquill-v1-Q2_K_L.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q2_K_L.gguf) | Q2_K_L | 6.53GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [TQ2.5-14B-Sugarquill-v1-IQ3_XS.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ3_XS.gguf) | IQ3_XS | 6.38GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [TQ2.5-14B-Sugarquill-v1-Q2_K.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-Q2_K.gguf) | Q2_K | 5.77GB | false | Very low quality but surprisingly usable. |
| [TQ2.5-14B-Sugarquill-v1-IQ2_M.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ2_M.gguf) | IQ2_M | 5.36GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
| [TQ2.5-14B-Sugarquill-v1-IQ2_S.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ2_S.gguf) | IQ2_S | 5.00GB | false | Low quality, uses SOTA techniques to be usable. |
| [TQ2.5-14B-Sugarquill-v1-IQ2_XS.gguf](https://huggingface.co/bartowski/TQ2.5-14B-Sugarquill-v1-GGUF/blob/main/TQ2.5-14B-Sugarquill-v1-IQ2_XS.gguf) | IQ2_XS | 4.70GB | false | Low quality, uses SOTA techniques to be usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
Some say that this improves the quality, others don't notice any difference. If you use these models PLEASE COMMENT with your findings. I would like feedback that these are actually used and useful so I don't keep uploading quants no one is using.
Thanks!
## Downloading using huggingface-cli
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/TQ2.5-14B-Sugarquill-v1-GGUF --include "TQ2.5-14B-Sugarquill-v1-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/TQ2.5-14B-Sugarquill-v1-GGUF --include "TQ2.5-14B-Sugarquill-v1-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (TQ2.5-14B-Sugarquill-v1-Q8_0) or download them all in place (./)
## Q4_0_X_X
These are *NOT* for Metal (Apple) offloading, only ARM chips.
If you're using an ARM chip, the Q4_0_X_X quants will have a substantial speedup. Check out Q4_0_4_4 speed comparisons [on the original pull request](https://github.com/ggerganov/llama.cpp/pull/5780#pullrequestreview-21657544660)
To check which one would work best for your ARM chip, you can check [AArch64 SoC features](https://gpages.juszkiewicz.com.pl/arm-socs-table/arm-socs.html) (thanks EloyOn!).
## Which file should I choose?
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggerganov/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU and Apple Metal, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
The I-quants are *not* compatible with Vulcan, which is also AMD, so if you have an AMD card double check if you're using the rocBLAS build or the Vulcan build. At the time of writing this, LM Studio has a preview with ROCm support, and other inference engines have specific builds for ROCm.
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset
Thank you ZeroWw for the inspiration to experiment with embed/output
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f5c34a419ed1123e765ab006a5b68fff1d7f621daeda6316e85417af5b0b6539
size 5356146912

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a6857100633b99756bafe1838217480ee66cb371b54fa9df083b4cbc54167c04
size 5003727072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:fa0822f7ef89a6a9eefc3d4dc3f5678504e8ca3c69df1d8bc730c31d4d105ce2
size 4704575712

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c69e030c770cdf18baee5c2839e562c9c0653bdfa09ae267aa439fa876393aa6
size 6916538592

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f13ad831f2a5a8762327b706ce327472c88098623ba228856b0c25abde4b4989
size 6383362272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:22155b46ec5257d602a2b953706e291bc81392a4152944d111b14166227359f9
size 8119840992

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:efb83d37bf3cb119a7420c5c76674099464dbf3c1b854317c684fb6fd5b0aeb7
size 5770498272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e8e572b18dc3333b4b75d683a4386133a043a2db5bfac9b9e62238fea3e1392b
size 6530818272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d6b6320b5d04881d01d16119e6b2abcde08d6c3444403dd6166ae66fbd47e9bd
size 7924768992

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:68d7006ead107ef668d95a3e73250e4b6fefd7b59f318a57d1c0a858c015a201
size 7339204832

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:661a455848ce083b739d5a1f32c1df28888dfc53e89b804abc1a526a5a7ad8d4
size 6659596512

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:04fb48dc63b7291d99922610595463a953952aea1e5be40f96e89131deb79606
size 8606015712

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3ccd3e50a023fcd538a5d1b425fe2ec176bb5f0a3308a8d4930f1dde3e9d6346
size 8544268512

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9ec14bf77218369ea5627ff3dacf0e62846d3718dc77f0015a889470b5fb038f
size 8517726432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b1d3492f7a8ade282bb5715b4a1d9b8baf6ed2f162cdf0c7dd7ce5f4657eaa42
size 8517726432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:70ad5848751eef5e8bfa81323f9898962e501841347f953c86debd3e2c91d952
size 8517726432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6f571c00181f057dece42c67492106e6a2980473850048f906e5574e919ee12b
size 9565954272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a654fe3f41e963d8ea6753fb9a06b9dd76893714ebf02605ef67827944a4025e
size 8988111072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e99eb6a3998494a96fec28c58077ae6c0b3714cc9da04ad68e274ad7e2985789
size 8573432032

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e737ac65884300d79014cf85c56f4f1d0511dab5cce90325d555d35cfaa311ff
size 10989396192

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d0c63e34a765d7560d686088d42b63b21f11780dd5422e4b0b9d1b2985e9eba5
size 10508873952

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8275e8670abaf81180165707d340be37bc9b4f1cb8d97d6cacb8e87b435e9934
size 10266554592

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:754a32be573adccb9b49ce7bcfe57956e9bbabd483f224331fe10a1c8e73776f
size 12124684512

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ec22903360a2168adb493e5aa76ebba84ab58801aec10b72d3c276b6b0b42275
size 12501803232

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7105a9c172c7183a9e093f7a74e871306809f3576c2abcd026cc26c723329724
size 15701598432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bc0b37c69b60f7242b301c16363d50251deec89f34c74bb06fbce0cdf3df6c15
size 29547716544

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:eeb72ea6904f01a02d3222336c5c8f2c959f69e362ce54b2bd0b07ad1ac22235
size 8563610

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}