初始化项目,由ModelHub XC社区提供模型

Model: bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-15 05:25:06 +08:00
commit bac26d0780
28 changed files with 268 additions and 0 deletions

60
.gitattributes vendored Normal file
View File

@@ -0,0 +1,60 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K_L.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_L.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_L.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_8_8.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_8.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_4.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_XL.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K_L.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-f16.gguf filter=lfs diff=lfs merge=lfs -text
WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B.imatrix filter=lfs diff=lfs merge=lfs -text

132
README.md Normal file
View File

@@ -0,0 +1,132 @@
---
base_model: WhiteRabbitNeo/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B
language:
- en
library_name: transformers
license: apache-2.0
pipeline_tag: text-generation
tags:
- code
- qwen-coder
- finetune
quantized_by: bartowski
---
## Llamacpp imatrix Quantizations of WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B
Using <a href="https://github.com/ggerganov/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggerganov/llama.cpp/releases/tag/b3878">b3878</a> for quantization.
Original model: https://huggingface.co/WhiteRabbitNeo/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8)
Run them in [LM Studio](https://lmstudio.ai/)
## Prompt format
```
<|im_start|>system
{system_prompt}<|im_end|>
<|im_start|>user
{prompt}<|im_end|>
<|im_start|>assistant
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-f16.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-f16.gguf) | f16 | 15.24GB | false | Full F16 weights. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q8_0.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q8_0.gguf) | Q8_0 | 8.10GB | false | Extremely high quality, generally unneeded but max available quant. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K_L.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K_L.gguf) | Q6_K_L | 6.52GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q6_K.gguf) | Q6_K | 6.25GB | false | Very high quality, near perfect, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_L.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_L.gguf) | Q5_K_L | 5.78GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_M.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_M.gguf) | Q5_K_M | 5.44GB | false | High quality, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_S.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q5_K_S.gguf) | Q5_K_S | 5.32GB | false | High quality, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_L.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_L.gguf) | Q4_K_L | 5.09GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_M.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_M.gguf) | Q4_K_M | 4.68GB | false | Good quality, default size for must use cases, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_XL.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_XL.gguf) | Q3_K_XL | 4.57GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_S.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_S.gguf) | Q4_K_S | 4.46GB | false | Slightly lower quality with more space savings, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0.gguf) | Q4_0 | 4.44GB | false | Legacy format, generally not worth using over similarly sized formats |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_8_8.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_8_8.gguf) | Q4_0_8_8 | 4.43GB | false | Optimized for ARM inference. Requires 'sve' support (see link below). |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_8.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_8.gguf) | Q4_0_4_8 | 4.43GB | false | Optimized for ARM inference. Requires 'i8mm' support (see link below). |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_4.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_0_4_4.gguf) | Q4_0_4_4 | 4.43GB | false | Optimized for ARM inference. Should work well on all ARM chips, pick this if you're unsure. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ4_XS.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ4_XS.gguf) | IQ4_XS | 4.22GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_L.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_L.gguf) | Q3_K_L | 4.09GB | false | Lower quality but usable, good for low RAM availability. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_M.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_M.gguf) | Q3_K_M | 3.81GB | false | Low quality. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_M.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_M.gguf) | IQ3_M | 3.57GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K_L.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K_L.gguf) | Q2_K_L | 3.55GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_S.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q3_K_S.gguf) | Q3_K_S | 3.49GB | false | Low quality, not recommended. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_XS.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ3_XS.gguf) | IQ3_XS | 3.35GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q2_K.gguf) | Q2_K | 3.02GB | false | Very low quality but surprisingly usable. |
| [WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ2_M.gguf](https://huggingface.co/bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF/blob/main/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-IQ2_M.gguf) | IQ2_M | 2.78GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
Some say that this improves the quality, others don't notice any difference. If you use these models PLEASE COMMENT with your findings. I would like feedback that these are actually used and useful so I don't keep uploading quants no one is using.
Thanks!
## Downloading using huggingface-cli
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF --include "WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-GGUF --include "WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (WhiteRabbitNeo-2.5-Qwen-2.5-Coder-7B-Q8_0) or download them all in place (./)
## Q4_0_X_X
These are *NOT* for Metal (Apple) offloading, only ARM chips.
If you're using an ARM chip, the Q4_0_X_X quants will have a substantial speedup. Check out Q4_0_4_4 speed comparisons [on the original pull request](https://github.com/ggerganov/llama.cpp/pull/5780#pullrequestreview-21657544660)
To check which one would work best for your ARM chip, you can check [AArch64 SoC features](https://gpages.juszkiewicz.com.pl/arm-socs-table/arm-socs.html) (thanks EloyOn!).
## Which file should I choose?
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggerganov/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU and Apple Metal, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
The I-quants are *not* compatible with Vulcan, which is also AMD, so if you have an AMD card double check if you're using the rocBLAS build or the Vulcan build. At the time of writing this, LM Studio has a preview with ROCm support, and other inference engines have specific builds for ROCm.
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset
Thank you ZeroWw for the inspiration to experiment with embed/output
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:95a76df0b81ac227a0306c4c2f0060a3d7c338dcd3e2bb601d75bb5ada052eab
size 2780340768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9b33de532f2952ac89ca7f31b42c07db37a6cae76c6c9870f54374813755542e
size 3574010400

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f3399f6af4512d46f9e1649d0f553eb4241f003f13bfb717d8e882af522f65ce
size 3346254368

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bf0bc5df40c0ef8ec4b293df55a2235a08c4ee009638bbd56a86ea611911b73b
size 4218470944

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7d781b8571c9f9692ff96cf49ee21f23a0fc6beb0f7bca54e4c8bc02e50c9ea7
size 3015938592

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:799d16b5bf58d0c8f8882e90bc2f9e9b9c463fff5591fd1c0fa75b8ef7211926
size 3548162592

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:858dbab8f97ae84754cb4ba1caf14e13f7465ff8da0310625c0249f26198d40e
size 4088457760

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:763640bebf125a93204798862fcbffea4a0e9870190126c64b99f881d0cb5e74
size 3808389664

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9f2572e809385219fd7cf5423d7e36457290c286b4e1383b194de58ce139f186
size 3492366880

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8602c5ed384bfb13f944b2c8b5cfd1767dcd344462ea1dc4d00afe432bd93d1c
size 4565330464

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5e6dd9fb1c9e92c4a2c2ea8889f5e5ec811b7213c4ed7afc373e9eca0ef25fdf
size 4444119584

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:530bfe2701c067f45134bb122b7f7c01aaf5a073a72e6e33127c5aa7d4b586df
size 4431389216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c6e9a457390c3c624b88616f07ea0fbde757b00efc87da90c486fe862f91f562
size 4431389216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e4616bad0836ba8dde6fb337ae7e56f56bd2fb0929d6244da34f2d3b9e690620
size 4431389216

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d897cd0e9e417701f27e2e2684ca9993ad5fe13fd7d3555f68489ced03d403ab
size 5087562272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3790b0bf2c505fcbd144b6b69354fe45a83ac09238a87469db0082027c127de4
size 4683072032

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bf1fb7e76ffb9c78cff6a3eac89f6a90d602ba7493a4a48ca02c7b7912ae3353
size 4457767456

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:07fc0ba7d6d5bba012caa47cdfca45a2a71e230aa9ed139b4df654585f9e082b
size 5781195296

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ba7ac15550b3e9672aa139d5669fcf8e624ad588872b9d4a68580f2eb363b7a2
size 5444829728

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6e8d95b819f3d25da66cedae35c05936740199aa0e35e8893c67af4ae851e29d
size 5315174944

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9e00e1d0c958bc3dce2d08f87efa2a0131d5fd11c584492c55f2401ab0111269
size 6254197280

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:cbf0f4e0753682ff82f306ddc1b98ee697bde0152711b31c969db96b928063ba
size 6518180384

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4bc6ed0f29a39d051164d6db7c6d7a937157f00a718d38cf4fb5cf8e80f368f6
size 8098523680

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5520738f18bcb8e0b766952245a0ef4f46f2e4d169eb5f7e245bbdd75177a587
size 15237851360

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c317f5da9c3cd0f40de6a1b085daac57466eba8f1fd44e485f328e4f29b9bcb0
size 4536678

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}