初始化项目,由ModelHub XC社区提供模型

Model: bartowski/TheDrummer_Precog-24B-v1-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-09-11 17:18:12 +08:00
commit 0a8636921b
30 changed files with 327 additions and 0 deletions

75
.gitattributes vendored Normal file
View File

@@ -0,0 +1,75 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bin.* filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zstandard filter=lfs diff=lfs merge=lfs -text
*.tfevents* filter=lfs diff=lfs merge=lfs -text
*.db* filter=lfs diff=lfs merge=lfs -text
*.ark* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*data* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.meta filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.index filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ggml filter=lfs diff=lfs merge=lfs -text
*.llamafile* filter=lfs diff=lfs merge=lfs -text
*.pt2 filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q2_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q4_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q3_K_XL.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-bf16.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-imatrix.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q6_K_L.gguf filter=lfs diff=lfs merge=lfs -text
TheDrummer_Precog-24B-v1-Q5_K_L.gguf filter=lfs diff=lfs merge=lfs -text

170
README.md Normal file
View File

@@ -0,0 +1,170 @@
---
quantized_by: bartowski
pipeline_tag: text-generation
base_model_relation: quantized
base_model: TheDrummer/Precog-24B-v1
---
## Llamacpp imatrix Quantizations of Precog-24B-v1 by TheDrummer
Using <a href="https://github.com/ggml-org/llama.cpp/">llama.cpp</a> release <a href="https://github.com/ggml-org/llama.cpp/releases/tag/b6907">b6907</a> for quantization.
Original model: https://huggingface.co/TheDrummer/Precog-24B-v1
All quants made using imatrix option with dataset from [here](https://gist.github.com/bartowski1182/eb213dccb3571f863da82e99418f81e8) combined with a subset of combined_all_small.parquet from Ed Addario [here](https://huggingface.co/datasets/eaddario/imatrix-calibration/blob/main/combined_all_small.parquet)
Run them in [LM Studio](https://lmstudio.ai/)
Run them directly with [llama.cpp](https://github.com/ggml-org/llama.cpp), or any other llama.cpp based project
## Prompt format
No chat template specified so default is used. This may be incorrect, check original model card for details.
```
<s>[SYSTEM_PROMPT]{system_prompt}[/SYSTEM_PROMPT][INST]{prompt}[/INST]
```
## Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| [Precog-24B-v1-bf16.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-bf16.gguf) | bf16 | 47.15GB | false | Full BF16 weights. |
| [Precog-24B-v1-Q8_0.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q8_0.gguf) | Q8_0 | 25.05GB | false | Extremely high quality, generally unneeded but max available quant. |
| [Precog-24B-v1-Q6_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q6_K_L.gguf) | Q6_K_L | 19.67GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [Precog-24B-v1-Q6_K.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q6_K.gguf) | Q6_K | 19.35GB | false | Very high quality, near perfect, *recommended*. |
| [Precog-24B-v1-Q5_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q5_K_L.gguf) | Q5_K_L | 17.18GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| [Precog-24B-v1-Q5_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q5_K_M.gguf) | Q5_K_M | 16.76GB | false | High quality, *recommended*. |
| [Precog-24B-v1-Q5_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q5_K_S.gguf) | Q5_K_S | 16.30GB | false | High quality, *recommended*. |
| [Precog-24B-v1-Q4_1.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q4_1.gguf) | Q4_1 | 14.87GB | false | Legacy format, similar performance to Q4_K_S but with improved tokens/watt on Apple silicon. |
| [Precog-24B-v1-Q4_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q4_K_L.gguf) | Q4_K_L | 14.83GB | false | Uses Q8_0 for embed and output weights. Good quality, *recommended*. |
| [Precog-24B-v1-Q4_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q4_K_M.gguf) | Q4_K_M | 14.33GB | false | Good quality, default size for most use cases, *recommended*. |
| [Precog-24B-v1-Q4_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q4_K_S.gguf) | Q4_K_S | 13.55GB | false | Slightly lower quality with more space savings, *recommended*. |
| [Precog-24B-v1-Q4_0.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q4_0.gguf) | Q4_0 | 13.49GB | false | Legacy format, offers online repacking for ARM and AVX CPU inference. |
| [Precog-24B-v1-IQ4_NL.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ4_NL.gguf) | IQ4_NL | 13.47GB | false | Similar to IQ4_XS, but slightly larger. Offers online repacking for ARM CPU inference. |
| [Precog-24B-v1-Q3_K_XL.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q3_K_XL.gguf) | Q3_K_XL | 12.99GB | false | Uses Q8_0 for embed and output weights. Lower quality but usable, good for low RAM availability. |
| [Precog-24B-v1-IQ4_XS.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ4_XS.gguf) | IQ4_XS | 12.76GB | false | Decent quality, smaller than Q4_K_S with similar performance, *recommended*. |
| [Precog-24B-v1-Q3_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q3_K_L.gguf) | Q3_K_L | 12.40GB | false | Lower quality but usable, good for low RAM availability. |
| [Precog-24B-v1-Q3_K_M.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q3_K_M.gguf) | Q3_K_M | 11.47GB | false | Low quality. |
| [Precog-24B-v1-IQ3_M.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ3_M.gguf) | IQ3_M | 10.65GB | false | Medium-low quality, new method with decent performance comparable to Q3_K_M. |
| [Precog-24B-v1-Q3_K_S.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q3_K_S.gguf) | Q3_K_S | 10.40GB | false | Low quality, not recommended. |
| [Precog-24B-v1-IQ3_XS.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ3_XS.gguf) | IQ3_XS | 9.91GB | false | Lower quality, new method with decent performance, slightly better than Q3_K_S. |
| [Precog-24B-v1-Q2_K_L.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q2_K_L.gguf) | Q2_K_L | 9.55GB | false | Uses Q8_0 for embed and output weights. Very low quality but surprisingly usable. |
| [Precog-24B-v1-IQ3_XXS.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ3_XXS.gguf) | IQ3_XXS | 9.28GB | false | Lower quality, new method with decent performance, comparable to Q3 quants. |
| [Precog-24B-v1-Q2_K.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-Q2_K.gguf) | Q2_K | 8.89GB | false | Very low quality but surprisingly usable. |
| [Precog-24B-v1-IQ2_M.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ2_M.gguf) | IQ2_M | 8.11GB | false | Relatively low quality, uses SOTA techniques to be surprisingly usable. |
| [Precog-24B-v1-IQ2_S.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ2_S.gguf) | IQ2_S | 7.48GB | false | Low quality, uses SOTA techniques to be usable. |
| [Precog-24B-v1-IQ2_XS.gguf](https://huggingface.co/bartowski/TheDrummer_Precog-24B-v1-GGUF/blob/main/TheDrummer_Precog-24B-v1-IQ2_XS.gguf) | IQ2_XS | 7.21GB | false | Low quality, uses SOTA techniques to be usable. |
## Embed/output weights
Some of these quants (Q3_K_XL, Q4_K_L etc) are the standard quantization method with the embeddings and output weights quantized to Q8_0 instead of what they would normally default to.
## Downloading using huggingface-cli
<details>
<summary>Click to view download instructions</summary>
First, make sure you have hugginface-cli installed:
```
pip install -U "huggingface_hub[cli]"
```
Then, you can target the specific file you want:
```
huggingface-cli download bartowski/TheDrummer_Precog-24B-v1-GGUF --include "TheDrummer_Precog-24B-v1-Q4_K_M.gguf" --local-dir ./
```
If the model is bigger than 50GB, it will have been split into multiple files. In order to download them all to a local folder, run:
```
huggingface-cli download bartowski/TheDrummer_Precog-24B-v1-GGUF --include "TheDrummer_Precog-24B-v1-Q8_0/*" --local-dir ./
```
You can either specify a new local-dir (TheDrummer_Precog-24B-v1-Q8_0) or download them all in place (./)
</details>
## ARM/AVX information
Previously, you would download Q4_0_4_4/4_8/8_8, and these would have their weights interleaved in memory in order to improve performance on ARM and AVX machines by loading up more data in one pass.
Now, however, there is something called "online repacking" for weights. details in [this PR](https://github.com/ggml-org/llama.cpp/pull/9921). If you use Q4_0 and your hardware would benefit from repacking weights, it will do it automatically on the fly.
As of llama.cpp build [b4282](https://github.com/ggml-org/llama.cpp/releases/tag/b4282) you will not be able to run the Q4_0_X_X files and will instead need to use Q4_0.
Additionally, if you want to get slightly better quality for , you can use IQ4_NL thanks to [this PR](https://github.com/ggml-org/llama.cpp/pull/10541) which will also repack the weights for ARM, though only the 4_4 for now. The loading time may be slower but it will result in an overall speed incrase.
<details>
<summary>Click to view Q4_0_X_X information (deprecated</summary>
I'm keeping this section to show the potential theoretical uplift in performance from using the Q4_0 with online repacking.
<details>
<summary>Click to view benchmarks on an AVX2 system (EPYC7702)</summary>
| model | size | params | backend | threads | test | t/s | % (vs Q4_0) |
| ------------------------------ | ---------: | ---------: | ---------- | ------: | ------------: | -------------------: |-------------: |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp512 | 204.03 ± 1.03 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp1024 | 282.92 ± 0.19 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | pp2048 | 259.49 ± 0.44 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg128 | 39.12 ± 0.27 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg256 | 39.31 ± 0.69 | 100% |
| qwen2 3B Q4_0 | 1.70 GiB | 3.09 B | CPU | 64 | tg512 | 40.52 ± 0.03 | 100% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp512 | 301.02 ± 1.74 | 147% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp1024 | 287.23 ± 0.20 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | pp2048 | 262.77 ± 1.81 | 101% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg128 | 18.80 ± 0.99 | 48% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg256 | 24.46 ± 3.04 | 83% |
| qwen2 3B Q4_K_M | 1.79 GiB | 3.09 B | CPU | 64 | tg512 | 36.32 ± 3.59 | 90% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp512 | 271.71 ± 3.53 | 133% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp1024 | 279.86 ± 45.63 | 100% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | pp2048 | 320.77 ± 5.00 | 124% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg128 | 43.51 ± 0.05 | 111% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg256 | 43.35 ± 0.09 | 110% |
| qwen2 3B Q4_0_8_8 | 1.69 GiB | 3.09 B | CPU | 64 | tg512 | 42.60 ± 0.31 | 105% |
Q4_0_8_8 offers a nice bump to prompt processing and a small bump to text generation
</details>
</details>
## Which file should I choose?
<details>
<summary>Click here for details</summary>
A great write up with charts showing various performances is provided by Artefact2 [here](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
The first thing to figure out is how big a model you can run. To do this, you'll need to figure out how much RAM and/or VRAM you have.
If you want your model running as FAST as possible, you'll want to fit the whole thing on your GPU's VRAM. Aim for a quant with a file size 1-2GB smaller than your GPU's total VRAM.
If you want the absolute maximum quality, add both your system RAM and your GPU's VRAM together, then similarly grab a quant with a file size 1-2GB Smaller than that total.
Next, you'll need to decide if you want to use an 'I-quant' or a 'K-quant'.
If you don't want to think too much, grab one of the K-quants. These are in format 'QX_K_X', like Q5_K_M.
If you want to get more into the weeds, you can check out this extremely useful feature chart:
[llama.cpp feature matrix](https://github.com/ggml-org/llama.cpp/wiki/Feature-matrix)
But basically, if you're aiming for below Q4, and you're running cuBLAS (Nvidia) or rocBLAS (AMD), you should look towards the I-quants. These are in format IQX_X, like IQ3_M. These are newer and offer better performance for their size.
These I-quants can also be used on CPU, but will be slower than their K-quant equivalent, so speed vs performance is a tradeoff you'll have to decide.
</details>
## Credits
Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset.
Thank you ZeroWw for the inspiration to experiment with embed/output.
Thank you to LM Studio for sponsoring my work.
Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a0daed47ad96f8956438aac5ac74c789d332f387d134c164c342c9b0b308ab09
size 8114054848

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:42c2ff096e6df3736473f4adc375d79534eddeb15ffc582d16812a36cf41162d
size 7478355648

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:942564c2e2a3753427055ce2788ce643dbeac7f6aecf3c6e1100187b839b914b
size 7207036608

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:fe4adb663eb5d565b2cf860a5c735012259f0693ba774cb60d3224bb76ee6b53
size 10650953408

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4e6ef75eb88f4f017e5cafcaff4b51ea15ea7c17a1dceb7093bcd1923f61f3e4
size 9907119808

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2c70edb1898bdd6ceb682560ac6511e218da3e6eee2d642ffcc73b7adf75cb9e
size 9280595648

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bfd5b42ded4e03bbebdb7f56e85886d825bbb2840c79e541bf11824c4eebc634
size 13468018368

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9b428fea02b10665b6e3477202f5891fd9961fe72dfe0928791d8c88c2e9f1f1
size 12758918848

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d86784dd3547f4e6c2781169ecde004ba1b9bf9da9dcf648c851e97b2a612b57
size 8890328768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f9f3f89028c6f3ce3c139602eb28474db6b97a0437487aa9d32e3df45bcb93b7
size 9545688768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8ffef8b861ba7d7c66ea671e44d6b6bc25a5702c5272bb9cd3b89c1e97250a94
size 12400764608

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f8d2b8f7b15132bdab7e678ae47b8b330dda3f8c03b52d8ff4e99937851c905b
size 11474085568

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:387fab458c201670729184848b81f744f3e00ac5cb166648d214379bd07d7d97
size 10400278208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3f4dfb2b8570d6bca5e8095cb700111dfa5aa2a8411b38ed10d7cc84118a6d53
size 12987967168

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0e0c54a1f472ab0877cccd0cd193705bd847f6af81ed1c3ebb02f3a20d1034de
size 13494232768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7d239b34414c03db98381c46558bb6a2c064a945d7c04b5267a33bf493854ab5
size 14873110208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:fc742ac6b551c00b50815563fa5db3391722d4aea0795b6431a589baee4216fc
size 14831986368

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:93675b8cae2a0d7bc6be30faa47d2cb4595821900c1aa141d0159fd262eb91d8
size 14333912768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c6731566c8e5f7289b82f9c5df520cbc4f93bae5d75c550bd60b377737042213
size 13549283008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2af2a8f080977f76a87da2108623728ce49e2d57bc949a7d67ff2353d9e66303
size 17178175168

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:708039e7794100a0c75b3248c255bf836a5d25c055e60317d0af2f5f916a6d85
size 16763987648

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8aa2a2202be1b475094fbd9ccdc122f0cbecb96e76040c1e960b4b41df9d2144
size 16304416448

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9f15f3d9b85c0c76bde30e75f3efc24e25c846cc991899ec6266fdd523f88782
size 19345942208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ff1264def0133eb08cdc0156324df34046da08c43c954938498efb7977e9b7b2
size 19671000768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:027815a41378592c0b3524ae7924bec680d5e1cc3c767939423b755bed8a0807
size 25054783168

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:aad0bfe4a91045054c2cfe1384e83654e5483fc1a9f73c7e7e6cae15b6c33aa1
size 47153522080

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:cb4a31eaad4fb33ebac369772ca14167828a418c94a375fdea5b7026a1de2087
size 10037344

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}