commit ca7ffd793e288e50daf0355458e3f368beac1780 Author: ModelHub XC Date: Wed Jun 3 20:49:22 2026 +0800 初始化项目,由ModelHub XC社区提供模型 Model: mradermacher/Llama3.1-8B-ShiningValiant2-GGUF Source: Original Platform diff --git a/.gitattributes b/.gitattributes new file mode 100644 index 0000000..75c8040 --- /dev/null +++ b/.gitattributes @@ -0,0 +1,48 @@ +*.7z filter=lfs diff=lfs merge=lfs -text +*.arrow filter=lfs diff=lfs merge=lfs -text +*.bin filter=lfs diff=lfs merge=lfs -text +*.bz2 filter=lfs diff=lfs merge=lfs -text +*.ckpt filter=lfs diff=lfs merge=lfs -text +*.ftz filter=lfs diff=lfs merge=lfs -text +*.gz filter=lfs diff=lfs merge=lfs -text +*.h5 filter=lfs diff=lfs merge=lfs -text +*.joblib filter=lfs diff=lfs merge=lfs -text +*.lfs.* filter=lfs diff=lfs merge=lfs -text +*.mlmodel filter=lfs diff=lfs merge=lfs -text +*.model filter=lfs diff=lfs merge=lfs -text +*.msgpack filter=lfs diff=lfs merge=lfs -text +*.npy filter=lfs diff=lfs merge=lfs -text +*.npz filter=lfs diff=lfs merge=lfs -text +*.onnx filter=lfs diff=lfs merge=lfs -text +*.ot filter=lfs diff=lfs merge=lfs -text +*.parquet filter=lfs diff=lfs merge=lfs -text +*.pb filter=lfs diff=lfs merge=lfs -text +*.pickle filter=lfs diff=lfs merge=lfs -text +*.pkl filter=lfs diff=lfs merge=lfs -text +*.pt filter=lfs diff=lfs merge=lfs -text +*.pth filter=lfs diff=lfs merge=lfs -text +*.rar filter=lfs diff=lfs merge=lfs -text +*.safetensors filter=lfs diff=lfs merge=lfs -text +saved_model/**/* filter=lfs diff=lfs merge=lfs -text +*.tar.* filter=lfs diff=lfs merge=lfs -text +*.tar filter=lfs diff=lfs merge=lfs -text +*.tflite filter=lfs diff=lfs merge=lfs -text +*.tgz filter=lfs diff=lfs merge=lfs -text +*.wasm filter=lfs diff=lfs merge=lfs -text +*.xz filter=lfs diff=lfs merge=lfs -text +*.zip filter=lfs diff=lfs merge=lfs -text +*.zst filter=lfs diff=lfs merge=lfs -text +*tfevents* filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.f16.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text +Llama3.1-8B-ShiningValiant2.Q4_0_4_4.gguf filter=lfs diff=lfs merge=lfs -text diff --git a/Llama3.1-8B-ShiningValiant2.IQ4_XS.gguf b/Llama3.1-8B-ShiningValiant2.IQ4_XS.gguf new file mode 100644 index 0000000..8038985 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.IQ4_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:98de0722b064eb92cc93b133a3a28bada025da5f86b5e9a8fdc57e79a3a7042c +size 4484368832 diff --git a/Llama3.1-8B-ShiningValiant2.Q2_K.gguf b/Llama3.1-8B-ShiningValiant2.Q2_K.gguf new file mode 100644 index 0000000..985422b --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q2_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:00ce7249f9ad795ca941e91043144796436c656cbb546b826c3ab81f6a968100 +size 3179137472 diff --git a/Llama3.1-8B-ShiningValiant2.Q3_K_L.gguf b/Llama3.1-8B-ShiningValiant2.Q3_K_L.gguf new file mode 100644 index 0000000..a54f4e1 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q3_K_L.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:7c4caa97b30e6e611f4218eeec5af039cb805145f28cd118b54e0eb11023ee1a +size 4321962432 diff --git a/Llama3.1-8B-ShiningValiant2.Q3_K_M.gguf b/Llama3.1-8B-ShiningValiant2.Q3_K_M.gguf new file mode 100644 index 0000000..ce355e1 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q3_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:0540ae63f274e47d38846ef036d48436b5f4c5136653eb850362affdf2baf57e +size 4018923968 diff --git a/Llama3.1-8B-ShiningValiant2.Q3_K_S.gguf b/Llama3.1-8B-ShiningValiant2.Q3_K_S.gguf new file mode 100644 index 0000000..9a5ebc5 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q3_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:17663bd7db7edc4049e875ea0f1ef1759aedf1878e6df90da70c7cec760fc746 +size 3664505280 diff --git a/Llama3.1-8B-ShiningValiant2.Q4_0_4_4.gguf b/Llama3.1-8B-ShiningValiant2.Q4_0_4_4.gguf new file mode 100644 index 0000000..e604643 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q4_0_4_4.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:cabd50fcc1b13bd48f31682bdfd76afec167583ad0ca689305d0a83c229dfba3 +size 4661217728 diff --git a/Llama3.1-8B-ShiningValiant2.Q4_K_M.gguf b/Llama3.1-8B-ShiningValiant2.Q4_K_M.gguf new file mode 100644 index 0000000..32e09d7 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q4_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:2a3c7e4b3af476fc41429c0810979d097dbcfe84b4ccf26d6a59586c093379d6 +size 4920740288 diff --git a/Llama3.1-8B-ShiningValiant2.Q4_K_S.gguf b/Llama3.1-8B-ShiningValiant2.Q4_K_S.gguf new file mode 100644 index 0000000..d993073 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q4_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:61f9be87b4c5e999f72f3e5888dc4770e264b19f164260856fb98688cc25b579 +size 4692675008 diff --git a/Llama3.1-8B-ShiningValiant2.Q5_K_M.gguf b/Llama3.1-8B-ShiningValiant2.Q5_K_M.gguf new file mode 100644 index 0000000..2ab2794 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q5_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:a3326f4588497f199d04166629ebf337c1491cbd60049c835748d73ee3e778c0 +size 5732993472 diff --git a/Llama3.1-8B-ShiningValiant2.Q5_K_S.gguf b/Llama3.1-8B-ShiningValiant2.Q5_K_S.gguf new file mode 100644 index 0000000..95ed89f --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q5_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:dd5dbe5ff89fd04e9b1c8c4c849941a225994ace6269415ced7fe1e0bc0786ef +size 5599300032 diff --git a/Llama3.1-8B-ShiningValiant2.Q6_K.gguf b/Llama3.1-8B-ShiningValiant2.Q6_K.gguf new file mode 100644 index 0000000..ba0b0fd --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q6_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:89a52a19616ef3386992759480b1b05166257ea9e4a286fc29fcd4b465176e66 +size 6596012480 diff --git a/Llama3.1-8B-ShiningValiant2.Q8_0.gguf b/Llama3.1-8B-ShiningValiant2.Q8_0.gguf new file mode 100644 index 0000000..db331f9 --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.Q8_0.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:c3f2e9f6c180b955c7be067f91515e7ba5f3cfb9bd8d8ac45e3b096dba3971d4 +size 8540776896 diff --git a/Llama3.1-8B-ShiningValiant2.f16.gguf b/Llama3.1-8B-ShiningValiant2.f16.gguf new file mode 100644 index 0000000..bdeb94f --- /dev/null +++ b/Llama3.1-8B-ShiningValiant2.f16.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:86d1cc50f0483c812b424af389b9227f9d9d00b290172aef6b27ea2758b85510 +size 16068897216 diff --git a/README.md b/README.md new file mode 100644 index 0000000..e4f5305 --- /dev/null +++ b/README.md @@ -0,0 +1,94 @@ +--- +base_model: ValiantLabs/Llama3.1-8B-ShiningValiant2 +datasets: +- sequelbox/Celestia +- sequelbox/Spurline +- sequelbox/Supernova +language: +- en +library_name: transformers +license: llama3.1 +model_type: llama +quantized_by: mradermacher +tags: +- shining-valiant +- shining-valiant-2 +- valiant +- valiant-labs +- llama +- llama-3.1 +- llama-3.1-instruct +- llama-3.1-instruct-8b +- llama-3 +- llama-3-instruct +- llama-3-instruct-8b +- 8b +- science +- physics +- biology +- chemistry +- compsci +- computer-science +- engineering +- technical +- conversational +- chat +- instruct +--- +## About + + + + + + +static quants of https://huggingface.co/ValiantLabs/Llama3.1-8B-ShiningValiant2 + + +weighted/imatrix quants are available at https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-i1-GGUF +## Usage + +If you are unsure how to use GGUF files, refer to one of [TheBloke's +READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for +more details, including on how to concatenate multi-part files. + +## Provided Quants + +(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants) + +| Link | Type | Size/GB | Notes | +|:-----|:-----|--------:|:------| +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q2_K.gguf) | Q2_K | 3.3 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q3_K_S.gguf) | Q3_K_S | 3.8 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q3_K_M.gguf) | Q3_K_M | 4.1 | lower quality | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q3_K_L.gguf) | Q3_K_L | 4.4 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.IQ4_XS.gguf) | IQ4_XS | 4.6 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q4_0_4_4.gguf) | Q4_0_4_4 | 4.8 | fast on arm, low quality | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q4_K_S.gguf) | Q4_K_S | 4.8 | fast, recommended | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q4_K_M.gguf) | Q4_K_M | 5.0 | fast, recommended | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q5_K_S.gguf) | Q5_K_S | 5.7 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q5_K_M.gguf) | Q5_K_M | 5.8 | | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q6_K.gguf) | Q6_K | 6.7 | very good quality | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.Q8_0.gguf) | Q8_0 | 8.6 | fast, best quality | +| [GGUF](https://huggingface.co/mradermacher/Llama3.1-8B-ShiningValiant2-GGUF/resolve/main/Llama3.1-8B-ShiningValiant2.f16.gguf) | f16 | 16.2 | 16 bpw, overkill | + +Here is a handy graph by ikawrakow comparing some lower-quality quant +types (lower is better): + +![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png) + +And here are Artefact2's thoughts on the matter: +https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9 + +## FAQ / Model Request + +See https://huggingface.co/mradermacher/model_requests for some answers to +questions you might have and/or if you want some other model quantized. + +## Thanks + +I thank my company, [nethype GmbH](https://www.nethype.de/), for letting +me use its servers and providing upgrades to my workstation to enable +this work in my free time. + +