commit 399e468ff26d32b559cdd4301d881eb7426cc204 Author: ModelHub XC Date: Thu Jul 2 00:08:14 2026 +0800 初始化项目,由ModelHub XC社区提供模型 Model: MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF Source: Original Platform diff --git a/.gitattributes b/.gitattributes new file mode 100644 index 0000000..4c19eb1 --- /dev/null +++ b/.gitattributes @@ -0,0 +1,64 @@ +*.7z filter=lfs diff=lfs merge=lfs -text +*.arrow filter=lfs diff=lfs merge=lfs -text +*.bin filter=lfs diff=lfs merge=lfs -text +*.bin.* filter=lfs diff=lfs merge=lfs -text +*.bz2 filter=lfs diff=lfs merge=lfs -text +*.ftz filter=lfs diff=lfs merge=lfs -text +*.gz filter=lfs diff=lfs merge=lfs -text +*.h5 filter=lfs diff=lfs merge=lfs -text +*.joblib filter=lfs diff=lfs merge=lfs -text +*.lfs.* filter=lfs diff=lfs merge=lfs -text +*.model filter=lfs diff=lfs merge=lfs -text +*.msgpack filter=lfs diff=lfs merge=lfs -text +*.onnx filter=lfs diff=lfs merge=lfs -text +*.ot filter=lfs diff=lfs merge=lfs -text +*.parquet filter=lfs diff=lfs merge=lfs -text +*.pb filter=lfs diff=lfs merge=lfs -text +*.pt filter=lfs diff=lfs merge=lfs -text +*.pth filter=lfs diff=lfs merge=lfs -text +*.rar filter=lfs diff=lfs merge=lfs -text +saved_model/**/* filter=lfs diff=lfs merge=lfs -text +*.tar.* filter=lfs diff=lfs merge=lfs -text +*.tflite filter=lfs diff=lfs merge=lfs -text +*.tgz filter=lfs diff=lfs merge=lfs -text +*.xz filter=lfs diff=lfs merge=lfs -text +*.zip filter=lfs diff=lfs merge=lfs -text +*.zstandard filter=lfs diff=lfs merge=lfs -text +*.tfevents* filter=lfs diff=lfs merge=lfs -text +*.db* filter=lfs diff=lfs merge=lfs -text +*.ark* filter=lfs diff=lfs merge=lfs -text +**/*ckpt*data* filter=lfs diff=lfs merge=lfs -text +**/*ckpt*.meta filter=lfs diff=lfs merge=lfs -text +**/*ckpt*.index filter=lfs diff=lfs merge=lfs -text +*.safetensors filter=lfs diff=lfs merge=lfs -text +*.ckpt filter=lfs diff=lfs merge=lfs -text + +*.ggml filter=lfs diff=lfs merge=lfs -text +*.llamafile* filter=lfs diff=lfs merge=lfs -text +*.pt2 filter=lfs diff=lfs merge=lfs -text +*.mlmodel filter=lfs diff=lfs merge=lfs -text +*.npy filter=lfs diff=lfs merge=lfs -text +*.npz filter=lfs diff=lfs merge=lfs -text +*.pickle filter=lfs diff=lfs merge=lfs -text +*.pkl filter=lfs diff=lfs merge=lfs -text +*.tar filter=lfs diff=lfs merge=lfs -text +*.wasm filter=lfs diff=lfs merge=lfs -text +*.zst filter=lfs diff=lfs merge=lfs -text +*tfevents* filter=lfs diff=lfs merge=lfs -text + +Mistral-7B-Instruct-v0.3.Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.fp16.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text +Mistral-7B-Instruct-v0.3.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text \ No newline at end of file diff --git a/Mistral-7B-Instruct-v0.3.IQ1_M.gguf b/Mistral-7B-Instruct-v0.3.IQ1_M.gguf new file mode 100644 index 0000000..9d3a2e4 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.IQ1_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:e241ecfd88517c9137220d5869ab461853ec80f27ec7241b7ba490e341e2e32f +size 1757663392 diff --git a/Mistral-7B-Instruct-v0.3.IQ1_S.gguf b/Mistral-7B-Instruct-v0.3.IQ1_S.gguf new file mode 100644 index 0000000..be567ef --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.IQ1_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:0bff52a0d005f697a8326214ae37a13429c6af911dc32fbde004bc23a0dca350 +size 1615319200 diff --git a/Mistral-7B-Instruct-v0.3.IQ2_XS.gguf b/Mistral-7B-Instruct-v0.3.IQ2_XS.gguf new file mode 100644 index 0000000..9dc0371 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.IQ2_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:213c75f43c21dca7f6fdf5e4eb7ad7e823727a56800a777d7db549c3e4f86cec +size 2201473184 diff --git a/Mistral-7B-Instruct-v0.3.IQ3_XS.gguf b/Mistral-7B-Instruct-v0.3.IQ3_XS.gguf new file mode 100644 index 0000000..ab1c6ac --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.IQ3_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:d70bef5e708d5245e7319391495b94e3952cb9b3cb2ce3b45f4701b17fc28219 +size 3022770336 diff --git a/Mistral-7B-Instruct-v0.3.IQ4_XS.gguf b/Mistral-7B-Instruct-v0.3.IQ4_XS.gguf new file mode 100644 index 0000000..8c84577 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.IQ4_XS.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:3fe60995f215969f08926dd61d8a7c81377b2935a6c8c865ba53885fe2e8b6d1 +size 3911962784 diff --git a/Mistral-7B-Instruct-v0.3.Q2_K.gguf b/Mistral-7B-Instruct-v0.3.Q2_K.gguf new file mode 100644 index 0000000..ec52774 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q2_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:f1e535fb88726ab341bbd2896041fe72d376ac3670d343caad7c2991ecb5b82f +size 2722877600 diff --git a/Mistral-7B-Instruct-v0.3.Q3_K_L.gguf b/Mistral-7B-Instruct-v0.3.Q3_K_L.gguf new file mode 100644 index 0000000..9a46ffc --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q3_K_L.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:ac7c3208985dc337fd9ad8c7674f3e4a141b2c27de62f291125e93345987d8c0 +size 3825979552 diff --git a/Mistral-7B-Instruct-v0.3.Q3_K_M.gguf b/Mistral-7B-Instruct-v0.3.Q3_K_M.gguf new file mode 100644 index 0000000..e88d658 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q3_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:b982905e4fccdf5894b4c597426fa71f7b528a4534719568a3ee806fe33a961f +size 3522941088 diff --git a/Mistral-7B-Instruct-v0.3.Q3_K_S.gguf b/Mistral-7B-Instruct-v0.3.Q3_K_S.gguf new file mode 100644 index 0000000..c5b2ffa --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q3_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:ad5a1c665f1a2c7df4e585c673c5d37338505f55290475338bfec4972069dd87 +size 3168522400 diff --git a/Mistral-7B-Instruct-v0.3.Q4_K_M.gguf b/Mistral-7B-Instruct-v0.3.Q4_K_M.gguf new file mode 100644 index 0000000..93bf05f --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q4_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:14850c84ff9f06e9b51d505d64815d5cc0cea0257380353ac0b3d21b21f6e024 +size 4372811936 diff --git a/Mistral-7B-Instruct-v0.3.Q4_K_S.gguf b/Mistral-7B-Instruct-v0.3.Q4_K_S.gguf new file mode 100644 index 0000000..c3b4929 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q4_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:45ef6dda7fe3a6058bc915ce3f5cfcd9a4eddf2ae2046aa91d07b68efff038f8 +size 4144746656 diff --git a/Mistral-7B-Instruct-v0.3.Q5_K_M.gguf b/Mistral-7B-Instruct-v0.3.Q5_K_M.gguf new file mode 100644 index 0000000..2c7e3a9 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q5_K_M.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:16617265368ffaa5f0c67b320a1df4b40b0c8e990a91cf5722fc2159e65f559f +size 5136175264 diff --git a/Mistral-7B-Instruct-v0.3.Q5_K_S.gguf b/Mistral-7B-Instruct-v0.3.Q5_K_S.gguf new file mode 100644 index 0000000..09f0b05 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q5_K_S.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:1dd8f3de99bb63143a3828f65ab4c81f54fcb1d1af4711904c50a2801b64e088 +size 5002481824 diff --git a/Mistral-7B-Instruct-v0.3.Q6_K.gguf b/Mistral-7B-Instruct-v0.3.Q6_K.gguf new file mode 100644 index 0000000..7c5727f --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q6_K.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:d58a20f828bca2e163342d43324f953f2edf9bdd5886bfe15c4b81b5b70a3b7b +size 5947248800 diff --git a/Mistral-7B-Instruct-v0.3.Q8_0.gguf b/Mistral-7B-Instruct-v0.3.Q8_0.gguf new file mode 100644 index 0000000..668d4cd --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.Q8_0.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:24df553dc0e725196fe8a3c7be1edfe6ff17a0fe855f508b3f4a0e444e2e4281 +size 7702565024 diff --git a/Mistral-7B-Instruct-v0.3.fp16.gguf b/Mistral-7B-Instruct-v0.3.fp16.gguf new file mode 100644 index 0000000..ab9c065 --- /dev/null +++ b/Mistral-7B-Instruct-v0.3.fp16.gguf @@ -0,0 +1,3 @@ +version https://git-lfs.github.com/spec/v1 +oid sha256:099a38de1702728b81aeedd337e704d649bfdedf5113952f9e0ade34ffa57022 +size 14497337312 diff --git a/README.md b/README.md new file mode 100644 index 0000000..e3e7739 --- /dev/null +++ b/README.md @@ -0,0 +1,55 @@ +--- +tags: +- quantized +- 2-bit +- 3-bit +- 4-bit +- 5-bit +- 6-bit +- 8-bit +- GGUF +- transformers +- safetensors +- mistral +- text-generation +- conversational +- license:apache-2.0 +- autotrain_compatible +- endpoints_compatible +- text-generation-inference +- region:us +- text-generation +model_name: Mistral-7B-Instruct-v0.3-GGUF +base_model: mistralai/Mistral-7B-Instruct-v0.3 +inference: false +model_creator: mistralai +pipeline_tag: text-generation +quantized_by: MaziyarPanahi +--- +# [MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF](https://huggingface.co/MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF) +- Model creator: [mistralai](https://huggingface.co/mistralai) +- Original model: [mistralai/Mistral-7B-Instruct-v0.3](https://huggingface.co/mistralai/Mistral-7B-Instruct-v0.3) + +## Description +[MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF](https://huggingface.co/MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF) contains GGUF format model files for [mistralai/Mistral-7B-Instruct-v0.3](https://huggingface.co/mistralai/Mistral-7B-Instruct-v0.3). + +### About GGUF + +GGUF is a new format introduced by the llama.cpp team on August 21st 2023. It is a replacement for GGML, which is no longer supported by llama.cpp. + +Here is an incomplete list of clients and libraries that are known to support GGUF: + +* [llama.cpp](https://github.com/ggerganov/llama.cpp). The source project for GGUF. Offers a CLI and a server option. +* [llama-cpp-python](https://github.com/abetlen/llama-cpp-python), a Python library with GPU accel, LangChain support, and OpenAI-compatible API server. +* [LM Studio](https://lmstudio.ai/), an easy-to-use and powerful local GUI for Windows and macOS (Silicon), with GPU acceleration. Linux available, in beta as of 27/11/2023. +* [text-generation-webui](https://github.com/oobabooga/text-generation-webui), the most widely used web UI, with many features and powerful extensions. Supports GPU acceleration. +* [KoboldCpp](https://github.com/LostRuins/koboldcpp), a fully featured web UI, with GPU accel across all platforms and GPU architectures. Especially good for story telling. +* [GPT4All](https://gpt4all.io/index.html), a free and open source local running GUI, supporting Windows, Linux and macOS with full GPU accel. +* [LoLLMS Web UI](https://github.com/ParisNeo/lollms-webui), a great web UI with many interesting and unique features, including a full model library for easy model selection. +* [Faraday.dev](https://faraday.dev/), an attractive and easy to use character-based chat GUI for Windows and macOS (both Silicon and Intel), with GPU acceleration. +* [candle](https://github.com/huggingface/candle), a Rust ML framework with a focus on performance, including GPU support, and ease of use. +* [ctransformers](https://github.com/marella/ctransformers), a Python library with GPU accel, LangChain support, and OpenAI-compatible AI server. Note, as of time of writing (November 27th 2023), ctransformers has not been updated in a long time and does not support many recent models. + +## Special thanks + +🙏 Special thanks to [Georgi Gerganov](https://github.com/ggerganov) and the whole team working on [llama.cpp](https://github.com/ggerganov/llama.cpp/) for making all of this possible. \ No newline at end of file diff --git a/config.json b/config.json new file mode 100644 index 0000000..9f0f76f --- /dev/null +++ b/config.json @@ -0,0 +1,3 @@ +{ + "model_type": "mistral" +} \ No newline at end of file diff --git a/configuration.json b/configuration.json new file mode 100644 index 0000000..bbeeda1 --- /dev/null +++ b/configuration.json @@ -0,0 +1 @@ +{"framework": "pytorch", "task": "text-generation", "allow_remote": true} \ No newline at end of file