初始化项目,由ModelHub XC社区提供模型

Model: Mungert/Qwen2.5-0.5B-Instruct-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-21 17:30:09 +08:00
commit 7a80082d73
26 changed files with 246 additions and 0 deletions

64
.gitattributes vendored Normal file
View File

@@ -0,0 +1,64 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q5_k_s.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq4_nl.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q4_0.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq4_xs.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq3_xs.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct.imatrix filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q5_k_m.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q5_k_l.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q3_k_s.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q3_k_l.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-bf16-q4_k.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q4_k_m.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q4_1.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q6_k_m.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q6_k_l.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-f16-q4_k.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-bf16-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q2_k_l.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q3_k_m.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-bf16-q6_k.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q4_k_l.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q4_k_s.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-f16-q6_k.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-f16-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-bf16.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq3_s.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq3_xxs.gguf filter=lfs diff=lfs merge=lfs -text
Qwen2.5-0.5B-Instruct-iq3_m.gguf filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8f7eb061d15b0376123e1e139571279a7d6bbd482df38b8f1d2e28f1826c3634
size 525434432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1cb09231cadda4432a48842ec39758fbd2d30f9ce062184de8699365e6b29584
size 633363008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c7f45fa6fbb4184b814246a506052001c3497c3dd77368b0a5b91d0d9545d18e
size 658694432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:00b0a4872e000e204aa019344419dcbd29ba53dc5cf1cca6f20844bd95ac927a
size 994156832

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:19bb1ea0596956459bfc96c62b799f926c8c9cb583b7a8f9263a0f64c4e004ac
size 525434432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7f5cdfb6ebabc8810d307a0e0510dd565f7868929abd6bea74aedf0829404749
size 633363008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ec0522470fd64a367c05a69030a28084f5b9ce559c47497556bec39e4ba3c21c
size 658694432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:10b8a4dee340ee040dc2f10eacfb327e4209f2c1a579f6be2155fd28c257d055
size 300210496

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e5a055431dc70683b368b9f06221e8270b727ca6a8534845f6cd8970652285e9
size 296065600

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4d50218705bc0d9337af295c82483ab23b8bc91bfaf021c6aa521680041799f1
size 296065600

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:097c2be61c1abc720db7e0457ac245cb30314a496506d6fb5a940eba93d0d3fe
size 291162688

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5949c2411a0c4b2d5c0369d89749e912df57e55327937b604ea92fe571d51fb7
size 352671296

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8f5eb44a9f43ae6cb06861388a40878d6bd1dc9529681fb49c7df8ff01003fbc
size 349402688

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6ef723d02ca972d0303501a29ce5d621b4b8071dd961d370d5ce6f219efd7bc7
size 355466816

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:08b9db4042cbd51641ddb6ef7ba4ec56df16a09ead752cee7200a0a201a6e5ad
size 338263616

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:32f6f7ec659438377387f380b4d7709a5972dafa25363710fa01dfda1e5cb83c
size 352155200

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7d367251cfdbe89c679a789c7605cbad8d0fa36195b89a4c8206af297f818b7c
size 374519360

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9c2eaa626149ec66d39e2ef3f7caf317db95b71123687148af3d8a2cf8c69287
size 397808192

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ac7710a4f8b68643af3d78ec64bb9980c52ba9fde08a38a5e5ed27bc3125e897
size 385472064

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f82490dbaa3ebcdf1ed378563f0293f88d24ed334d039e6f75d8f3f84c6ef25e
size 420086336

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c9b2f386c456c24a3aeea7db21d33badea92254f7539d83b5b86364fe01a84b8
size 412710464

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7d629db4087aa947c614e7939f600cb1fae2bc0cbc0fcf56b417e4c82f7b6664
size 505736768

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2a2d5b9827435252c972e50e75606a42117c7df4153c88f12b8d30ca965b4020
size 531068192

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0714eccb3c4ddd904d007db51e4375f161e649f739774b866aed875c7c8f7e00
size 988610

110
README.md Normal file
View File

@@ -0,0 +1,110 @@
---
license: apache-2.0
license_link: https://huggingface.co/Qwen/Qwen2.5-0.5B-Instruct/blob/main/LICENSE
language:
- en
pipeline_tag: text-generation
base_model: Qwen/Qwen2.5-0.5B
tags:
- chat
library_name: transformers
---
# Qwen2.5-0.5B-Instruct
## Introduction
Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:
- Significantly **more knowledge** and has greatly improved capabilities in **coding** and **mathematics**, thanks to our specialized expert models in these domains.
- Significant improvements in **instruction following**, **generating long texts** (over 8K tokens), **understanding structured data** (e.g, tables), and **generating structured outputs** especially JSON. **More resilient to the diversity of system prompts**, enhancing role-play implementation and condition-setting for chatbots.
- **Long-context Support** up to 128K tokens and can generate up to 8K tokens.
- **Multilingual support** for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more.
**This repo contains the instruction-tuned 0.5B Qwen2.5 model**, which has the following features:
- Type: Causal Language Models
- Training Stage: Pretraining & Post-training
- Architecture: transformers with RoPE, SwiGLU, RMSNorm, Attention QKV bias and tied word embeddings
- Number of Parameters: 0.49B
- Number of Paramaters (Non-Embedding): 0.36B
- Number of Layers: 24
- Number of Attention Heads (GQA): 14 for Q and 2 for KV
- Context Length: Full 32,768 tokens and generation 8192 tokens
For more details, please refer to our [blog](https://qwenlm.github.io/blog/qwen2.5/), [GitHub](https://github.com/QwenLM/Qwen2.5), and [Documentation](https://qwen.readthedocs.io/en/latest/).
## Requirements
The code of Qwen2.5 has been in the latest Hugging face `transformers` and we advise you to use the latest version of `transformers`.
With `transformers<4.37.0`, you will encounter the following error:
```
KeyError: 'qwen2'
```
## Quickstart
Here provides a code snippet with `apply_chat_template` to show you how to load the tokenizer and model and how to generate contents.
```python
from transformers import AutoModelForCausalLM, AutoTokenizer
model_name = "Qwen/Qwen2.5-0.5B-Instruct"
model = AutoModelForCausalLM.from_pretrained(
model_name,
torch_dtype="auto",
device_map="auto"
)
tokenizer = AutoTokenizer.from_pretrained(model_name)
prompt = "Give me a short introduction to large language model."
messages = [
{"role": "system", "content": "You are Qwen, created by Alibaba Cloud. You are a helpful assistant."},
{"role": "user", "content": prompt}
]
text = tokenizer.apply_chat_template(
messages,
tokenize=False,
add_generation_prompt=True
)
model_inputs = tokenizer([text], return_tensors="pt").to(model.device)
generated_ids = model.generate(
**model_inputs,
max_new_tokens=512
)
generated_ids = [
output_ids[len(input_ids):] for input_ids, output_ids in zip(model_inputs.input_ids, generated_ids)
]
response = tokenizer.batch_decode(generated_ids, skip_special_tokens=True)[0]
```
## Evaluation & Performance
Detailed evaluation results are reported in this [📑 blog](https://qwenlm.github.io/blog/qwen2.5/).
For requirements on GPU memory and the respective throughput, see results [here](https://qwen.readthedocs.io/en/latest/benchmark/speed_benchmark.html).
## Citation
If you find our work helpful, feel free to give us a cite.
```
@misc{qwen2.5,
title = {Qwen2.5: A Party of Foundation Models},
url = {https://qwenlm.github.io/blog/qwen2.5/},
author = {Qwen Team},
month = {September},
year = {2024}
}
@article{qwen2,
title={Qwen2 Technical Report},
author={An Yang and Baosong Yang and Binyuan Hui and Bo Zheng and Bowen Yu and Chang Zhou and Chengpeng Li and Chengyuan Li and Dayiheng Liu and Fei Huang and Guanting Dong and Haoran Wei and Huan Lin and Jialong Tang and Jialin Wang and Jian Yang and Jianhong Tu and Jianwei Zhang and Jianxin Ma and Jin Xu and Jingren Zhou and Jinze Bai and Jinzheng He and Junyang Lin and Kai Dang and Keming Lu and Keqin Chen and Kexin Yang and Mei Li and Mingfeng Xue and Na Ni and Pei Zhang and Peng Wang and Ru Peng and Rui Men and Ruize Gao and Runji Lin and Shijie Wang and Shuai Bai and Sinan Tan and Tianhang Zhu and Tianhao Li and Tianyu Liu and Wenbin Ge and Xiaodong Deng and Xiaohuan Zhou and Xingzhang Ren and Xinyu Zhang and Xipin Wei and Xuancheng Ren and Yang Fan and Yang Yao and Yichang Zhang and Yu Wan and Yunfei Chu and Yuqiong Liu and Zeyu Cui and Zhenru Zhang and Zhihao Fan},
journal={arXiv preprint arXiv:2407.10671},
year={2024}
}
```