Model: kms7530/qwen2.5-0.5B-RAG-ko Source: Original Platform
license, license_link, language, pipeline_tag, base_model, tags, library_name
| license | license_link | language | pipeline_tag | base_model | tags | library_name | |||
|---|---|---|---|---|---|---|---|---|---|
| apache-2.0 | https://huggingface.co/Qwen/Qwen2.5-0.5B-Instruct/blob/main/LICENSE |
|
text-generation | Qwen/Qwen2.5-0.5B-Instruct |
|
mlx |
kms7530/qwen2.5-0.5B-RAG-ko
This model kms7530/qwen2.5-0.5B-RAG-ko was converted to MLX format from Qwen/Qwen2.5-0.5B-Instruct using mlx-lm version 0.22.2.
Use with mlx
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("kms7530/qwen2.5-0.5B-RAG-ko")
prompt = "hello"
if tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)
Description