Files
qwen3-4b-think-baseline-lor…/README.md
ModelHub XC 16182a07e7 初始化项目,由ModelHub XC社区提供模型
Model: modrill/qwen3-4b-think-baseline-lora-sft
Source: Original Platform
2026-07-27 03:36:10 +08:00

1.4 KiB

license, base_model, tags, language, library_name, pipeline_tag
license base_model tags language library_name pipeline_tag
apache-2.0 Qwen/Qwen3-4B-Base
qwen3
code
sft
lora
lora-merged
think
en
zh
transformers text-generation

Qwen3-4B Code SFT - Think Baseline (LoRA Merged)

LoRA supervised fine-tuning of Qwen/Qwen3-4B-Base, with adapters merged into full weights for direct inference.

This repo is LoRA-based SFT (rank 64), not full-parameter fine-tuning. For native full SFT weights, see modrill/qwen3-4b-think-baseline-full-sft.

Model Details

  • Base model: Qwen/Qwen3-4B-Base
  • Fine-tuning: LoRA SFT (rank 64, alpha 128), merged into full weights
  • Mode: Think (enable_thinking=true)
  • Training cutoff length: 24576 tokens

Usage

from transformers import AutoModelForCausalLM, AutoTokenizer

model_id = "modrill/qwen3-4b-think-baseline-lora-sft"
tok = AutoTokenizer.from_pretrained(model_id, trust_remote_code=True)
model = AutoModelForCausalLM.from_pretrained(
    model_id, trust_remote_code=True, torch_dtype="auto", device_map="auto"
)

Inference Tips

  • Set enable_thinking=true in chat template
  • Recommended max_tokens: 24576

License

Apache 2.0, consistent with the Qwen3 base model license.