Files
safe-flat-lora-baseline-qwe…/README.md
ModelHub XC 922f4289ab 初始化项目,由ModelHub XC社区提供模型
Model: mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged
Source: Original Platform
2026-09-04 17:51:36 +08:00

841 B

base_model, tags, license
base_model tags license
Qwen/Qwen2.5-1.5B-Instruct
safe-flat-lora
baseline
lora
qwen
apache-2.0

mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged

Part of the Safe-Flat-LoRA project (the Safety-Flatness Paradox in 4-bit quantization of LoRA-fine-tuned LLMs -- see the project blueprint for the full motivation and methodology).

  • Base model: Qwen/Qwen2.5-1.5B-Instruct
  • Method: baseline (plain LoRA fine-tune, used as the pre-Flat-LoRA baseline)
  • Artifact: merged
  • Fine-tuning task: SQL generation (b-mc2/sql-create-context), used as a downstream-task proxy to study how post-training 4-bit (NF4) quantization shifts safety-refusal behavior after LoRA fine-tuning.

Base model Qwen2.5-1.5B-Instruct is Apache-2.0.