922f4289ab16205709ce5b03e4d4356ef4fa11b8
Model: mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged Source: Original Platform
base_model, tags, license
| base_model | tags | license | ||||
|---|---|---|---|---|---|---|
| Qwen/Qwen2.5-1.5B-Instruct |
|
apache-2.0 |
mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged
Part of the Safe-Flat-LoRA project (the Safety-Flatness Paradox in 4-bit quantization of LoRA-fine-tuned LLMs -- see the project blueprint for the full motivation and methodology).
- Base model:
Qwen/Qwen2.5-1.5B-Instruct - Method:
baseline(plain LoRA fine-tune, used as the pre-Flat-LoRA baseline) - Artifact: merged
- Fine-tuning task: SQL generation (
b-mc2/sql-create-context), used as a downstream-task proxy to study how post-training 4-bit (NF4) quantization shifts safety-refusal behavior after LoRA fine-tuning.
Base model Qwen2.5-1.5B-Instruct is Apache-2.0.
Description
Model synced from source: mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged
Languages
Jinja
100%