--- base_model: Qwen/Qwen2.5-1.5B-Instruct tags: - safe-flat-lora - baseline - lora - qwen license: apache-2.0 --- # mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged Part of the **Safe-Flat-LoRA** project (the Safety-Flatness Paradox in 4-bit quantization of LoRA-fine-tuned LLMs -- see the project blueprint for the full motivation and methodology). - **Base model:** [`Qwen/Qwen2.5-1.5B-Instruct`](https://huggingface.co/Qwen/Qwen2.5-1.5B-Instruct) - **Method:** `baseline` (plain LoRA fine-tune, used as the pre-Flat-LoRA baseline) - **Artifact:** merged - **Fine-tuning task:** SQL generation (`b-mc2/sql-create-context`), used as a downstream-task proxy to study how post-training 4-bit (NF4) quantization shifts safety-refusal behavior after LoRA fine-tuning. Base model Qwen2.5-1.5B-Instruct is Apache-2.0.