初始化项目,由ModelHub XC社区提供模型

Model: mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-09-04 17:51:36 +08:00
commit 922f4289ab
8 changed files with 224 additions and 0 deletions

24
README.md Normal file
View File

@@ -0,0 +1,24 @@
---
base_model: Qwen/Qwen2.5-1.5B-Instruct
tags:
- safe-flat-lora
- baseline
- lora
- qwen
license: apache-2.0
---
# mahdiboughrous/safe-flat-lora-baseline-qwen2.5-1.5b-instruct-merged
Part of the **Safe-Flat-LoRA** project (the Safety-Flatness Paradox in 4-bit
quantization of LoRA-fine-tuned LLMs -- see the project blueprint for the
full motivation and methodology).
- **Base model:** [`Qwen/Qwen2.5-1.5B-Instruct`](https://huggingface.co/Qwen/Qwen2.5-1.5B-Instruct)
- **Method:** `baseline` (plain LoRA fine-tune, used as the pre-Flat-LoRA baseline)
- **Artifact:** merged
- **Fine-tuning task:** SQL generation (`b-mc2/sql-create-context`), used as a
downstream-task proxy to study how post-training 4-bit (NF4) quantization
shifts safety-refusal behavior after LoRA fine-tuning.
Base model Qwen2.5-1.5B-Instruct is Apache-2.0.