Measured from our SLURM logs for this configuration. These are training-loss
observations only — no downstream benchmark evaluation has been run on this
model, so they should not be read as a quality claim.
SLURM job
Steps
First loss
Final loss
45987994
1,554
3.3731
1.1936
46021015
1,554
3.3731
1.1904
Limitations
No benchmark evaluation has been run on this checkpoint. The only reported
numbers are training-loss observations.
Inherits the biases, knowledge cutoff and failure modes of the base model.
Fine-tuned on a single instruction-following dataset; behaviour outside that
distribution is untested.
LoRA adapters were merged into the base weights, so the merged model cannot
be detached from this fine-tune.
Reproducing
Trained by notebooks/fable_distillation_vibethinker-3b_fable-glint_unsloth.ipynb, executed non-interactively with
papermill on a SLURM H100 partition (Unsloth + TRL, LoRA).
Card generated from the training run's own configuration and logs byscripts/generate_hub_model_card.py.