Model: Raghav-Singhal/pretrain-normal-smollm-1p7b-100B-20n-2048sl-960gbsz Source: Original Platform
library_name, pipeline_tag, tags
| library_name | pipeline_tag | tags | |||
|---|---|---|---|---|---|
| transformers | text-generation |
|
pretrain-normal-smollm-1p7b-100B-20n-2048sl-960gbsz
Converted Hugging Face base checkpoint from the Model Raising pretraining run.
Details
- Architecture:
LlamaForCausalLM - Base model size:
1.7B - Precision on disk:
bfloat16 - Tokenizer:
HuggingFaceTB/SmolLM2-1.7B-Instruct
This repo contains the verified Hugging Face export of the final pretraining checkpoint before SFT.
Description
Model synced from source: Raghav-Singhal/pretrain-normal-smollm-1p7b-100B-20n-2048sl-960gbsz
Languages
Jinja
100%