Files
Barcenas-Tiny-1.1b-DPO/README.md
ModelHub XC cf997b6907 初始化项目,由ModelHub XC社区提供模型
Model: Danielbrdz/Barcenas-Tiny-1.1b-DPO
Source: Original Platform
2026-07-03 11:55:17 +08:00

636 B

license, datasets, language
license datasets language
apache-2.0
Intel/orca_dpo_pairs
en
es

Barcenas Tiny 1.1b DPO

It is a model based on the famous TinyLlama/TinyLlama-1.1B-Chat-v1.0 and trained with DPO using the Intel/orca_dpo_pairs dataset.

With its reinforcement based training we hope to improve the Tiny model in a huge way and have a better model with better responses with a small size and accessible to most people.

Many thanks to Maxime Labonne (mlabonne) for his tutorial on how to train a LLM model using DPO, without his tutorial this model would not have been possible.

Made with ❤️ in Guadalupe, Nuevo Leon, Mexico 🇲🇽