Files
OLMoE-1B-7B-0924-SFT/README.md
ModelHub XC 80397fa378 初始化项目,由ModelHub XC社区提供模型
Model: LLM-Research/OLMoE-1B-7B-0924-SFT
Source: Original Platform
2026-05-23 23:02:12 +08:00

2.7 KiB

license, language, tags, co2_eq_emissions, datasets, base_model
license language tags co2_eq_emissions datasets base_model
apache-2.0
en
moe
olmo
olmoe
1
allenai/tulu-v3.1-mix-preview-4096-OLMoE
allenai/OLMoE-1B-7B-0924
OLMoE Logo.

Model Summary

This model is an intermediate training checkpoint during post-training, after the Supervised Fine-Tuning (SFT) step. For best performance, we recommend you use the OLMoE-Instruct version.

Branches:

Citation

@misc{muennighoff2024olmoeopenmixtureofexpertslanguage,
      title={OLMoE: Open Mixture-of-Experts Language Models}, 
      author={Niklas Muennighoff and Luca Soldaini and Dirk Groeneveld and Kyle Lo and Jacob Morrison and Sewon Min and Weijia Shi and Pete Walsh and Oyvind Tafjord and Nathan Lambert and Yuling Gu and Shane Arora and Akshita Bhagia and Dustin Schwenk and David Wadden and Alexander Wettig and Binyuan Hui and Tim Dettmers and Douwe Kiela and Ali Farhadi and Noah A. Smith and Pang Wei Koh and Amanpreet Singh and Hannaneh Hajishirzi},
      year={2024},
      eprint={2409.02060},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2409.02060}, 
}