ModelHub XC a747a31756 初始化项目,由ModelHub XC社区提供模型
Model: sequelbox/Qwen3-8B-Esper3-PREVIEW
Source: Original Platform
2026-09-05 18:34:14 +08:00

library_name, license, pipeline_tag, base_model
library_name license pipeline_tag base_model
transformers apache-2.0 text-generation
Qwen/Qwen3-8B

Click here to support our open-source dataset and model releases!

This is an early alpha preview of the upcoming Esper 3 series for Qwen 3 - use at your own discretion! The full model is now available, click here!

Esper 3 is a reasoning-chat finetune focused on coding, architecture, DevOps, and general reasoning chat.

All training data generated synthetically by Deepseek-R1 685b model. This sneak preview uses training data from our Titanium, Tachibana, and Raiden series of datasets. Final datasets used will be provided along the full release of Esper 3 - this preview release is only trained on a subselection of the data for early testing. Full model release coming soon!

See the Qwen 3 8b page for sample prompting scripts or further information on the base model. Esper 3 is a reasoning finetune: enable_thinking=True is recommended for all chats.

Try the preview release out, see what you think, tell your friends :)

Please consider supporting our releases if you can. There's still time for a bottom-up AI revolution: the time to make a difference in how this turns out is now!

More Qwen 3 releases to come soon!

Do as you will.

Description
Model synced from source: sequelbox/Qwen3-8B-Esper3-PREVIEW
Readme 12 MiB