ModelHub XC 846c9a78c7 初始化项目,由ModelHub XC社区提供模型
Model: OpenPipe/Qwen3-14B-Instruct
Source: Original Platform
2026-07-17 08:05:06 +08:00

library_name, license, license_link, pipeline_tag, base_model
library_name license license_link pipeline_tag base_model
transformers apache-2.0 https://huggingface.co/Qwen/Qwen3-14B/blob/main/LICENSE text-generation
Qwen/Qwen3-14B-Base

Qwen3-14B

Chat

Qwen3-14B-Instruct Highlights

OpenPipe/Qwen3-14B-Instruct is a finetune friendly instruct variant of Qwen3-14B. Qwen3 release does not include a 14B Instruct (non-thinking) model, this fork introduces an updated chat template that makes Qwen3-14B non-thinking by default and be highly compatible with OpenPipe and other finetuning frameworks.

The default Qwen3 chat template does not render <think></think> tags on the previous assistant message, which can lead to inconsistencies between training and generation. This version resolves that issue by adding <think></think> tags to all assistant prompts and generation templates to ensure message format consistency during both training and inference.

The model retains the strong general capabilities of Qwen3-14B while providing a more finetuning friendly chat template.

Model Overview

Qwen3-14B has the following features:

  • Type: Causal Language Models
  • Training Stage: Pretraining & Post-training
  • Number of Parameters: 14.8B
  • Number of Paramaters (Non-Embedding): 13.2B
  • Number of Layers: 40
  • Number of Attention Heads (GQA): 40 for Q and 8 for KV
  • Context Length: 32,768 natively and 131,072 tokens with YaRN.

For more details, including benchmark evaluation, hardware requirements, and inference performance, please refer to our blog, GitHub, and Documentation.

Description
Model synced from source: OpenPipe/Qwen3-14B-Instruct
Readme 15 MiB
Languages
Jinja 100%