library_name, tags, base_model, model-index
library_name tags base_model model-index
transformers
mergekit
merge
prithivMLmods/QwQ-LCoT-3B-Instruct
bunnycore/Qwen2.5-3B-RP-Mix
name results
QwQen-3B-LCoT
task dataset metrics source
type name
text-generation Text Generation
name type args
IFEval (0-Shot) HuggingFaceH4/ifeval
num_few_shot
0
type value name
inst_level_strict_acc and prompt_level_strict_acc 60.25 strict accuracy
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard
task dataset metrics source
type name
text-generation Text Generation
name type args
BBH (3-Shot) BBH
num_few_shot
3
type value name
acc_norm 28.5 normalized accuracy
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard
task dataset metrics source
type name
text-generation Text Generation
name type args
MATH Lvl 5 (4-Shot) hendrycks/competition_math
num_few_shot
4
type value name
exact_match 0.91 exact match
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard
task dataset metrics source
type name
text-generation Text Generation
name type args
GPQA (0-shot) Idavidrein/gpqa
num_few_shot
0
type value name
acc_norm 2.24 acc_norm
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard
task dataset metrics source
type name
text-generation Text Generation
name type args
MuSR (0-shot) TAUR-Lab/MuSR
num_few_shot
0
type value name
acc_norm 10.76 acc_norm
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard
task dataset metrics source
type name
text-generation Text Generation
name type config split args
MMLU-PRO (5-shot) TIGER-Lab/MMLU-Pro main test
num_few_shot
5
type value name
acc 29.99 accuracy
url name
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=bunnycore/QwQen-3B-LCoT Open LLM Leaderboard

merge

This is a merge of pre-trained language models created using mergekit.

Merge Details

Merge Method

This model was merged using the linear merge method using bunnycore/Qwen2.5-3B-RP-Mix as a base.

Models Merged

The following models were included in the merge:

Configuration

The following YAML configuration was used to produce this model:

merge_method: linear
dtype: bfloat16
normalize: true
base_model: bunnycore/Qwen2.5-3B-RP-Mix
models:
  - model: bunnycore/Qwen2.5-3B-RP-Mix
    parameters:
      weight: 10
      density: 1
  - model: prithivMLmods/QwQ-LCoT-3B-Instruct
    parameters:
      weight: 7
      density: 0.8

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 22.11
IFEval (0-Shot) 60.25
BBH (3-Shot) 28.50
MATH Lvl 5 (4-Shot) 0.91
GPQA (0-shot) 2.24
MuSR (0-shot) 10.76
MMLU-PRO (5-shot) 29.99
Description
Model synced from source: bunnycore/QwQen-3B-LCoT
Readme 2 MiB