language, license, tags, model-index
language
license
tags
model-index
mit
name
results
FusionNet_7Bx2_MoE_14B
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
config
split
args
AI2 Reasoning Challenge (25-Shot)
ai2_arc
ARC-Challenge
test
type
value
name
acc_norm
73.55
normalized accuracy
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
split
args
HellaSwag (10-Shot)
hellaswag
validation
type
value
name
acc_norm
88.84
normalized accuracy
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
config
split
args
MMLU (5-Shot)
cais/mmlu
all
test
type
value
name
acc
64.68
accuracy
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
config
split
args
TruthfulQA (0-shot)
truthful_qa
multiple_choice
validation
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
config
split
args
Winogrande (5-shot)
winogrande
winogrande_xl
validation
type
value
name
acc
88.16
accuracy
task
dataset
metrics
source
type
name
text-generation
Text Generation
name
type
config
split
args
GSM8k (5-shot)
gsm8k
main
test
type
value
name
acc
70.66
accuracy
FusionNet
Fine-tuned model on English language using MoE method.
Model description
The FusionNet is a model to experiment with the MoE method, which could significantly increase the performance of the original model. The FusionNet has 12.9B parameters, and this model is fine-tuned. Enjoy!
Detailed results can be found here
Metric
Value
Avg.
75.91
AI2 Reasoning Challenge (25-Shot)
73.55
HellaSwag (10-Shot)
88.84
MMLU (5-Shot)
64.68
TruthfulQA (0-shot)
69.60
Winogrande (5-shot)
88.16
GSM8k (5-shot)
70.66