Files
RoMistral-7b-Instruct/README.md
ModelHub XC dbbb1a8001 初始化项目,由ModelHub XC社区提供模型
Model: OpenLLM-Ro/RoMistral-7b-Instruct
Source: Original Platform
2026-07-26 03:45:11 +08:00

26 KiB

license, language, base_model, datasets, model-index
license language base_model datasets model-index
cc-by-nc-4.0
ro
mistralai/Mistral-7B-v0.3
OpenLLM-Ro/ro_sft_alpaca
OpenLLM-Ro/ro_sft_alpaca_gpt4
OpenLLM-Ro/ro_sft_dolly
OpenLLM-Ro/ro_sft_selfinstruct_gpt4
OpenLLM-Ro/ro_sft_norobots
OpenLLM-Ro/ro_sft_orca
OpenLLM-Ro/ro_sft_camel
OpenLLM-Ro/ro_sft_oasst
OpenLLM-Ro/ro_sft_ultrachat
OpenLLM-Ro/ro_sft_magpie_mt
OpenLLM-Ro/ro_sft_magpie_reasoning
name results
OpenLLM-Ro/RoMistral-7b-Instruct-2025-04-23
task dataset metrics
type
text-generation
name type
RoMT-Bench RoMT-Bench
name type value
Score Score 6.24
task dataset metrics
type
text-generation
name type
RoCulturaBench RoCulturaBench
name type value
Score Score 4.36
task dataset metrics
type
text-generation
name type
Romanian_Academic_Benchmarks Romanian_Academic_Benchmarks
name type value
Average accuracy accuracy 54.40
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_arc_challenge OpenLLM-Ro/ro_arc_challenge
name type value
Average accuracy accuracy 52.86
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_mmlu OpenLLM-Ro/ro_mmlu
name type value
Average accuracy accuracy 52.33
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_winogrande OpenLLM-Ro/ro_winogrande
name type value
Average accuracy accuracy 68.57
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_hellaswag OpenLLM-Ro/ro_hellaswag
name type value
Average accuracy accuracy 63.50
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_gsm8k OpenLLM-Ro/ro_gsm8k
name type value
Average accuracy accuracy 38.15
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_truthfulqa OpenLLM-Ro/ro_truthfulqa
name type value
Average accuracy accuracy 51.01
task dataset metrics
type
text-generation
name type
LaRoSeDa_binary LaRoSeDa_binary
name type value
Average macro-f1 macro-f1 97.67
task dataset metrics
type
text-generation
name type
LaRoSeDa_multiclass LaRoSeDa_multiclass
name type value
Average macro-f1 macro-f1 61.79
task dataset metrics
type
text-generation
name type
WMT_EN-RO WMT_EN-RO
name type value
Average bleu bleu 28.69
task dataset metrics
type
text-generation
name type
WMT_RO-EN WMT_RO-EN
name type value
Average bleu bleu 19.23
task dataset metrics
type
text-generation
name type
XQuAD XQuAD
name type value
Average exact_match exact_match 49.05
task dataset metrics
type
text-generation
name type
XQuAD XQuAD
name type value
Average f1 f1 69.11
task dataset metrics
type
text-generation
name type
STS STS
name type value
Average spearman spearman 78.67
task dataset metrics
type
text-generation
name type
STS STS
name type value
Average pearson pearson 77.08
task dataset metrics
type
text-generation
name type
RoMT-Bench RoMT-Bench
name type value
First turn Score 6.78
name type value
Second turn Score 5.70
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_arc_challenge OpenLLM-Ro/ro_arc_challenge
name type value
0-shot accuracy 50.04
name type value
1-shot accuracy 50.99
name type value
3-shot accuracy 53.30
name type value
5-shot accuracy 53.73
name type value
10-shot accuracy 54.07
name type value
25-shot accuracy 55.01
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_mmlu OpenLLM-Ro/ro_mmlu
name type value
0-shot accuracy 51.04
name type value
1-shot accuracy 52.53
name type value
3-shot accuracy 53.22
name type value
5-shot accuracy 52.52
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_winogrande OpenLLM-Ro/ro_winogrande
name type value
0-shot accuracy 66.38
name type value
1-shot accuracy 68.90
name type value
3-shot accuracy 68.82
name type value
5-shot accuracy 70.17
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_hellaswag OpenLLM-Ro/ro_hellaswag
name type value
0-shot accuracy 62.61
name type value
1-shot accuracy 63.19
name type value
3-shot accuracy 63.46
name type value
5-shot accuracy 63.92
name type value
10-shot accuracy 64.34
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_gsm8k OpenLLM-Ro/ro_gsm8k
name type value
1-shot accuracy 27.98
name type value
3-shot accuracy 40.46
name type value
5-shot accuracy 46.02
task dataset metrics
type
text-generation
name type
LaRoSeDa_binary LaRoSeDa_binary
name type value
0-shot macro-f1 97.87
name type value
1-shot macro-f1 96.73
name type value
3-shot macro-f1 98.20
name type value
5-shot macro-f1 97.87
task dataset metrics
type
text-generation
name type
LaRoSeDa_multiclass LaRoSeDa_multiclass
name type value
0-shot macro-f1 45.15
name type value
1-shot macro-f1 65.77
name type value
3-shot macro-f1 66.57
name type value
5-shot macro-f1 69.66
task dataset metrics
type
text-generation
name type
WMT_EN-RO WMT_EN-RO
name type value
0-shot bleu 28.92
name type value
1-shot bleu 28.42
name type value
3-shot bleu 28.85
name type value
5-shot bleu 28.58
task dataset metrics
type
text-generation
name type
WMT_RO-EN WMT_RO-EN
name type value
0-shot bleu 3.56
name type value
1-shot bleu 9.60
name type value
3-shot bleu 29.53
name type value
5-shot bleu 34.25
task dataset metrics
type
text-generation
name type
XQuAD_EM XQuAD_EM
name type value
0-shot exact_match 45.21
name type value
1-shot exact_match 49.83
name type value
3-shot exact_match 50.34
name type value
5-shot exact_match 50.84
task dataset metrics
type
text-generation
name type
XQuAD_F1 XQuAD_F1
name type value
0-shot f1 66.40
name type value
1-shot f1 68.92
name type value
3-shot f1 70.68
name type value
5-shot f1 70.44
task dataset metrics
type
text-generation
name type
STS_Spearman STS_Spearman
name type value
1-shot spearman 79.08
name type value
3-shot spearman 78.65
name type value
5-shot spearman 78.29
task dataset metrics
type
text-generation
name type
STS_Pearson STS_Pearson
name type value
1-shot pearson 77.79
name type value
3-shot pearson 76.89
name type value
5-shot pearson 76.57

Model Card for Model ID

This model points/is identical to RoMistral-7b-Instruct-2025-04-23.

RoMistral is a family of pretrained and fine-tuned generative text models for Romanian. This is the repository for the instruct 7B model. Links to other models can be found at the bottom of this page.

Model Details

Model Description

OpenLLM-Ro represents the first open-source effort to build a LLM specialized for Romanian. OpenLLM-Ro developed and publicly releases a collection of Romanian LLMs, both in the form of foundational model and instruct and chat variants.

  • Developed by: OpenLLM-Ro

Model Sources

Intended Use

Intended Use Cases

RoMistral is intented for research use in Romanian. Base models can be adapted for a variety of natural language tasks while instruction and chat tuned models are intended for assistant-like chat.

Out-of-Scope Use

Use in any manner that violates the license, any applicable laws or regluations, use in languages other than Romanian.

How to Get Started with the Model

Use the code below to get started with the model.

from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("OpenLLM-Ro/RoMistral-7b-Instruct")
model = AutoModelForCausalLM.from_pretrained("OpenLLM-Ro/RoMistral-7b-Instruct")

instruction = "Ce jocuri de societate pot juca cu prietenii mei?"
chat = [
        {"role": "user", "content": instruction},
        ]
prompt = tokenizer.apply_chat_template(chat, tokenize=False, system_message="")

inputs = tokenizer.encode(prompt, add_special_tokens=False, return_tensors="pt")
outputs = model.generate(input_ids=inputs, max_new_tokens=128)
print(tokenizer.decode(outputs[0]))

Academic Benchmarks

Model Average ARC MMLU Winogrande Hellaswag GSM8k TruthfulQA
Mistral-7B-Instruct-v0.247.4046.2947.0058.7854.2713.4764.59
RoMistral-7b-Instruct-2024-05-1752.5450.4151.6166.4860.2734.1952.30
RoMistral-7b-Instruct-2024-10-0952.9152.2749.3370.0362.8832.4250.51
RoMistral-7b-Instruct-2025-04-2354.4052.8652.3368.5763.5038.1551.01
RoMistral-7b-Instruct-DPO-2024-10-0951.9550.7347.8868.4162.2732.2750.12
RoMistral-7b-Instruct-DPO-2025-04-2356.6255.5152.6168.0464.9741.0757.55

Downstream tasks

LaRoSeDa WMT
Few-shot Finetuned Few-shot Finetuned
Model Binary
(Macro F1)
Multiclass
(Macro F1)
Binary
(Macro F1)
Multiclass
(Macro F1)
EN-RO
(Bleu)
RO-EN
(Bleu)
EN-RO
(Bleu)
RO-EN
(Bleu)
Mistral-7B-Instruct-v0.296.9756.6698.8387.3218.6033.9926.1939.88
RoMistral-7b-Instruct-2024-05-1797.3667.5598.8088.2827.9313.2128.7240.86
RoMistral-7b-Instruct-2024-10-0995.5667.8399.0087.5728.286.1027.7040.36
RoMistral-7b-Instruct-2025-04-2397.6761.79--28.6919.23--
RoMistral-7b-Instruct-DPO-2024-10-0982.1365.24--26.256.09--
RoMistral-7b-Instruct-DPO-2025-04-2397.9466.13--27.2418.41--
XQuAD STS
Few-shot Finetuned Few-shot Finetuned
Model (EM) (F1) (EM) (F1) (Spearman) (Pearson) (Spearman) (Pearson)
Mistral-7B-Instruct-v0.227.9250.7165.4679.7362.6260.8684.9285.44
RoMistral-7b-Instruct-2024-05-1743.6663.7055.0472.3177.4378.4387.2587.79
RoMistral-7b-Instruct-2024-10-0941.0963.2147.5662.6978.4777.2487.2887.88
RoMistral-7b-Instruct-2025-04-2349.0569.11--78.6777.08--
RoMistral-7b-Instruct-DPO-2024-10-0923.4045.80--77.3376.60--
RoMistral-7b-Instruct-DPO-2025-04-2340.8662.24--77.8976.40--

MT-Bench

Model Average 1st turn 2nd turn Answers in Ro
Mistral-7B-Instruct-v0.25.035.055.00154/160
RoMistral-7b-Instruct-2024-05-174.995.464.53160/160
RoMistral-7b-Instruct-2024-10-095.295.864.72160/160
RoMistral-7b-Instruct-2025-04-236.246.785.70160/160
RoMistral-7b-Instruct-DPO-2024-10-095.886.445.33160/160
RoMistral-7b-Instruct-DPO-2025-04-236.616.866.35160/160

RoCulturaBench

Model Average Answers in Ro
Mistral-7B-Instruct-v0.23.6897/100
RoMistral-7b-Instruct-2024-05-173.38100/100
RoMistral-7b-Instruct-2024-10-093.99100/100
RoMistral-7b-Instruct-2025-04-234.36100/100
RoMistral-7b-Instruct-DPO-2024-10-094.72100/100
RoMistral-7b-Instruct-DPO-2025-04-234.93100/100

RoMistral Model Family

Model Link
RoMistral-7b-Instruct-2024-05-17 link
RoMistral-7b-Instruct-2024-10-09 link
RoMistral-7b-Instruct-2025-04-23 link
RoMistral-7b-Instruct-DPO-2024-10-09 link
RoMistral-7b-Instruct-DPO-2025-04-23 link

Citation

@inproceedings{masala-etal-2024-vorbesti,
    title = "``Vorbe\c{s}ti Rom{\^a}ne\c{s}te?'' A Recipe to Train Powerful {R}omanian {LLM}s with {E}nglish Instructions",
    author = "Masala, Mihai and Ilie-Ablachim, Denis and Dima, Alexandru and Corlatescu, Dragos Georgian and Zavelca, Miruna-Andreea and Olaru, Ovio and Terian, Simina-Maria and Terian, Andrei and Leordeanu, Marius and Velicu, Horia and Popescu, Marius and Dascalu, Mihai and Rebedea, Traian",
    editor = "Al-Onaizan, Yaser and Bansal, Mohit and Chen, Yun-Nung",
    booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2024",
    month = nov,
    year = "2024",
    address = "Miami, Florida, USA",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2024.findings-emnlp.681/",
    doi = "10.18653/v1/2024.findings-emnlp.681",
    pages = "11632--11647"
}