license, language, base_model, datasets, model-index
license language base_model datasets model-index
cc-by-nc-4.0
ro
OpenLLM-Ro/RoLlama2-7b-Base
OpenLLM-Ro/ro_sft_alpaca
OpenLLM-Ro/ro_sft_alpaca_gpt4
OpenLLM-Ro/ro_sft_dolly
OpenLLM-Ro/ro_sft_selfinstruct_gpt4
OpenLLM-Ro/ro_sft_norobots
OpenLLM-Ro/ro_sft_orca
OpenLLM-Ro/ro_sft_camel
OpenLLM-Ro/ro_sft_oasst
OpenLLM-Ro/ro_sft_ultrachat
OpenLLM-Ro/ro_sft_magpie_mt
OpenLLM-Ro/ro_sft_magpie_reasoning
name results
OpenLLM-Ro/RoLlama2-7b-Instruct-2025-04-23
task dataset metrics
type
text-generation
name type
RoMT-Bench RoMT-Bench
name type value
Score Score 4.97
task dataset metrics
type
text-generation
name type
RoCulturaBench RoCulturaBench
name type value
Score Score 4.56
task dataset metrics
type
text-generation
name type
Romanian_Academic_Benchmarks Romanian_Academic_Benchmarks
name type value
Average accuracy accuracy 45.51
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_arc_challenge OpenLLM-Ro/ro_arc_challenge
name type value
Average accuracy accuracy 45.7
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_mmlu OpenLLM-Ro/ro_mmlu
name type value
Average accuracy accuracy 40.36
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_winogrande OpenLLM-Ro/ro_winogrande
name type value
Average accuracy accuracy 63.26
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_hellaswag OpenLLM-Ro/ro_hellaswag
name type value
Average accuracy accuracy 60.25
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_gsm8k OpenLLM-Ro/ro_gsm8k
name type value
Average accuracy accuracy 18.02
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_truthfulqa OpenLLM-Ro/ro_truthfulqa
name type value
Average accuracy accuracy 45.48
task dataset metrics
type
text-generation
name type
LaRoSeDa_binary LaRoSeDa_binary
name type value
Average macro-f1 macro-f1 97.6
task dataset metrics
type
text-generation
name type
LaRoSeDa_multiclass LaRoSeDa_multiclass
name type value
Average macro-f1 macro-f1 60.22
task dataset metrics
type
text-generation
name type
WMT_EN-RO WMT_EN-RO
name type value
Average bleu bleu 27.21
task dataset metrics
type
text-generation
name type
WMT_RO-EN WMT_RO-EN
name type value
Average bleu bleu 22.15
task dataset metrics
type
text-generation
name type
XQuAD XQuAD
name type value
Average exact_match exact_match 47.39
task dataset metrics
type
text-generation
name type
XQuAD XQuAD
name type value
Average f1 f1 65.77
task dataset metrics
type
text-generation
name type
STS STS
name type value
Average spearman spearman 59.05
task dataset metrics
type
text-generation
name type
STS STS
name type value
Average pearson pearson 56.45
task dataset metrics
type
text-generation
name type
RoMT-Bench RoMT-Bench
name type value
First turn Score 5.56
name type value
Second turn Score 4.39
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_arc_challenge OpenLLM-Ro/ro_arc_challenge
name type value
0-shot accuracy 43.02
name type value
1-shot accuracy 45.84
name type value
3-shot accuracy 45.24
name type value
5-shot accuracy 46.19
name type value
10-shot accuracy 46.7
name type value
25-shot accuracy 47.22
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_mmlu OpenLLM-Ro/ro_mmlu
name type value
0-shot accuracy 38.64
name type value
1-shot accuracy 40.77
name type value
3-shot accuracy 41.19
name type value
5-shot accuracy 40.86
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_winogrande OpenLLM-Ro/ro_winogrande
name type value
0-shot accuracy 63.61
name type value
1-shot accuracy 62.75
name type value
3-shot accuracy 63.46
name type value
5-shot accuracy 63.22
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_hellaswag OpenLLM-Ro/ro_hellaswag
name type value
0-shot accuracy 59.79
name type value
1-shot accuracy 59.62
name type value
3-shot accuracy 60.12
name type value
5-shot accuracy 60.71
name type value
10-shot accuracy 61.01
task dataset metrics
type
text-generation
name type
OpenLLM-Ro/ro_gsm8k OpenLLM-Ro/ro_gsm8k
name type value
1-shot accuracy 6.14
name type value
3-shot accuracy 22.52
name type value
5-shot accuracy 25.4
task dataset metrics
type
text-generation
name type
LaRoSeDa_binary LaRoSeDa_binary
name type value
0-shot macro-f1 98.17
name type value
1-shot macro-f1 96.3
name type value
3-shot macro-f1 97.8
name type value
5-shot macro-f1 98.13
task dataset metrics
type
text-generation
name type
LaRoSeDa_multiclass LaRoSeDa_multiclass
name type value
0-shot macro-f1 49.8
name type value
1-shot macro-f1 56.03
name type value
3-shot macro-f1 65.33
name type value
5-shot macro-f1 69.7
task dataset metrics
type
text-generation
name type
WMT_EN-RO WMT_EN-RO
name type value
0-shot bleu 19.34
name type value
1-shot bleu 29.89
name type value
3-shot bleu 29.99
name type value
5-shot bleu 29.62
task dataset metrics
type
text-generation
name type
WMT_RO-EN WMT_RO-EN
name type value
0-shot bleu 2.29
name type value
1-shot bleu 14.74
name type value
3-shot bleu 34.82
name type value
5-shot bleu 36.75
task dataset metrics
type
text-generation
name type
XQuAD_EM XQuAD_EM
name type value
0-shot exact_match 42.86
name type value
1-shot exact_match 47.82
name type value
3-shot exact_match 48.32
name type value
5-shot exact_match 50.59
task dataset metrics
type
text-generation
name type
XQuAD_F1 XQuAD_F1
name type value
0-shot f1 63.66
name type value
1-shot f1 65.27
name type value
3-shot f1 66.04
name type value
5-shot f1 68.12
task dataset metrics
type
text-generation
name type
STS_Spearman STS_Spearman
name type value
1-shot spearman 54.51
name type value
3-shot spearman 60.98
name type value
5-shot spearman 61.65
task dataset metrics
type
text-generation
name type
STS_Pearson STS_Pearson
name type value
1-shot pearson 54.35
name type value
3-shot pearson 57.88
name type value
5-shot pearson 57.13

Model Card for Model ID

This model points/is identical to RoLlama2-7b-Instruct-2025-04-23.

RoLlama2 is a family of pretrained and fine-tuned generative text models for Romanian. This is the repository for the instruct 7B model. Links to other models can be found at the bottom of this page.

Model Details

Model Description

OpenLLM represents the first open-source effort to build a LLM specialized for Romanian. OpenLLM-Ro developed and publicly releases a collection of Romanian LLMs, both in the form of foundational model and instruct and chat variants.

  • Developed by: OpenLLM-Ro

Model Sources

Intended Use

Intended Use Cases

RoLlama2 is intented for research use in Romanian. Base models can be adapted for a variety of natural language tasks while instruction and chat tuned models are intended for assistant-like chat.

Out-of-Scope Use

Use in any manner that violates the license, any applicable laws or regluations, use in languages other than Romanian.

How to Get Started with the Model

Use the code below to get started with the model.

from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("OpenLLM-Ro/RoLlama2-7b-Instruct")
model = AutoModelForCausalLM.from_pretrained("OpenLLM-Ro/RoLlama2-7b-Instruct")

instruction = "Care este cel mai înalt vârf muntos din România?"
chat = [
        {"role": "system", "content": "Ești un asistent folositor, respectuos și onest. Încearcă să ajuți cât mai mult prin informațiile oferite, excluzând răspunsuri toxice, rasiste, sexiste, periculoase și ilegale."},
        {"role": "user", "content": instruction},
        ]
prompt = tokenizer.apply_chat_template(chat, tokenize=False)

inputs = tokenizer.encode(prompt, add_special_tokens=False, return_tensors="pt")
outputs = model.generate(input_ids=inputs, max_new_tokens=128)
print(tokenizer.decode(outputs[0]))

Academic Benchmarks

Model Average ARC MMLU Winogrande Hellaswag GSM8k TruthfulQA
Llama-2-7b-chat36.8437.0333.8055.8745.364.9044.09
RoLlama2-7b-Instruct-2024-05-1445.7143.6639.7070.3457.3618.7844.44
RoLlama2-7b-Instruct-2024-10-0944.5044.7340.3963.6759.1213.2945.78
RoLlama2-7b-Instruct-2025-04-2345.5145.7040.3663.2660.2518.0245.48
RoLlama2-7b-Instruct-DPO-2024-10-0943.2044.2438.3962.5759.2015.7239.07
RoLlama2-7b-Instruct-DPO-2025-04-2346.7748.1641.3864.1561.3718.3547.20

Downstream tasks

LaRoSeDa WMT
Few-shot Finetuned Few-shot Finetuned
Model Binary
(Macro F1)
Multiclass
(Macro F1)
Binary
(Macro F1)
Multiclass
(Macro F1)
EN-RO
(Bleu)
RO-EN
(Bleu)
EN-RO
(Bleu)
RO-EN
(Bleu)
Llama-2-7b-chat87.7852.8197.2782.0215.5528.5319.9931.48
RoLlama2-7b-Instruct-2024-05-1497.4865.2698.8387.2827.3810.3227.5940.13
RoLlama2-7b-Instruct-2024-10-0997.6662.4197.9760.8927.1319.3927.6339.75
RoLlama2-7b-Instruct-2025-04-2397.6060.22--27.2122.15--
RoLlama2-7b-Instruct-DPO-2024-10-0997.3160.56--26.5621.68--
RoLlama2-7b-Instruct-DPO-2025-04-2397.7765.21--25.4822.75--
XQuAD STS
Few-shot Finetuned Few-shot Finetuned
Model (EM) (F1) (EM) (F1) (Spearman) (Pearson) (Spearman) (Pearson)
Llama-2-7b-chat32.3554.0060.3475.9832.5631.9974.0872.64
RoLlama2-7b-Instruct-2024-05-1444.5264.7554.9670.2065.5067.7984.4484.76
RoLlama2-7b-Instruct-2024-10-0945.7165.0859.2474.2559.6957.1684.6685.07
RoLlama2-7b-Instruct-2025-04-2347.3965.77--59.0556.45--
RoLlama2-7b-Instruct-DPO-2024-10-0935.7859.31--61.2258.41--
RoLlama2-7b-Instruct-DPO-2025-04-2338.2860.88--66.7664.72--

Romanian MT-Bench

Model Average 1st turn 2nd turn Answers in Ro
Llama-2-7b-chat1.081.440.7345/160
RoLlama2-7b-Instruct-2024-05-143.864.673.04160/160
RoLlama2-7b-Instruct-2024-10-094.434.923.94160/160
RoLlama2-7b-Instruct-2025-04-234.975.564.39160/160
RoLlama2-7b-Instruct-DPO-2024-10-094.615.154.06160/160
RoLlama2-7b-Instruct-DPO-2025-04-235.555.845.26160/160

RoCulturaBench

Model Average Answers in Ro
Llama-2-7b-chat1.2133/100
RoLlama2-7b-Instruct-2024-05-143.77100/100
RoLlama2-7b-Instruct-2024-10-094.08100/100
RoLlama2-7b-Instruct-2025-04-234.56100/100
RoLlama2-7b-Instruct-DPO-2024-10-094.80100/100
RoLlama2-7b-Instruct-DPO-2025-04-235.24100/100

RoLlama2 Model Family

Model Link
RoLlama2-7b-Base-2024-05-14 link
RoLlama2-7b-Instruct-2024-05-14 link
RoLlama2-7b-Instruct-2024-10-09 link
RoLlama2-7b-Instruct-2025-04-23 link
RoLlama2-7b-Instruct-DPO-2024-10-09 link
RoLlama2-7b-Instruct-DPO-2025-04-23 link

Citation

@inproceedings{masala-etal-2024-vorbesti,
    title = "``Vorbe\c{s}ti Rom{\^a}ne\c{s}te?'' A Recipe to Train Powerful {R}omanian {LLM}s with {E}nglish Instructions",
    author = "Masala, Mihai and Ilie-Ablachim, Denis and Dima, Alexandru and Corlatescu, Dragos Georgian and Zavelca, Miruna-Andreea and Olaru, Ovio and Terian, Simina-Maria and Terian, Andrei and Leordeanu, Marius and Velicu, Horia and Popescu, Marius and Dascalu, Mihai and Rebedea, Traian",
    editor = "Al-Onaizan, Yaser and Bansal, Mohit and Chen, Yun-Nung",
    booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2024",
    month = nov,
    year = "2024",
    address = "Miami, Florida, USA",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2024.findings-emnlp.681/",
    doi = "10.18653/v1/2024.findings-emnlp.681",
    pages = "11632--11647"
}
Description
Model synced from source: OpenLLM-Ro/RoLlama2-7b-Instruct
Readme 1.1 MiB