ModelHub XC 10ff83d2d1 初始化项目,由ModelHub XC社区提供模型
Model: OwenArli/ArliAI-Llama-3-8B-Instruct-ORPO-v0.1
Source: Original Platform
2026-10-07 01:31:28 +08:00

license
license
llama3

Based on Meta-Llama-3-8b-Instruct, and is governed by Meta Llama 3 License agreement: https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

ORPO fine tuning method using the following datasets:

Despite the toxic datasets to reduce refusals, this model is still relatively safe but refuses less than the original Meta model.

As of now ORPO fine tuning seems to improve some metrics while reducing other metrics by a lot:

OpenLLM Leaderboard

Instruct format:

<|begin_of_text|><|start_header_id|>system<|end_header_id|>

{{ system_prompt }}<|eot_id|><|start_header_id|>user<|end_header_id|>

{{ user_message_1 }}<|eot_id|><|start_header_id|>assistant<|end_header_id|>

{{ model_answer_1 }}<|eot_id|><|start_header_id|>user<|end_header_id|>

{{ user_message_2 }}<|eot_id|><|start_header_id|>assistant<|end_header_id|>

Quants:

Description
Model synced from source: OwenArli/ArliAI-Llama-3-8B-Instruct-ORPO-v0.1
Readme 2.8 MiB