初始化项目,由ModelHub XC社区提供模型

Model: JaydeepR/SmolLM-135M-neuraltxt-dpo-v1
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-08-02 05:49:18 +08:00
commit 62b47dc323
61 changed files with 1487206 additions and 0 deletions

2884
logs/dpo_20260601_0917.log Normal file

File diff suppressed because one or more lines are too long

View File

@@ -0,0 +1,18 @@
==========================================
method=dpo run_id=dpo_default
base=paperbd/smollm_135M_neuraltxt_v1
dataset=paperbd/paper_preference_150K-v1
batch=16 grad_accum=8 (eff ~128)
==========================================
== uv sync ==
Resolved 148 packages in 5ms
Checked 130 packages in 414ms
== Step 0: baseline diversity (SFT model) ==
Warning: You are sending unauthenticated requests to the HF Hub. Please set a HF_TOKEN to enable higher rate limits and faster downloads.
Generating train split: 0%| | 0/250000 [00:00<?, ? examples/s]
Generating train split: 2%|▏ | 6006/250000 [00:00<00:04, 50480.11 examples/s]
Generating train split: 7%|▋ | 18026/250000 [00:00<00:02, 81013.14 examples/s]
Generating train split: 12%|█▏ | 30072/250000 [00:00<00:02, 94133.40 examples/s]
Generating train split: 17%|█▋ | 42104/250000 [00:00<00:02, 100499.83 examples/s]
Generating train split: 22%|██▏ | 54154/250000 [00:00<00:01, 105437.55 examples/s]

File diff suppressed because one or more lines are too long