初始化项目,由ModelHub XC社区提供模型
Model: JaydeepR/SmolLM-135M-neuraltxt-dpo-v1 Source: Original Platform
This commit is contained in:
2884
logs/dpo_20260601_0917.log
Normal file
2884
logs/dpo_20260601_0917.log
Normal file
File diff suppressed because one or more lines are too long
18
logs/dpo_default_20260601_0805.log
Normal file
18
logs/dpo_default_20260601_0805.log
Normal file
@@ -0,0 +1,18 @@
|
||||
==========================================
|
||||
method=dpo run_id=dpo_default
|
||||
base=paperbd/smollm_135M_neuraltxt_v1
|
||||
dataset=paperbd/paper_preference_150K-v1
|
||||
batch=16 grad_accum=8 (eff ~128)
|
||||
==========================================
|
||||
== uv sync ==
|
||||
Resolved 148 packages in 5ms
|
||||
Checked 130 packages in 414ms
|
||||
== Step 0: baseline diversity (SFT model) ==
|
||||
Warning: You are sending unauthenticated requests to the HF Hub. Please set a HF_TOKEN to enable higher rate limits and faster downloads.
|
||||
|
||||
Generating train split: 0%| | 0/250000 [00:00<?, ? examples/s]
|
||||
Generating train split: 2%|▏ | 6006/250000 [00:00<00:04, 50480.11 examples/s]
|
||||
Generating train split: 7%|▋ | 18026/250000 [00:00<00:02, 81013.14 examples/s]
|
||||
Generating train split: 12%|█▏ | 30072/250000 [00:00<00:02, 94133.40 examples/s]
|
||||
Generating train split: 17%|█▋ | 42104/250000 [00:00<00:02, 100499.83 examples/s]
|
||||
Generating train split: 22%|██▏ | 54154/250000 [00:00<00:01, 105437.55 examples/s]
|
||||
356
logs/dpo_default_20260601_0821.log
Normal file
356
logs/dpo_default_20260601_0821.log
Normal file
File diff suppressed because one or more lines are too long
Reference in New Issue
Block a user