Model synced from source: Mrigank005/SLM-sentiment-crosslingual-seed-42
Updated 2026-04-29 20:28:46 +08:00
Model synced from source: N-Bot-Int/ElaNore3-4B_ADJUSTED_merged
Updated 2026-04-29 20:28:10 +08:00
Model synced from source: sstoica12/influence_metamath_qwen2.5-3b_repeat_regularized_1k_scaled_e1
Updated 2026-04-29 19:49:10 +08:00
Model synced from source: Lyraix-AI/LyraixGuard-v0
Updated 2026-04-29 19:45:21 +08:00
Model synced from source: xw1234gan/Fixed_Merging_Qwen2.5-3B-Instruct_MedQA_lr1e-05_mb2_ga128_n2048_seed42
Updated 2026-04-29 19:39:23 +08:00
Model synced from source: Mohith202/brainrl-grpo-single-m
Updated 2026-04-29 19:17:22 +08:00
Model synced from source: sebsigma/SemanticCite-Refiner-Qwen3-1B
Updated 2026-04-29 19:16:20 +08:00
Model synced from source: ccui46/cookingworld_per_chunk_act_q3_tokfix_diffPrompt_lowerLR_tformerPin_8000
Updated 2026-04-29 18:37:10 +08:00
Model synced from source: LorenaYannnnn/general_reward-Qwen3-0.6B_7168-OURS_self-seed_0
Updated 2026-04-29 18:30:16 +08:00
Model synced from source: LorenaYannnnn/general_reward-Qwen3-0.6B_7168-baseline_all_tokens-seed_0
Updated 2026-04-29 18:30:11 +08:00
Model synced from source: Hyeongwon/P2-split2_prob_Qwen3-14B-Base_0405_1e-5
Updated 2026-04-29 18:23:52 +08:00
Model synced from source: tikeape/Llama-3.2-3B-Hunter-Alpha-Distill
Updated 2026-04-29 18:23:49 +08:00
Model synced from source: DCAgent/d1_constrain_top4_seq_glm47
Updated 2026-04-29 18:23:12 +08:00
Model synced from source: JoaoReiz/Llama3.2_1B_HAREM
Updated 2026-04-29 18:17:39 +08:00
Model synced from source: gjyotin305/Qwen2.5-3B-Instruct_adaptive_tune_no_ref
Updated 2026-04-29 18:11:12 +08:00
Model synced from source: IAAR-Shanghai/MemReader-4B-thinking
Updated 2026-04-29 18:11:11 +08:00
Model synced from source: JamesGern/lorel.ai_cherrypicked
Updated 2026-04-29 18:11:11 +08:00
Model synced from source: laion/swesmith-316__Qwen3-8B
Updated 2026-04-29 18:04:50 +08:00
Model synced from source: JamesGern/lorel.ai_long_train
Updated 2026-04-29 18:04:47 +08:00