Files
ModelHub XC 24ecb38c2f 初始化项目,由ModelHub XC社区提供模型
Model: matthewagi/essay-author-rewrite-smollm3-3b-pg-jasper-drift080-step100
Source: Original Platform
2026-10-06 10:39:16 +08:00

735 B

base_model, library_name, tags
base_model library_name tags
HuggingFaceTB/SmolLM3-3B transformers
prime-rl
essay-rewrite
paul-graham
jasper
lora

Essay Author Rewrite - SmolLM3 3B - Step 100

Prime-RL essay author rewrite checkpoint trained for 100 steps with Jasper PG author reward and hard embedding drift gate at cosine >= 0.8.

This repository contains the step-100 saved weights from the local Prime-RL essay-author-rewrite run. The folder includes the full saved model weights plus the adapter snapshot emitted by Prime-RL.

Training setup:

  • Dataset:
  • Target author: Paul Graham
  • Author/semantic embedding model:
  • N-gram discriminator:
  • Source truncation: first 1024 model-tokenized source tokens
  • Max completion tokens: 1400