初始化项目,由ModelHub XC社区提供模型
Model: matthewagi/essay-author-rewrite-qwen3-0.6b-pg-jasper-step100 Source: Original Platform
This commit is contained in:
24
README.md
Normal file
24
README.md
Normal file
@@ -0,0 +1,24 @@
|
||||
---
|
||||
base_model: Qwen/Qwen3-0.6B
|
||||
library_name: transformers
|
||||
tags:
|
||||
- prime-rl
|
||||
- essay-rewrite
|
||||
- paul-graham
|
||||
- jasper
|
||||
- lora
|
||||
---
|
||||
|
||||
# Essay Author Rewrite - Qwen3 0.6B - Step 100
|
||||
|
||||
Prime-RL essay author rewrite checkpoint trained for 100 steps with Jasper PG author reward and original drift gate.
|
||||
|
||||
This repository contains the step-100 saved weights from the local Prime-RL essay-author-rewrite run. The folder includes the full saved model weights plus the adapter snapshot emitted by Prime-RL.
|
||||
|
||||
Training setup:
|
||||
- Dataset:
|
||||
- Target author: Paul Graham
|
||||
- Author/semantic embedding model:
|
||||
- N-gram discriminator:
|
||||
- Source truncation: first 1024 model-tokenized source tokens
|
||||
- Max completion tokens: 1400
|
||||
Reference in New Issue
Block a user