Model: longtermrisk/Llama-3.1-8B-school-of-reward-hacks-last-third-sft Source: Original Platform
4 lines
135 B
Plaintext
4 lines
135 B
Plaintext
version https://git-lfs.github.com/spec/v1
|
|
oid sha256:caf1b70fd8216d73d168dbef4622136b2b114c3dea486505c8e7184f3e1d9b6f
|
|
size 1168138808
|