license, language, pipeline_tag, library_name, datasets
license language pipeline_tag library_name datasets
mit
en
text-generation transformers
croqaz/vintage-v1

Vintage-LLM

This is a 340M params Llama3-based model, trained on text pre-1900.

This is a base model. It has very limited chat capabilities.

Dataset used: https://huggingface.co/datasets/croqaz/vintage-v1

The data was de-duplicated and only the decent quality texts were used.

There are a few checkpoints in the repo:

Description
Model synced from source: croqaz/vintage-LLM-340m-v1-base
Readme 586 KiB
Languages
Jinja 100%