library_name, license, license_name, license_link, base_model, pipeline_tag, tags, language
library_name
license
license_name
license_link
base_model
pipeline_tag
tags
language
transformers
other
llama3
LICENSE
NousResearch/Meta-Llama-3-8B
text-generation
llama-3
text-generation
quantization
post-training-quantization
warpquant
hadamard-transform
output-fisher
pytorch
llm
en
WarpQuant Llama 3 8B R16E4H4
Llama 3 8B quantized with signed Hadamard rotation, block-GPTQ, and Output-Fisher weak-column recovery. Projection weights use a 3.5-bpw INT3 base, selected columns are restored in BF16, and the embedding and output head use group-128 INT4.
The repository stores the quantized values in BF16-compatible safetensors. The reported payload is the packed-equivalent analytical size including codes, scales, recovery values, and column indices.