Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face fine-tuned Llama3 8B models to 1.58-bit quantization using the BitNet architecture, released the models under the HF1BitLLM organization, and introduced a new quantization method called 'bitnet' in Transformers.
From the source
We have successfully fine-tuned a Llama3 8B model using the BitNet architecture, achieving strong performance on downstream tasks. The 8B models we developed are released under the HF1BitLLM organization.
huggingface.co