# Hugging Face — A Gentle Introduction to 8-bit Matrix Multiplication for transformers at scale using transformers, accelerate and bitsandbytes

- Company: Hugging Face (huggingface.co)
- Announced: 2022-08-17
- Category: capability-change
- Coverage: not counted
- Announcement: yes
- Group: announcements
- Source: https://huggingface.co/blog/hf-bitsandbytes-integration
- Record: https://forck.live/items/2163-a-gentle-introduction-to-8-bit-matrix-multiplication-for-transformers-at-scale
- Subject: Platform

Hugging Face integrates LLM.int8() 8-bit matrix multiplication into its transformers library to reduce memory footprint of large language models without degrading performance.

## Evidence

Verbatim from https://huggingface.co/blog/hf-bitsandbytes-integration:

> we offer LLM.int8() integration for all Hugging Face models

---

Record: https://forck.live/items/2163-a-gentle-introduction-to-8-bit-matrix-multiplication-for-transformers-at-scale
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
