# Hugging Face — 🚀 Accelerating LLM Inference with TGI on Intel Gaudi

- Company: Hugging Face (huggingface.co)
- Announced: 2025-03-28T00:00:00+00:00
- Category: infrastructure-release
- Subject: Platform
- Source: https://huggingface.co/blog/intel-gaudi-backend-for-tgi
- Record: https://forck.live/items/1723-accelerating-llm-inference-with-tgi-on-intel-gaudi

Hugging Face announces native integration of Intel Gaudi hardware support into Text Generation Inference (TGI), enabling deployment of LLMs on Gaudi accelerators with features like multi-card inference, vision-language models, and FP8 precision, and supporting models such as Llama 3.1, Mixtral, Mistral, and more.

## Evidence

Verbatim from https://huggingface.co/blog/intel-gaudi-backend-for-tgi:

> We're excited to announce the native integration of Intel Gaudi hardware support directly into Text Generation Inference (TGI), our production-ready serving solution for Large Language Models (LLMs).

---

Record: https://forck.live/items/1723-accelerating-llm-inference-with-tgi-on-intel-gaudi
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
