# Hugging Face — Accelerate a World of LLMs on Hugging Face with NVIDIA NIM

- Company: Hugging Face (huggingface.co)
- Announced: 2025-07-21T18:01:30+00:00
- Category: capability-change
- Subject: Platform
- Source: https://huggingface.co/blog/nvidia/multi-llm-nim
- Record: https://forck.live/items/1652-accelerate-a-world-of-llms-on-hugging-face-with-nvidia-nim

NVIDIA NIM now supports deploying a broad range of LLMs from Hugging Face using a single Docker container that automatically selects and optimizes an inference backend (TensorRT-LLM, vLLM, or SGLang).

## Evidence

Verbatim from https://huggingface.co/blog/nvidia/multi-llm-nim:

> NIM now provides a single docker container for deploying a broad range of LLMs supported by leading inference frameworks from NVIDIA and the community including NVIDIA TensorRT-LLM, vLLM and SGLang.

---

Record: https://forck.live/items/1652-accelerate-a-world-of-llms-on-hugging-face-with-nvidia-nim
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
