Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face and NVIDIA launch a serverless inference service called NVIDIA NIM API (serverless) on the Hugging Face Hub, available to Enterprise Hub organizations, providing pay-as-you-go access to open models like Llama and Mistral using NVIDIA DGX Cloud accelerated compute.
From the source
Today, we are thrilled to announce the launch of Hugging Face NVIDIA NIM API (serverless), a new service on the Hugging Face Hub, available to Enterprise Hub organizations.
huggingface.co