Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face announces that users can now spin up a private, OpenAI-compatible LLM endpoint on HF Jobs with a single command, using vLLM, with pay-per-second billing and no server provisioning.
From the source
You can spin up a private, OpenAI-compatible LLM endpoint on Hugging Face infrastructure with a single command — no servers to provision, no Kubernetes, pay-per-second.
huggingface.co