Lead story
Ask your AI
Top stories
Models & availability
Latest
Lead story
Ask your AI
Top stories
Models & availability
Latest
NVIDIA and Microsoft announced partnerships to accelerate local AI inference on NVIDIA hardware at IFA 2026, including new optimizations for llama.cpp and vLLM, and simplified local model setup in Hermes Agent, OpenClaw, and Perplexity Portable Computer. NVIDIA also introduced the Personal AI Router tool (NVIDIA PAIR) and the RTX Spark compact Windows PC line coming in October from Lenovo and Acer.
From the source
Three of the most widely used agent apps will offer simplified local model setup on Windows, each built on llama.cpp and incorporating NVIDIA’s latest inference optimizations.
blogs.nvidia.com