# NVIDIA — NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide

- Company: NVIDIA (nvidia.com)
- Announced: 2026-07-21T15:36:43+00:00
- Category: new-model
- Subject: AI platform
- Models affected: Vera Rubin, Vera Rubin NVL72, Vera CPU, Groq 3 LPX, Spectrum-6 SPX, Vera BlueField-4 STX, Mistral Medium 3.5, OCR 4
- Pricing: one-tenth the cost per million tokens
- Source: https://blogs.nvidia.com/blog/vera-rubin/
- Record: https://forck.live/items/1438-nvidia-vera-rubin-driving-performance-per-watt-lowest-token-cost-for-partners

NVIDIA announces the Vera Rubin platform, a chip-to-grid AI supercomputer featuring the Vera CPU and six other chips, claiming 10x more tokens per megawatt and one-tenth the cost per million tokens compared to GB200 NVL72, with partners like CoreWeave, Google Cloud, Microsoft Azure, Oracle Cloud Infrastructure, and Nebius ramping production. The post also mentions Mistral Medium 3.5 and OCR 4 models available in Microsoft Foundry, and a new partnership between Microsoft and Mistral for European AI infrastructure using thousands of Vera Rubin GPUs.

## Evidence

Verbatim from https://blogs.nvidia.com/blog/vera-rubin/:

> The Vera Rubin platform is built from chip to grid to deliver the highest performance per watt and the lowest token cost.

---

Record: https://forck.live/items/1438-nvidia-vera-rubin-driving-performance-per-watt-lowest-token-cost-for-partners
Catalogue: https://forck.live/llms.txt
Feed: https://forck.live/feed.md
