Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Cursor and Together AI partner to deliver real-time, low-latency inference for Cursor's AI coding platform using NVIDIA Blackwell GPUs, custom kernels, and a quantization pipeline to enable fast model iteration.
From the source
Cursor partnered with Together AI, the AI Native Cloud, to build infrastructure for this loop — using the NVIDIA Blackwell architecture and tuning the inference stack to meet strict latency targets.
together.ai