Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Google Research introduces TurboQuant, a compression algorithm for large language models and vector search engines, achieving massive compression with zero accuracy loss, along with the QJL and PolarQuant methods.
From the source
Today, we introduce TurboQuant (to be presented at ICLR 2026), a compression algorithm that optimally addresses the challenge of memory overhead in vector quantization.
research.google