From the source
Lead story
Ask your AI
Top stories
Models & availability
Latest
Lead story
Top stories
Models & availability
Latest
From the source
AMD and Cerebras partner on disaggregated AI inference combining Helios and Wafer-Scale Engine.
AMD and Cerebras announced a technical partnership to deliver a disaggregated AI inference solution combining AMD Helios rackscale systems with the Cerebras Wafer-Scale Engine.
The joint solution is expected to deliver up to 5x higher tokens per second per watt, with availability initially through Cerebras Cloud in the second half of 2026.
From the source
Together, the two compute engines are expected to deliver up to 5x higher tokens per second per watt (T/s/W) i .
ir.amd.com