# AMD — AMD and Cerebras Announce Industry-Leading Ultra-Low-Latency and High Throughput AI Inference Solution

- Company: AMD
- Announced: 2026-07-23T17:45:00+00:00
- Category: partnership-acquisition
- Coverage: not counted
- Announcement: yes
- Group: announcements
- Source: https://ir.amd.com/news-events/press-releases/detail/1293/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution
- Record: https://forck.live/items/15036-amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high
- Subject: AMD Instinct / ROCm

AMD and Cerebras announced a technical partnership to deliver a disaggregated AI inference solution combining AMD Helios rackscale systems with the Cerebras Wafer-Scale Engine. The joint solution is expected to deliver up to 5x higher tokens per second per watt, with availability initially through Cerebras Cloud in the second half of 2026.

## Evidence

Verbatim from https://ir.amd.com/news-events/press-releases/detail/1293/amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high-throughput-ai-inference-solution:

> Together, the two compute engines are expected to deliver up to 5x higher tokens per second per watt (T/s/W) i .

---

Record: https://forck.live/items/15036-amd-and-cerebras-announce-industry-leading-ultra-low-latency-and-high
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
