Key specifications (per single Helios rack) — from AMD's official product page and brochure :
| Component | Spec |
|---|---|
| GPUs | 72x AMD Instinct MI455X (CDNA 5, TSMC N2) |
| CPUs | AMD EPYC "Venice" (Zen 6, 6th Gen) |
| Pooled memory | 31 TB HBM4 |
| Aggregate memory bandwidth | 1.4 PB/s |
| Per-GPU memory | 432 GB HBM4, 19.6 TB/s bandwidth |
| Per-GPU compute | 40 FP4 PFLOPS / 20 FP8 PFLOPS |
| Rack-level compute | 2.9 FP4 exaFLOPS (inference), 1.4 FP8 exaFLOPS (training) |
| Networking | AMD Pensando "Vulcano" with UALink |
| Form factor | 18 compute trays (4 GPUs + 1 EPYC Venice per tray), double-wide rack |
| Estimated price | ~$5.25M per rack |
The MI455X delivers roughly 2x the FP4 compute of the previous MI350 series . The Helios system is AMD's primary rack-scale competitor to Nvidia's Vera Rubin platform and has already lined up customers including Microsoft, Meta, OpenAI, Oracle, and Anthropic .
Announced on stage at Advancing AI 2026 (July 23), AMD and Cerebras are building a disaggregated inference platform that splits AI inference into two phases :
The two systems operate as a single disaggregated solution via high-speed interconnect .
Claimed performance: Up to 5x higher tokens per second per watt compared to conventional architectures (based on July 2026 modeling using the Kimi 2.6 1T model) .
Availability: The joint AMD-Cerebras AI inference solution is expected to be available through cloud providers in H2 2026 . Cerebras will deploy Helios systems in its own data centers, with the solution available first through Cerebras Cloud .
Cerebras CEO Andrew Feldman said at the event that the architecture directly targets Nvidia's Groq LPUs and is positioned for agentic AI workloads where low-latency inference is critical .
Announced July 22, 2026, via AMD official press release :
Deal terms:
This adds Anthropic to AMD's growing roster of major AI customers alongside OpenAI, Meta, Oracle, and Microsoft . The MI450 Series mentioned in this deal is the generation of GPU that powers the Helios racks — the MI455X is the specific SKU used in the first Helios production units .
| Product / Program | Availability |
|---|---|
| Helios rack (MI455X + EPYC Venice) | Full production now; shipping end of Q3 2026 |
| AMD-Cerebras disaggregated inference | Cloud availability H2 2026 |
| Anthropic MI450 deployment (1st GW) | Begins H1 2027 |
| MI455X engineering samples | H2 2026; mass production Q2 2027 |
All performance figures are AMD-reported and not yet independently benchmarked . The Helios pricing of ~$5.25M/rack and the Cerebras 5x throughput-per-watt claim are based on vendor modeling.