At its July 22–23 event in San Francisco, AMD announced its Helios rack scale platform is in full production and shipping by end of Q3 2026, a disaggregated inference partnership with Cerebras targeting 5x tokens per... AMD's Helios rack packs 72 Instinct MI455X GPUs, 31 TB of HBM4, and 2.9 FP4 exaFLOPS of inference...

Create a landscape editorial hero image for this Studio Global article: Search & fact-check with cited sources for What did AMD announce at its Advancing AI 2026 event regarding the Helios server platform enterin. Article summary: At its **Advancing AI 2026** event (July 22–23, San Francisco), AMD made three major announcements covering the Helios rack-scale platform, a Cerebras technical partnership, and an Anthropic strategic deal. Here is the b. Topic tags: general, general web, user generated, news. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts w
At its Advancing AI 2026 event (July 22–23, San Francisco), AMD made three major announcements: the Helios rack-scale platform entering full production, a technical partnership with Cerebras for low-latency AI inference, and a strategic deal with Anthropic. Here is the breakdown with key specifications and timelines.
CEO Lisa Su confirmed the Helios servers are in full production and will begin shipping at the end of Q3 2026 . AMD describes Helios as the industry's highest-performance AI rack for training and running frontier models at massive scale
.
Key specifications (per single Helios rack) — from AMD's official product page and brochure :
The MI455X delivers roughly 2x the FP4 compute of the previous MI350 series . The Helios system is AMD's primary rack-scale competitor to Nvidia's Vera Rubin platform and has already lined up customers including Microsoft, Meta, OpenAI, Oracle, and Anthropic
.
Announced on stage at Advancing AI 2026 (July 23), AMD and Cerebras are building a disaggregated inference platform that splits AI inference into two phases :
The two systems operate as a single disaggregated solution via high-speed interconnect .
Claimed performance: Up to 5x higher tokens per second per watt compared to conventional architectures (based on July 2026 modeling using the Kimi 2.6 1T model) .
Availability: The joint AMD-Cerebras AI inference solution is expected to be available through cloud providers in H2 2026 . Cerebras will deploy Helios systems in its own data centers, with the solution available first through Cerebras Cloud
.
Cerebras CEO Andrew Feldman said at the event that the architecture directly targets Nvidia's Groq LPUs and is positioned for agentic AI workloads where low-latency inference is critical .
Announced July 22, 2026, via AMD official press release :
Deal terms:
This adds Anthropic to AMD's growing roster of major AI customers alongside OpenAI, Meta, Oracle, and Microsoft . The MI450 Series mentioned in this deal is the generation of GPU that powers the Helios racks — the MI455X is the specific SKU used in the first Helios production units
.
All performance figures are AMD-reported and not yet independently benchmarked . The Helios pricing of ~$5.25M/rack and the Cerebras 5x throughput-per-watt claim are based on vendor modeling.
Studio Global AI
Use this topic as a starting point for a fresh source-backed answer, then compare citations before you share it.
At its July 22–23 event in San Francisco, AMD announced its Helios rack scale platform is in full production and shipping by end of Q3 2026, a disaggregated inference partnership with Cerebras targeting 5x tokens per...
At its July 22–23 event in San Francisco, AMD announced its Helios rack scale platform is in full production and shipping by end of Q3 2026, a disaggregated inference partnership with Cerebras targeting 5x tokens per... AMD's Helios rack packs 72 Instinct MI455X GPUs, 31 TB of HBM4, and 2.9 FP4 exaFLOPS of inference compute, with an estimated price of $5.25 million per rack.
The Cerebras joint solution splits inference into prefill (Helios) and decode (Wafer Scale Engine), with cloud availability expected in H2 2026.