Vultr Brings AMD's 72-GPU Helios Rack to the Cloud — Specs, Timeline, and Why It Matters
On July 23, 2026, Vultr announced support for AMD's Helios rackscale system, which packs 72 Instinct MI455X GPUs, 31 TB of HBM4 memory, and 1.67 PB/s memory bandwidth into a single rack. The Helios rack includes 18 sixth generation EPYC 'Venice' CPUs, AMD Pensando networking, and ROCm 7 software, and is one of AMD's...
Published byEdited with DeepSeek-V4-FlashImages generated with GPT Image 1.5
On July 23, 2026, Vultr announced support for AMD's Helios rackscale system, which packs 72 Instinct MI455X GPUs, 31 TB of HBM4 memory, and 1.67 PB/s memory bandwidth into a single rack.
The Helios rack includes 18 sixth generation EPYC 'Venice' CPUs, AMD Pensando networking, and ROCm 7 software, and is one of AMD's most direct challenges yet to Nvidia's Vera Rubin NVL72.
All performance figures are vendor reported and not yet independently benchmarked — a key caveat as AMD pushes its first true rack scale platform into production.
Search & fact-check with cited sources for What are the key details of Vultr's July 23, 2026, announcement to support AMD's Instinct MI455XA conceptual rendering of the AMD Helios rack-scale AI system, featuring 72 Instinct MI455X GPUs and 31 TB of HBM4 memory.
AI Prompt
Create a landscape editorial hero image for this Studio Global article: Search & fact-check with cited sources for What are the key details of Vultr's July 23, 2026, announcement to support AMD's Instinct MI455X. Article summary: Here is a comprehensive, source-cited breakdown of Vultr's July 23, 2026, announcement and the broader AMD Helios ecosystem.. Topic tags: general, general web, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake numbers, clickbait thumbnails, icons, and tiny thumbnail layouts. Make it useful as an illustr
openai.com
On July 23, 2026, Vultr announced support for the AMD Helios rackscale solution powered by AMD Instinct MI455X GPUs, making it one of the first cloud providers — alongside Microsoft Azure — to offer the new hardware . The announcement, timed with AMD's Advancing AI 2026 event in San Francisco, positions Vultr to serve as a global provider for the system .
Studio Global AI
Continue your research
This page includes a source-backed answer you can continue inside Studio Global.
What is the short answer to "Vultr Brings AMD's 72-GPU Helios Rack to the Cloud — Specs, Timeline, and Why It Matters"?
On July 23, 2026, Vultr announced support for AMD's Helios rackscale system, which packs 72 Instinct MI455X GPUs, 31 TB of HBM4 memory, and 1.67 PB/s memory bandwidth into a single rack.
What are the key points to validate first?
On July 23, 2026, Vultr announced support for AMD's Helios rackscale system, which packs 72 Instinct MI455X GPUs, 31 TB of HBM4 memory, and 1.67 PB/s memory bandwidth into a single rack. The Helios rack includes 18 sixth generation EPYC 'Venice' CPUs, AMD Pensando networking, and ROCm 7 software, and is one of AMD's most direct challenges yet to Nvidia's Vera Rubin NVL72.
What should I do next in practice?
All performance figures are vendor reported and not yet independently benchmarked — a key caveat as AMD pushes its first true rack scale platform into production.
The rack is a 7,000-pound double-wide unit drawing up to 245 kW .
Integration of EPYC 'Venice' CPUs and Pensando Networking
The Helios rack integrates 18 sixth-generation AMD EPYC "Venice" CPUs — one per four GPUs across 18 compute trays — linked via AMD Pensando networking hardware for lossless, high-throughput AI cluster communication . This co-optimized silicon approach, combining GPUs, CPUs, networking, and software into a single integrated system, is AMD's first true rack-scale AI platform and its most direct answer to Nvidia's Vera Rubin NVL72 .
Deployment Timeline: Pre-Orders in Q4 2026, Production from 2027 to 2028
Pre-orders for reserved capacity opened in Q4 2026
The MI455X GPU is due in Q4 2026
Helios deployments on Vultr are planned from 2027 through 2028
Engineering samples and limited-volume Helios production ship in the second half of 2026, with full mass-production ramp and first production AI tokens not expected until Q2 2027, per industry timelines .
Deployment Options: Bare Metal, Virtualized Instances, and Kubernetes
Vultr offers three deployment models for the Helios platform :
Bare metal — direct access to physical GPU servers
Virtualized instances — composable cloud infrastructure via the Vultr Console and API
Kubernetes — via the free Vultr Kubernetes Engine (VKE) control plane for repeatable, cloud-native AI deployment and scaling
Customers can also use Vultr Clusters (self-service) with a Slurm or Kubernetes workload scheduler and an observability suite for GPU health monitoring . Vultr is positioning the new systems for model training, inference, fine-tuning, and agentic AI as businesses move from pilot projects to production .
Broader Cloud Momentum for AMD Helios
Microsoft Azure
On July 20, 2026 — three days before Vultr's announcement — Microsoft announced it will deploy the AMD Helios rackscale solution at scale on Azure for frontier-model AI inference, customer applications, and Microsoft's own AI services . Volume deployments are planned for the second half of 2026. The expanded partnership covers Instinct GPUs, EPYC CPUs, Pensando networking, and ROCm software . Microsoft's official blog noted three new Azure VM families: HDv2, HXv2, and ND MI455X v7 .
Other Named Customers
Anthropic announced a strategic partnership to deploy up to 2 gigawatts of AMD Instinct MI450 Series GPUs in Helios rack-scale systems
Meta and OpenAI are also named as Helios customers
TensorWave will operate an all-AMD AI cloud built on Helios racks
Alignment Engine plans Helios deployments at an Ohio data center
Schneider Electric co-developed a validated reference design for Helios deployment
HPE
The available evidence does not directly confirm a December 2025 HPE agreement or its 2026 availability timeline. No specific source for that claim was found.
Competitive Context: AMD's Rack-Scale Strategy vs. Nvidia
AMD Helios directly competes with Nvidia's Vera Rubin NVL72 . Key differentiating elements:
Memory leadership: 31 TB HBM4 per rack (72 GPUs × 432 GB each) versus Nvidia's Rubin at roughly 20 TB — a ~50% capacity advantage per GPU . Per-rack memory bandwidth is roughly 1.67 PB/s .
Paper specs vs. reality: The Register notes that Helios is bigger and faster than Vera Rubin "on paper" but that AMD has yet to prove performance in production deployments .
UALink over Ethernet (260 TB/s scale-up): AMD's open-standard alternative to Nvidia's proprietary NVLink. UALink is built as an Ethernet-based, open industry standard .
UEC-aligned Ethernet (43 TB/s scale-out): Helios uses Ultra Ethernet Consortium (UEC)-aligned Ethernet for open scale-out networking, avoiding proprietary InfiniBand .
Open standards positioning: AMD is positioning Helios as a more open, standards-based alternative to Nvidia's vertically integrated rack-scale systems, aiming to appeal to hyperscalers and cloud providers seeking vendor diversity .
Training gap acknowledged: Multiple reports note that while Helios leads in memory and inference specs, its training performance trails Nvidia's offerings, and all figures are vendor-reported and not yet independently benchmarked .
Pricing: The Helios rack is listed at approximately $5.25 million per unit .
webronaq.com
AMD Helios Rack-Scale AI System: Specs, UALink, and ...