The Helios rack integrates 18 sixth-generation AMD EPYC "Venice" CPUs — one per four GPUs across 18 compute trays — linked via AMD Pensando networking hardware for lossless, high-throughput AI cluster communication . This co-optimized silicon approach, combining GPUs, CPUs, networking, and software into a single integrated system, is AMD's first true rack-scale AI platform and its most direct answer to Nvidia's Vera Rubin NVL72
.
Engineering samples and limited-volume Helios production ship in the second half of 2026, with full mass-production ramp and first production AI tokens not expected until Q2 2027, per industry timelines .
Customers can also use Vultr Clusters (self-service) with a Slurm or Kubernetes workload scheduler and an observability suite for GPU health monitoring . Vultr is positioning the new systems for model training, inference, fine-tuning, and agentic AI as businesses move from pilot projects to production
.
On July 20, 2026 — three days before Vultr's announcement — Microsoft announced it will deploy the AMD Helios rackscale solution at scale on Azure for frontier-model AI inference, customer applications, and Microsoft's own AI services . Volume deployments are planned for the second half of 2026
. The expanded partnership covers Instinct GPUs, EPYC CPUs, Pensando networking, and ROCm software
. Microsoft's official blog noted three new Azure VM families: HDv2, HXv2, and ND MI455X v7
.
The available evidence does not directly confirm a December 2025 HPE agreement or its 2026 availability timeline. No specific source for that claim was found.