Particle.news

AMD Puts Helios Rack Into Production to Challenge Nvidia at Cloud Scale

The move signals AMD's push to sell complete 72‑GPU server cabinets and will be judged by early shipments, independent performance tests, and customer rollouts.

Overview

  • AMD announced at its Advancing AI event that Helios is in full production and that shipments will begin late in the third quarter with expansion through Q4, positioning the company as a rack‑scale systems supplier.
  • Each Helios cabinet pairs 72 Instinct MI455X GPUs with 6th‑Gen EPYC Venice CPUs and up to 432 GB of HBM4 per accelerator, delivering roughly 31 TB of HBM4 per rack and higher memory capacity than comparable Nvidia systems.
  • Major cloud and AI customers have contractual commitments totalling multi‑gigawatt scale, including OpenAI and Meta at about 6 GW each and Anthropic up to 2 GW, with Microsoft planning Helios use on Azure.
  • AMD has assembled an ecosystem of partners — Samsung for HBM4, Supermicro for OEM systems, Pensando for networking and TSMC for Venice CPUs — but near‑term execution risks include memory and GPU supply, system integration and ROCm software readiness.
  • Analysts raised forecasts after the launch and AMD framed a larger compute market opportunity, yet independent benchmarks and real‑world cost‑per‑token measures remain the critical tests that will determine whether Helios erodes Nvidia's dominance.