
AMD just got Microsoft on Helios, its first rack-scale AI system, shipping later this year. Meta, OpenAI, and Oracle were already on the early customer list. Azure says the racks will power frontier inference for Microsoft and customers, plus two new Venice CPU instance types for agent pipelines and chip design.
Futurum puts Helios around $5–5.5M per rack vs ~$3.5–4M for the Nvidia rack comps they use. AMD itself is selling TCO and “lowest cost per token,” not a public price card. Meta’s earlier plan still sits at up to 6 GW of AMD GPUs over time, with 1 GW on Helios later this year if that schedule holds.
For anyone renting GPU hours, this is more second-source rack supply and memory/bandwidth competition than a reason to re-rack a homelab. Ship dates and undisclosed capacity still matter more than the press photo from the Rockdale lab. I’ll believe the TCO claim when independent tokens/$ show up next to GB/VR class systems.

English