AMD Helios AI rack system takes direct aim at Nvidia with Microsoft, OpenAI backing

Reviewed byNidhi Govil

35 Sources

Share

AMD launched Helios, its first rack-scale AI platform designed to compete with Nvidia's dominance in data center computing. The system combines 72 Instinct MI455X GPUs and has already secured major customers including Microsoft, OpenAI, Meta, and Anthropic. With superior performance metrics and gigawatt-scale deployments planned, AMD positions itself as a serious Nvidia competitor in the AI accelerator market.

AMD Helios Emerges as First True Nvidia Competitor in Rack-Scale AI

At AMD's sold-out Advancing AI event in San Francisco, Chair and CEO Dr. Lisa Su unveiled Helios, calling it the tech industry's "highest performance AI rack" built to train and run the most demanding frontier AI models at massive scale

1

. The AMD Helios AI rack system represents the company's first genuine attempt to compete with Nvidia's established Grace Blackwell and Vera Rubin rack-scale systems that have dominated AI data center infrastructure for years

3

.

Source: TweakTown

Source: TweakTown

The rack-scale AI platform measures 1.2 meters wide and 44 rack units high, nearly twice the size of Nvidia's NVL72 system, and AMD has utilized this extra space strategically

3

. Helios boasts 50 percent more HBM4 memory and scale-out bandwidth compared to Vera Rubin, with between 15 and 25 percent higher performance for AI training workloads

3

. AMD estimates this performance advantage will translate to a 30 percent performance-per-dollar lead over the competition, a metric that could prove decisive for cost-conscious hyperscalers

3

.

Instinct MI455X Powers AMD's Challenge to Nvidia

The Instinct MI455X GPU serves as the foundation of Helios, with AMD calling it "by leaps and bounds the most advanced AI accelerator we've ever built"

2

. This massive chip encompasses 320 billion transistors and employs advanced chiplet design using TSMC's cutting-edge process technologies

2

.

Source: Guru3D

Source: Guru3D

The CDNA 5 architecture underlying the MI455X brings substantial improvements, with four Accelerator Complex Dies fabricated on TSMC's 2nm gate-all-around process technology stacked atop Fabric and Cache Dies built on TSMC N3P

2

. Each of the two Fabric and Cache Dies features 96MB of L2 cache, delivering 1.5 times higher bandwidth per die compared to CDNA 4's Infinity Cache

3

.

The MI455X delivers particularly impressive gains for lower-precision floating-point formats now common in inference workloads. Performance for OCP MXFP8 and MXFP4 formats reaches up to four times faster than the previous-generation MI355X

2

. The complete Helios system offers an aggregate 1.4 exaFLOPS of FP8 compute and 2.9 exaFLOPS of FP4 across its 72 GPUs, with 31.1TB of HBM4 memory capacity

4

.

Microsoft Azure Joins Growing Customer Base

Microsoft announced Monday it will expand its cloud infrastructure with AMD Helios to give customers the performance, scale, and choice needed to build next-generation AI applications, according to CEO Satya Nadella

5

. Microsoft Azure will deploy the system to power frontier model inference for its AI customers and support Azure AI services

4

.

Source: DT

Source: DT

The partnership extends a longtime relationship between the companies, with AMD chips already powering Microsoft's Surface PCs and Xbox gaming consoles

5

. Microsoft joins an impressive roster of Helios customers including OpenAI, Meta, Oracle, and Anthropic, all planning to deploy the system

1

. Anthropic and AMD announced a strategic partnership Wednesday to deploy up to two gigawatts of GPUs via the new AI rack system

1

.

AMD says eight of the top 10 AI companies now run workloads on its Instinct GPUs, including Cohere and Elon Musk's SpaceXAI

5

. The company will begin shipping Helios to customers, including Microsoft, later this year, though financial terms and exact compute capacity commitments remain undisclosed

5

.

Agentic AI Drives Trillion-Dollar Market Forecast

During her remarks at the Advancing AI event, Dr. Lisa Su outlined the trajectory of the chip industry, predicting that by 2030, the AI accelerator market will reach approximately $1.4 trillion

1

. This represents a market approaching the size of the entire semiconductor industry today, driven largely by the rise of agentic AI

1

.

Su explained that agentic AI creates a step change in compute demand because when you ask an agent to complete a task, "it actually has dozens of steps, and it has to reason, and it has to call tools, and it has to access data, and it has to keep doing it over and over until it solves the problem, and so you need lots of GPUs to do all that"

1

. GPUs are expected to constitute the vast majority of this market because algorithms remain in their infancy and workloads continue evolving, favoring programmability in the silicon ecosystem

1

.

Microsoft will also leverage AMD's upcoming Epyc Venice CPUs, adding two new Azure VM series: the HDv2 series for agentic AI and data pipelines, and the HXv2 for semiconductor design workflows

4

. AMD also introduced its Venice-X CPU designed for AI data center workloads, expected to launch in 2027

1

.

Today's Top Stories

© 2026 TheOutpost.AI All rights reserved