End-to-End Hardware for Next-Gen Intelligence: AMD Delivers Full-Stack Compute for Agentic AI
AMD launches broad high-performance computing and physical AI portfolios, including its first rack-scale AI solution and the world’s most powerful AI rack, AMD Helios.
AMD (NASDAQ: AMD) today launched its next-generation AI infrastructure and physical AI portfolio at Advancing AI 2026, led by AMD Helios rackscale solutions, now in production to be deployed by leading AI companies at gigawatt scale.
As AI expands from training to inference and agentic workloads, compute demand is accelerating rapidly. AMD delivers an open, full-stack AI platform that gives customers the flexibility to deploy the right compute for every workload.
“The next phase of AI will span frontier models, agents and physical AI, creating new opportunities to bring intelligence everywhere,” said Dr. Lisa Su, chair and CEO, AMD. “Realizing that potential will take the entire industry working together. AMD is partnering across the ecosystem to deliver leadership compute and open platforms that give customers the performance, flexibility and choice to scale AI from the data center to the edge.”
AMD Helios: The Highest Performance Rack-Scale AI Solution
Delivering frontier AI requires a fully integrated rack architecture, with every part of the stack pushing the boundaries of performance. AMD Helios rackscale solutions are built for this, with co-optimized silicon spanning 72 high-performance AMD Instinct MI455X GPUs and 18 powerful 6th Gen AMD EPYC “Venice” CPUs, connected by AMD Pensando front-end, scale-up and scale-out networking, and accelerated by AMD ROCm open software. AMD Helios combines leadership compute performance, memory capacity and networking bandwidth to deliver up to 30% more tokens per dollar than the leading competitive solution.1
Leading AI labs and cloud providers are choosing AMD Helios for its open, full-stack performance. They include OpenAI, Anthropic, Meta, Microsoft, Oracle, HUMAIN, Tensorwave, Vultr, Cirrascale and others. Systems will be available from leading OEMs, including Bull, HPE, Lenovo and Supermicro, as well as infrastructure partners Sanmina and Wiwynn.
At Advancing AI, AMD partners detailed how they deploy AMD AI infrastructure at scale for frontier training and inference:
- Anthropic and AMD further outlined Wednesday’s strategic partnership announcement to deploy up to 2 gigawatts of AMD Instinct MI455X GPUs in AMD Helios rackscale solutions. The companies are launching a multiyear engineering collaboration to use Claude to accelerate AMD software development. Specifically, the teams will use Claude to optimize workloads for AMD Instinct GPUs and accelerate ROCm software development. AMD will also broadly adopt Claude across its engineering and product development teams.
- OpenAI and AMD are partnering to optimize the full AI stack, from silicon to software. Leveraging OpenAI’s Triton framework with AMD ROCm software, the companies are optimizing GPT-class workloads on AMD Instinct MI455X GPUs and AMD Helios racks. OpenAI expects to bring Helios online beginning in the fourth quarter of 2026, with deployments accelerating throughout 2027.
- Meta and AMD are co-designing for gigawatt-scale deployments, optimizing AMD’s full AI compute stack for Meta workloads. Meta is now validating 6th Gen EPYC CPU platforms in its labs and has begun testing and validating workloads on AMD Helios racks as they prepare to deploy at scale.
- Cerebras and AMD are collaborating to deliver a combined solution of Cerebras ultra-low-latency AI compute and AMD Helios high-throughput rack-scale infrastructure to help improve inference efficiency, scalability and economics for ultra-low-latency inference serving.
Comments
Post a Comment