Skip to content
TECH CEO Daily

Microsoft to deploy AMD Helios on Azure for AI inference and add two EPYC Venice VM series

AMD said Helios shipments to customers including Microsoft start in the second half of 2026; neither company disclosed deployment scale or financial terms.

By · Editor

· Archive story, added · 3 min read · ✓ Fact-checked

The 60-second brief

  • 1Microsoft will deploy AMD Helios racks on Azure for frontier model inference; AMD said shipments begin in the second half of 2026.
  • 2Azure is adding HDv2 VMs for agentic AI and data pipelines and HXv2 VMs for chip design, both on EPYC Venice.
  • 3Neither company disclosed deployment scale or financial terms, and Microsoft gave no availability dates or pricing.

The news

On July 20, 2026, Advanced Micro Devices (AMD) and Microsoft (MSFT) said Microsoft will deploy AMD Helios on Azure, putting AMD's rack-scale AI system to work serving frontier AI models for Microsoft's own services and its cloud customers.

Helios packages AMD Instinct MI455X graphics processing units (GPUs), 6th Gen AMD EPYC "Venice" central processing units (CPUs), AMD Pensando networking and AMD's ROCm software into one integrated rack, according to AMD. The company said it will begin shipping Helios to customers, including Microsoft, in the second half of 2026.

In a blog post the same day, Scott Guthrie, Microsoft's executive vice president of Cloud + AI, said Helios will power a new Azure instance, ND MI455X v7, which Microsoft described as designed for reasoning, search and agentic AI workloads. AMD said enterprise customers can run production AI workloads on AMD-powered infrastructure through Azure Foundry Managed Compute.

The agreement also covers CPUs. Azure is adding two virtual machine (VM) series on EPYC Venice: HDv2 for agentic AI and data pipelines, and HXv2 for semiconductor design. Microsoft said HDv2 offers nearly 500 physical CPU cores, 4 terabytes of RAM, 32 terabytes of local NVMe storage and 400 Gb Azure Boost networking, while HXv2 offers 176 cores running above 5 GHz and 800 Gb InfiniBand.

On networking, AMD said Microsoft is broadening its use of AMD Pensando data processing units (DPUs), chips that take networking work off a server's main processors, and that the companies are integrating AMD silicon with Azure Boost to lift networking performance across Azure's fleet. Microsoft CEO Satya Nadella said the collaboration gives customers "the performance, scale and choice they need."

Neither company disclosed how many Helios racks Microsoft will deploy or what the agreement is worth, a gap WinBuzzer also noted in its July 23 coverage.

The numbers

Helios shipments to customers begin
Second half of 2026 (AMD)
New Azure VM series on EPYC Venice
2 (HDv2, HXv2)
HDv2 physical CPU cores
Nearly 500
HDv2 memory and local storage
4 TB RAM, 32 TB NVMe
HXv2 CPU cores
176, above 5 GHz
Deployment scale and deal value
Not disclosed

Why CEOs should care

For technology buyers, the practical change is another large pool of AI inference capacity inside a cloud many enterprises already use. Teams serving models on Azure should ask their Microsoft account team when ND MI455X v7 capacity will be offered, in which regions, and at what price compared with Nvidia-based instances. They should also test whether their models and serving software run on ROCm without major rework, because software portability decides whether a second chip supplier is usable in practice.

For CFOs negotiating multiyear Azure commitments, competing accelerator families inside one cloud can create pricing leverage, but only if workloads can move. Ask whether committed spend can shift between GPU families as new instances launch, and avoid locking long reservations to hardware whose pricing Microsoft has not yet published. Chip and hardware design firms should look at HXv2, which Microsoft aimed at register-transfer level (RTL) simulation and other engineering workloads.

For CISOs and boards, new silicon in the network path deserves the same scrutiny as any new cloud layer. Security teams can ask Microsoft how Pensando DPUs and the Azure Boost integration affect tenant isolation and firmware update practices. Boards should treat the announcement as a signal of direction rather than a capacity guarantee, since no deployment volumes were disclosed.

The bigger picture

The announcement fits a broader effort by cloud providers and AI developers to spread AI compute across more than one chip supplier. WinBuzzer reported that a Futurum Group estimate would put Nvidia's (NVDA) GPU market share at more than 95% and AMD's at about 4.5%, and quoted Counterpoint Research analyst Neil Shah, who attributes Nvidia's ecosystem lead to its CUDA software. AMD pitches Helios as an integrated rack of compute, networking and software, which lets it compete for whole-system AI deployments rather than individual chip sales.

What happened next

On July 23, 2026, at its Advancing AI 2026 event, AMD said Helios was in production and named Microsoft among the companies deploying it at scale. AMD said each Helios rack holds 72 MI455X GPUs and 18 EPYC CPUs. On August 4, 2026, AMD's second-quarter results listed the Azure Helios deployment, the new EPYC VM series and the Pensando expansion among its highlights, and CEO Lisa Su said Helios was beginning to ramp.

Microsoft's July 20 post did not give availability dates, regions or pricing for ND MI455X v7, HDv2 or HXv2. Those are the details to watch as AMD's second-half 2026 shipments proceed.

Written by

Editor · Technology & Business Writer

Hussein is a writer and business technology enthusiast focused on the intersection of technology, entrepreneurship, finance, artificial intelligence, and digital innovation.

CoversAICybersecurityBig TechSaaSStartupsFintech

How this story was made. Researched from primary sources such as company announcements and filings, with the help of technology tools, fact-checked twice, and approved for publication by Hussein Mukhtar.

Published by Tech CEO Daily, an independent publication. Masthead · Editorial standards · Report an error

Free newsletters

The technology briefing for people running businesses.

Daily, weekly, bi-weekly or monthly. You choose.

How often

The Daily Brief · Weekdays, 6 a.m. ET

Free forever. One click to unsubscribe. We never sell your email.

More in Big Tech