News
CAPACITY TEST:

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size

Newsroom brief

Microsoft will deploy AMD Helios rack-scale AI accelerators for Azure AI workloads, with watts, dollars and rack counts still absent from the public terms of the commitment.

Verified against source materialEdited by SendTech Times Chips & Compute Desk
Microsoft Adds AMD Helios AI Racks To Azure Without Order Size
Image source: tomshardware.com

AMD's Helios rack-scale AI accelerator is moving onto Microsoft Azure for frontier-model workloads, giving the chipmaker a named hyperscale deployment before the order size is public.

Tom's Hardware reported that Microsoft and AMD announced the plan on July 20, 2026, with the systems intended for Microsoft's own data centres, Azure AI infrastructure customers and Microsoft Foundry users.

The commitment adds another accelerator option for cloud customers that need training and inference capacity.

The public terms stop before watts, dollars or rack counts, so the deployment is a cloud availability proof point for Helios instead of a measured share shift against Nvidia.

Helios Brings MI455X GPUs To Cloud AI Workloads

Tom's Hardware reported that the Helios rack joins 72 next-generation Instinct MI455X GPUs with 31.1TB of HBM4 memory across the system.

The same specification lists up to 1.4 exaFLOPS of FP8 compute and 2.9 exaFLOPS of FP4 compute for AI models using OCP AI data types.

AI labs are expected to use the AMD systems for training and inference serving, while enterprise workloads would run through Microsoft Foundry.

That delivery path puts the hardware inside cloud services, not only in a direct server procurement channel.

Interconnect performance is part of the rack-level design.

AMD is targeting 260 TB/s of scale-up bandwidth inside the Helios rack and 43 TB/s of scale-out bandwidth using UALink over Ethernet, Tom's Hardware's specification shows.

The article compares the scale-up figure with Nvidia's Vera Rubin NVL72 rack-scale system and puts the scale-out figure at about twice Vera Rubin's level.

UALink-over-Ethernet performance still has to be proven in practice.

Venice CPUs Add VM Series For AI And Chip Design

The Azure update extends beyond GPUs.

Microsoft and AMD also announced two VM series built on AMD's upcoming sixth-generation Epyc Venice CPUs: HDv2 for agentic AI and data pipelines, and HXv2 for semiconductor design workflows.

HDv2 targets data movement and agentic AI infrastructure.

HXv2 is aimed at electronic design and chip engineering tasks that need cloud capacity.

The cloud provider also plans to use its existing AMD Pensando DPU deployment with Azure Boost to accelerate networking and storage processing.

The package gives AMD a broader Azure footprint across GPUs, CPUs and DPUs.

Cloud customers trying to secure model-training and inference capacity would gain another hardware route through service contracts, while chip-design teams would get a Venice-based VM option for semiconductor workflows.

Helios Order Size Remains Outside The Public Terms

The order size is the main public-record gap.

Tom's Hardware wrote that the companies did not indicate the exact size of the Helios deployment in watts or dollars, even as the article placed the deal beside AMD's recent OpenAI and Meta partnerships involving gigawatts of compute installations and hundreds of billions of dollars potentially at stake.

The public record still lacks the first Helios wattage for the Azure rollout.

Share this article
inXf

Related articles

More
IBM Adds Rackmount z17 And LinuxONE 5 Systems Without Customer Counts
Chips & Semiconductors

IBM Adds Rackmount z17 And LinuxONE 5 Systems Without Customer Counts

IBM is extending its z17 and fifth-generation LinuxONE families with single-frame and rackmount systems, including LinuxONE Express as an 18U entry point. The smaller hardware brings Telum II and Spyre AI capability to lower-scale deployments, Customer counts, pricing and shipment volumes for the new models remain outside the public record.

Qualcomm AI250 Stacks DRAM Over Compute But Leaves FLOPS Undisclosed
Chips & Semiconductors

Qualcomm AI250 Stacks DRAM Over Compute But Leaves FLOPS Undisclosed

Qualcomm is pitching high-bandwidth compute for AI inference, with AI250 cards claiming 768 GB of memory and 133 TB/s of effective bandwidth, while the public record still lacks peak FLOPS or named customers.

Imec Says AI Inference Pushes Optical I/O Closer To Chips
Chips & Semiconductors

Imec Says AI Inference Pushes Optical I/O Closer To Chips

Imec researchers told EE Times that AI inference is shifting the interconnect problem from rack-scale optics towards 2.5D and 3D optical I/O near processors. The group discussed a 250 Tb/s bandwidth projection while the public record still lacks commercial product timing, customer deployments or cooling test results.

Unconventional AI Tests Oscillator Models Before Power-Efficient Chip Proof
Chips & Semiconductors

Unconventional AI Tests Oscillator Models Before Power-Efficient Chip Proof

Unconventional AI has released the Un-0 model series to test oscillator-based image generation, but the work still runs on simulated oscillators rather than a physical AI accelerator.

Micron And GM Sign Memory Supply Deal Without Volumes Or Pricing
Chips & Semiconductors

Micron And GM Sign Memory Supply Deal Without Volumes Or Pricing

Micron and General Motors announced a strategic customer agreement covering long-term supply of LPDRAM, NOR and UFS NAND for future vehicle platforms. The company statement links the deal to Micron’s $2 billion Manassas fab modernisation and says it is one of 16 strategic customer agreements discussed on Micron’s fiscal third-quarter 2026 call, The public record still lacks order volumes, pricing or covered vehicle lines.

AMD EPYC 8005 Raises SP6 Core Counts Without Customer Rollout Data
Chips & Semiconductors

AMD EPYC 8005 Raises SP6 Core Counts Without Customer Rollout Data

ServeTheHome reported that AMD EPYC 8005 “Sorano” keeps the SP6 server socket while reaching 84 cores, DDR5-6400 memory and CXL 2.0. The sponsored test material disclosed AMD-supplied CPUs, leaving customer deployments and order evidence outside the public record.

Keep Reading

More Stories

Latest
e& UAE And Core42 Launch Sovereign AI Compute PlatformCloud & Data CentersJul 21, 2026e& UAE And Core42 Launch Sovereign AI Compute PlatformMiddle East AI News reported that e& UAE and Core42 launched Sovereign AI Compute, giving UAE enterprises and government bodies in-country GPU access with data residency, connectivity and vendor-claimed zero egress fees.AI Coding Agents Face Sandbox-Escape Findings Across Four ToolsCybersecurityJul 21, 2026AI Coding Agents Face Sandbox-Escape Findings Across Four ToolsBleepingComputer reported that Pillar Security reproduced sandbox-escape paths in Cursor, OpenAI Codex, Gemini CLI and Google Antigravity, shifting attention from agent containment to trusted developer tools around the workspace.AliExpress Hit With Record €550m EU Fine Over Illegal GoodsCapital & PolicyJul 21, 2026AliExpress Hit With Record €550m EU Fine Over Illegal GoodsBBC reported that the European Commission imposed a record €550m Digital Services Act penalty on AliExpress and ordered the Alibaba-owned marketplace to file a corrective action plan by 20 October.Neo Raises $100M To Control Enterprise AI Software ActionsCybersecurityJul 21, 2026Neo Raises $100M To Control Enterprise AI Software ActionsSecurityWeek reported that Neo emerged from stealth with $100 million for a platform that governs AI agents, MCP servers and software actions across enterprise systems.Z.ai Tests Gigawatt AI Data Centre Built On Domestic ChipsCloud & Data CentersJul 21, 2026Z.ai Tests Gigawatt AI Data Centre Built On Domestic ChipsUnite.AI reported that Z.ai has begun operating part of a gigawatt-class AI data centre built on Chinese-made chips, highlighting how export controls are pushing large-scale model training toward domestic compute stacks.Digital Takumi AWS Route 53 Outage Exposes Root-Account Recovery RiskCloud & Data CentersJul 21, 2026Digital Takumi AWS Route 53 Outage Exposes Root-Account Recovery RiskThe Register links a suspended AWS Route 53 account to Digital Takumi-managed websites and connected Google Workspace email going offline after billing warnings, MFA recovery and DNS hosting sat inside one operating loop.SBI Takes Majority Control Of Coinhako After MAS ApprovalFintech & Digital PaymentsJul 21, 2026SBI Takes Majority Control Of Coinhako After MAS Approvale27 reported that SBI Holdings acquired a majority stake in Singapore crypto exchange Coinhako after MAS approval, using a capital injection and share purchase to turn the platform into a consolidated subsidiary. Financial terms and product timelines were not disclosed.PULSE Gives 10 US Health Agencies OpenAI And Anthropic AI PilotsAIJul 21, 2026PULSE Gives 10 US Health Agencies OpenAI And Anthropic AI PilotsPULSE will let 10 US public health jurisdictions test OpenAI and Anthropic enterprise AI tools, with capacity for up to 2,000 practitioners and playbooks expected in 2027, AI News reported.PJM Backup-Generator Warnings Test Large-Load Grid ToolCloud & Data CentersJul 21, 2026PJM Backup-Generator Warnings Test Large-Load Grid ToolData Center Knowledge reported that PJM issued emergency backup-generator warnings during a July heat wave but did not dispatch customer-owned generators, leaving total available large-load backup capacity undisclosed.Taiwan Mobile Extends Nokia 5G Deal Around AI-Native Network OperationsTelco & ConnectivityJul 21, 2026Taiwan Mobile Extends Nokia 5G Deal Around AI-Native Network OperationsRCR Wireless News reported that Taiwan Mobile and Nokia have extended their 5G partnership with AI-native network operations, including AirScale equipment, MantaRay SON automation, RedCap support and a 100% renewable-electricity target by 2040.AI Agent Rollouts Require Testing Before Live CustomersAIJul 21, 2026AI Agent Rollouts Require Testing Before Live CustomersNo Jitter reported that enterprise AI agents need guardrails, simulations, answer checks and visibility before customer-facing deployment, as vendors add tools to catch regressions and rollback failures.FCC Names Three Unlicensed Bands For Satellite D2D VoteTelco & ConnectivityJul 21, 2026FCC Names Three Unlicensed Bands For Satellite D2D VoteAn FCC draft order names 902MHz-928MHz, 2400MHz-2483.5MHz and 5725MHz-5850MHz for a proposed unlicensed direct-to-device satellite proceeding ahead of an August 6 commissioner vote.