News
CAPACITY TEST:

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size

Newsroom brief

Microsoft will deploy AMD Helios rack-scale AI accelerators for Azure AI workloads, with watts, dollars and rack counts still absent from the public terms of the commitment.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: tomshardware.com
Microsoft Adds AMD Helios AI Racks To Azure Without Order Size
Image source: tomshardware.com

AMD's Helios rack-scale AI accelerator is moving onto Microsoft Azure for frontier-model workloads, giving the chipmaker a named hyperscale deployment before the order size is public.

Tom's Hardware reported that Microsoft and AMD announced the plan on July 20, 2026, with the systems intended for Microsoft's own data centres, Azure AI infrastructure customers and Microsoft Foundry users.

The commitment adds another accelerator option for cloud customers that need training and inference capacity.

The public terms stop before watts, dollars or rack counts, so the deployment is a cloud availability proof point for Helios instead of a measured share shift against Nvidia.

Helios Brings MI455X GPUs To Cloud AI Workloads

Tom's Hardware reported that the Helios rack joins 72 next-generation Instinct MI455X GPUs with 31.1TB of HBM4 memory across the system.

The same specification lists up to 1.4 exaFLOPS of FP8 compute and 2.9 exaFLOPS of FP4 compute for AI models using OCP AI data types.

AI labs are expected to use the AMD systems for training and inference serving, while enterprise workloads would run through Microsoft Foundry.

That delivery path puts the hardware inside cloud services, not only in a direct server procurement channel.

Interconnect performance is part of the rack-level design.

AMD is targeting 260 TB/s of scale-up bandwidth inside the Helios rack and 43 TB/s of scale-out bandwidth using UALink over Ethernet, Tom's Hardware's specification shows.

The article compares the scale-up figure with Nvidia's Vera Rubin NVL72 rack-scale system and puts the scale-out figure at about twice Vera Rubin's level.

UALink-over-Ethernet performance still has to be proven in practice.

Venice CPUs Add VM Series For AI And Chip Design

The Azure update extends beyond GPUs.

Microsoft and AMD also announced two VM series built on AMD's upcoming sixth-generation Epyc Venice CPUs: HDv2 for agentic AI and data pipelines, and HXv2 for semiconductor design workflows.

HDv2 targets data movement and agentic AI infrastructure.

HXv2 is aimed at electronic design and chip engineering tasks that need cloud capacity.

The cloud provider also plans to use its existing AMD Pensando DPU deployment with Azure Boost to accelerate networking and storage processing.

The package gives AMD a broader Azure footprint across GPUs, CPUs and DPUs.

Cloud customers trying to secure model-training and inference capacity would gain another hardware route through service contracts, while chip-design teams would get a Venice-based VM option for semiconductor workflows.

Helios Order Size Remains Outside The Public Terms

The order size is the main public-record gap.

Tom's Hardware wrote that the companies did not indicate the exact size of the Helios deployment in watts or dollars, even as the article placed the deal beside AMD's recent OpenAI and Meta partnerships involving gigawatts of compute installations and hundreds of billions of dollars potentially at stake.

The public record still lacks the first Helios wattage for the Azure rollout.

Share this article
inXf

Related articles

More
AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Chips & Semiconductors

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

Anthropic Builds Custom Silicon Team For Claude Scaling
Chips & Semiconductors

Anthropic Builds Custom Silicon Team For Claude Scaling

The Next Web reported that Anthropic publicly confirmed an in-house silicon team for Claude, while saying AWS, Google, Nvidia and AMD hardware remain part of its multi-chip scaling strategy.

Etched Raises $300 Million As $1 Billion Pre-Orders Test Inference Rack Plan
Chips & Semiconductors

Etched Raises $300 Million As $1 Billion Pre-Orders Test Inference Rack Plan

EE Times reported that Etched raised $300 million at a $10 billion pre-money valuation and has $1 billion in pre-orders, but the AI chip startup has not made public performance figures for racks due to ship this summer.

TSMC Raises Arizona Plan To $265 Billion After Profit Jump
Chips & Semiconductors

TSMC Raises Arizona Plan To $265 Billion After Profit Jump

CNBC reported that TSMC posted a 77.4% second-quarter profit jump and said its Arizona investment will rise to $265 billion after an additional $100 billion commitment. The report did not name the U.S. customers, fab opening dates, order volumes or pricing terms attached to the expanded plan.

Nvidia AI Push Names Banks, Carmakers And RIKEN
Chips & Semiconductors

Nvidia AI Push Names Banks, Carmakers And RIKEN

Tech Wire Asia reported that Nvidia used a country visit by Jensen Huang to connect Nemotron open models, DGX B200 systems and Blackwell supercomputers with local banks, carmakers, hospitals and research groups. The report said pilot-to-production conversion, audited benchmark results and chip-order volumes remain outside the public record.

Lenovo Whitsett AI Server Line Doubles Plant To 890,000 Square Feet
Chips & Semiconductors

Lenovo Whitsett AI Server Line Doubles Plant To 890,000 Square Feet

A Tom's Hardware tour details Lenovo's expanded Whitsett AI server line, including 890,000 square feet, up to 20 MW of power and liquid-cooling tests for GB300 and Vera Rubin racks.

Hugging Face Adds 4-Bit Nunchaku Loading To Diffusers
Chips & Semiconductors

Hugging Face Adds 4-Bit Nunchaku Loading To Diffusers

Hugging Face added Nunchaku Lite support to Diffusers, letting developers load 4-bit diffusion checkpoints with `from_pretrained()` while using Hub-delivered CUDA kernels.

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch
Chips & Semiconductors

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch

NVIDIA introduced Jetson T3000 and T2000 modules for edge AI and robotics systems, with Q1 2027 availability planned and source-listed performance, memory and emulation details still tied to company-provided figures.

Keep Reading

More Stories

Latest
Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaCloud & Data CentersAug 8, 2026Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaData Center Dynamics reported that Indosat, Ooredoo Group, Nokia and Nvidia launched Zankore by Indosat with a plan for up to 1GW of AI data centre capacity in Indonesia.Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightAIAug 8, 2026Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightCNBC reported that Black Hat cybersecurity leaders treated the Hugging Face AI-agent breach as a turning point for governing autonomous cyber models rather than a one-off failure.Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAIAug 8, 2026Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAI News reported that Alibaba plans revenue-sharing terms for some commercial users of its next Qwen open-weight AI model, following a licensing pattern already used by Moonshot for Kimi K3.Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanCapital & PolicyAug 8, 2026Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanArs Technica reported that a New Mexico judge ordered Meta to provide $567 million for treatment, screening, awareness and prevention after finding that its platforms contributed to a public nuisance.Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationAIAug 8, 2026Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationSiliconANGLE reported that Harvey AI is seeking at least $500 million in new funding that could value the legal AI startup at $15.5 billion after annualized revenue passed $350 million.Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaScience & TechAug 7, 2026Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaTech Collective SEA wrote that Shopee’s Vietnam share fell from 61% to 53% between May 2025 and April 2026 as TikTok Shop rose from 33% to 44%, showing how social commerce is reshaping regional ecommerce infrastructure.China Opens Security Review Of Palo Alto Networks ProductsCybersecurityAug 7, 2026China Opens Security Review Of Palo Alto Networks ProductsChina's cyberspace regulator opened a security review of Palo Alto Networks products, with no named product line, technical flaw or decision timetable disclosed.AI Pioneers Split Over Risk As Compute Buildout AcceleratesAIAug 7, 2026AI Pioneers Split Over Risk As Compute Buildout AcceleratesData Center Knowledge reported that Geoffrey Hinton, Fei-Fei Li and Andrew Ng disagreed at Ai4 over AI risk, jobs, openness and regulation, leaving infrastructure investors to plan capacity amid unsettled deployment rules.SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportTelco & ConnectivityAug 7, 2026SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportLight Reading reported that SpaceX urged the FCC to sunset High-Cost rural broadband subsidies, while rural telecom and electric-cooperative groups said LEO satellite coverage cannot replace terrestrial network support.OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutAIAug 7, 2026OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutBleepingComputer reported that OpenAI is rolling out GPT-5.6 Sol for paid ChatGPT users and GPT-5.6 Luna for Free and Go users, pairing unlimited free text chats with a new reasoning control and additional safeguards for users believed to be under 18.JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsCapital & PolicyAug 7, 2026JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsData Center Dynamics reported that JLL's EMEA Mid-Year Data Centre Report 2026 put FLAPD live capacity at 3.8GW, while the Middle East had 2.6GW in development paused and 13.8GW in planning.AWS Adds Persistent Runtime Instances For Production AI AgentsCloud & Data CentersAug 7, 2026AWS Adds Persistent Runtime Instances For Production AI AgentsAWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.