News
CAPACITY TEST:

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Newsroom brief

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: datacenterknowledge.com
AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Image source: datacenterknowledge.com

AMD has moved its Helios rack-scale AI system into full production, giving the chipmaker an integrated platform to compete with Nvidia beyond individual accelerators.

Data Center Knowledge reported that hardware partners are scheduled to begin shipments by the end of the third quarter of 2026.

The system combines Instinct MI455X accelerators, Epyc server processors, Pensando networking and ROCm software.

AMD said Microsoft and Anthropic are the first named deployment routes for the integrated architecture.

A Rack-Scale Platform Built Around Venice

Helios follows the integrated design now shaping high-end AI infrastructure: processors, accelerators, networking and software are engineered as one rack rather than assembled as separate components.

The approach is intended for increasingly complex agentic workloads that require CPUs to coordinate data retrieval, tool calls and other tasks around GPU reasoning.

At the July 23 Advancing AI event in San Francisco, Chief Executive Lisa Su said AI agents could expand from millions to billions.

Her presentation positioned server processors as an important part of that growth, not simply as hosts for accelerators.

AMD said the Venice family includes configurations with up to 256 cores: SP8 variants span 8 to 128 cores, the SP7 range reaches 256 cores and the LP host-node processor reaches 72 cores.

The lineup covers enterprise servers, cloud systems and specialised AI host nodes.

Performance Claims Centre On The Full System

Su said Helios delivers 15% more AI compute, 50% more memory, 50% more scale-out bandwidth and up to 30% more tokens per dollar than Nvidia's competing platform.

AMD's launch material also compared a 256-core Venice configuration with Nvidia's 88-core Vera CPU, claiming 2.2 times the throughput, and cited a 20% per-core advantage for a 96-core Venice chip.

According to AMD's launch benchmarks, MI455X provides 34 times the token throughput and up to 18 times the token-cost performance of the previous MI355X generation.

Taken together, the vendor's figures present a performance and capacity case for evaluating the rack as an alternative to Nvidia's integrated designs.

Microsoft And Anthropic Establish Deployment Routes

AMD identified Microsoft as a customer for frontier-model inference and Azure AI services.

Su also said Anthropic intends to deploy as much as 2 GW of accelerators in the rack-scale systems.

Those commitments move the launch beyond a component roadmap and support a broader effort to sell CPUs, accelerators, networking and software as one infrastructure stack.

In the same presentation, Su projected a $1.4 trillion AI infrastructure market by 2030, replacing her earlier $500 billion estimate for 2028.

She put the server CPU opportunity at $200 billion by 2030, up from $25 billion today.

The presentation also attributed 46% of latest-quarter server-market revenue to the processor family and listed deployments at more than 60% of Fortune 100 companies.

The Processor Roadmap Runs Through 2030

AMD said its roadmap schedules the SP7 version for the fourth quarter of 2026, SP8 for the first half of 2027 and the LP host-node processor for the second half of that year.

It places seventh-generation Florence CPUs in 2028 and an eighth generation in 2030.

On the accelerator side, the roadmap calls for a new GPU and Helios generation each year, including Instinct MI500 in 2027 and MI600 in 2028.

ROCm.AI is intended to support that cadence with command-line tools, reusable skills and AI-assisted software optimisation across the hardware portfolio.

The remaining evidence gap is operational.

Public material establishes production status, shipment timing, product specifications and customer commitments, but it does not provide order sizes, deployed-cluster economics or independently reproduced benchmarks.

Share this article
inXf

Related articles

More
US Export-Control Shift Opens UAE AI Chip Access, DCD Reports
Chips & Semiconductors

US Export-Control Shift Opens UAE AI Chip Access, DCD Reports

Data Center Dynamics reported that the US government has moved the UAE into a lower-restriction export-control category, giving UAE companies broader access to advanced AI chips from Nvidia and AMD. The account names G42 and hyperscaler data-centre projects, but it does not list chip volumes, approved licences or specific UAE projects.

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs
Chips & Semiconductors

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs

The Register reported that Alibaba launched Qwen 3.8-Max as a downloadable open-weight model while DeepSeek refreshed V4 Flash, sharpening the price and performance challenge facing US AI labs.

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace
Chips & Semiconductors

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace

SiliconANGLE reported that Compute Exchange has launched a marketplace for used and refurbished Nvidia H100 and A100 GPUs. Compute Exchange said requests range from hundreds to tens of thousands of GPUs, but the public launch did not name suppliers, customers or pricing benchmarks.

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size
Chips & Semiconductors

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size

Microsoft will deploy AMD Helios rack-scale AI accelerators for Azure AI workloads, with watts, dollars and rack counts still absent from the public terms of the commitment.

Maybank Sees AI Demand Driving Singapore Semiconductor Growth
Chips & Semiconductors

Maybank Sees AI Demand Driving Singapore Semiconductor Growth

Maybank Investment Bank said AI-related demand should remain the main growth driver for Singapore semiconductor companies, while warning that large AI infrastructure spending still needs stronger revenue proof.

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push
Chips & Semiconductors

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push

Rest of World reported that Saudi Arabia and the UAE are trying to diversify AI chip supply while major projects still rely on Nvidia hardware. The report cited Humain data-centre plans, G42’s Stargate project and analyst warnings that U.S. approvals, TSMC capacity and high-bandwidth memory remain constraints.

Nvidia-SK AI Deal Carries $500 Billion Value With Few Build Details
Chips & Semiconductors

Nvidia-SK AI Deal Carries $500 Billion Value With Few Build Details

Tom's Hardware reported that Nvidia and SK Group signed letters of intent for a strategic relationship valued at more than $500 billion, anchored by SK Telecom's planned 2-gigawatt South Korea AI data centre and a long-term SK hynix memory supply agreement.

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch
Chips & Semiconductors

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch

NVIDIA introduced Jetson T3000 and T2000 modules for edge AI and robotics systems, with Q1 2027 availability planned and source-listed performance, memory and emulation details still tied to company-provided figures.

Keep Reading

More Stories

Latest
Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaCloud & Data CentersAug 8, 2026Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaData Center Dynamics reported that Indosat, Ooredoo Group, Nokia and Nvidia launched Zankore by Indosat with a plan for up to 1GW of AI data centre capacity in Indonesia.Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightAIAug 8, 2026Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightCNBC reported that Black Hat cybersecurity leaders treated the Hugging Face AI-agent breach as a turning point for governing autonomous cyber models rather than a one-off failure.Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAIAug 8, 2026Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAI News reported that Alibaba plans revenue-sharing terms for some commercial users of its next Qwen open-weight AI model, following a licensing pattern already used by Moonshot for Kimi K3.Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanCapital & PolicyAug 8, 2026Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanArs Technica reported that a New Mexico judge ordered Meta to provide $567 million for treatment, screening, awareness and prevention after finding that its platforms contributed to a public nuisance.Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationAIAug 8, 2026Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationSiliconANGLE reported that Harvey AI is seeking at least $500 million in new funding that could value the legal AI startup at $15.5 billion after annualized revenue passed $350 million.Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaScience & TechAug 7, 2026Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaTech Collective SEA wrote that Shopee’s Vietnam share fell from 61% to 53% between May 2025 and April 2026 as TikTok Shop rose from 33% to 44%, showing how social commerce is reshaping regional ecommerce infrastructure.China Opens Security Review Of Palo Alto Networks ProductsCybersecurityAug 7, 2026China Opens Security Review Of Palo Alto Networks ProductsChina's cyberspace regulator opened a security review of Palo Alto Networks products, with no named product line, technical flaw or decision timetable disclosed.AI Pioneers Split Over Risk As Compute Buildout AcceleratesAIAug 7, 2026AI Pioneers Split Over Risk As Compute Buildout AcceleratesData Center Knowledge reported that Geoffrey Hinton, Fei-Fei Li and Andrew Ng disagreed at Ai4 over AI risk, jobs, openness and regulation, leaving infrastructure investors to plan capacity amid unsettled deployment rules.SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportTelco & ConnectivityAug 7, 2026SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportLight Reading reported that SpaceX urged the FCC to sunset High-Cost rural broadband subsidies, while rural telecom and electric-cooperative groups said LEO satellite coverage cannot replace terrestrial network support.OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutAIAug 7, 2026OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutBleepingComputer reported that OpenAI is rolling out GPT-5.6 Sol for paid ChatGPT users and GPT-5.6 Luna for Free and Go users, pairing unlimited free text chats with a new reasoning control and additional safeguards for users believed to be under 18.JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsCapital & PolicyAug 7, 2026JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsData Center Dynamics reported that JLL's EMEA Mid-Year Data Centre Report 2026 put FLAPD live capacity at 3.8GW, while the Middle East had 2.6GW in development paused and 13.8GW in planning.AWS Adds Persistent Runtime Instances For Production AI AgentsCloud & Data CentersAug 7, 2026AWS Adds Persistent Runtime Instances For Production AI AgentsAWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.