SendTech Times
News
DEPLOYMENT WATCH:

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Newsroom brief

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: datacenterknowledge.com
AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Image source: datacenterknowledge.com

AMD has moved its Helios rack-scale AI system into full production, giving the chipmaker an integrated platform to compete with Nvidia beyond individual accelerators.

Data Center Knowledge reported that hardware partners are scheduled to begin shipments by the end of the third quarter of 2026.

The system combines Instinct MI455X accelerators, Epyc server processors, Pensando networking and ROCm software.

AMD said Microsoft and Anthropic are the first named deployment routes for the integrated architecture.

A Rack-Scale Platform Built Around Venice

Helios follows the integrated design now shaping high-end AI infrastructure: processors, accelerators, networking and software are engineered as one rack rather than assembled as separate components.

The approach is intended for increasingly complex agentic workloads that require CPUs to coordinate data retrieval, tool calls and other tasks around GPU reasoning.

At the July 23 Advancing AI event in San Francisco, Chief Executive Lisa Su said AI agents could expand from millions to billions.

Her presentation positioned server processors as an important part of that growth, not simply as hosts for accelerators.

AMD said the Venice family includes configurations with up to 256 cores: SP8 variants span 8 to 128 cores, the SP7 range reaches 256 cores and the LP host-node processor reaches 72 cores.

The lineup covers enterprise servers, cloud systems and specialised AI host nodes.

Performance Claims Centre On The Full System

Su said Helios delivers 15% more AI compute, 50% more memory, 50% more scale-out bandwidth and up to 30% more tokens per dollar than Nvidia's competing platform.

AMD's launch material also compared a 256-core Venice configuration with Nvidia's 88-core Vera CPU, claiming 2.2 times the throughput, and cited a 20% per-core advantage for a 96-core Venice chip.

According to AMD's launch benchmarks, MI455X provides 34 times the token throughput and up to 18 times the token-cost performance of the previous MI355X generation.

Taken together, the vendor's figures present a performance and capacity case for evaluating the rack as an alternative to Nvidia's integrated designs.

Microsoft And Anthropic Establish Deployment Routes

AMD identified Microsoft as a customer for frontier-model inference and Azure AI services.

Su also said Anthropic intends to deploy as much as 2 GW of accelerators in the rack-scale systems.

Those commitments move the launch beyond a component roadmap and support a broader effort to sell CPUs, accelerators, networking and software as one infrastructure stack.

In the same presentation, Su projected a $1.4 trillion AI infrastructure market by 2030, replacing her earlier $500 billion estimate for 2028.

She put the server CPU opportunity at $200 billion by 2030, up from $25 billion today.

The presentation also attributed 46% of latest-quarter server-market revenue to the processor family and listed deployments at more than 60% of Fortune 100 companies.

The Processor Roadmap Runs Through 2030

AMD said its roadmap schedules the SP7 version for the fourth quarter of 2026, SP8 for the first half of 2027 and the LP host-node processor for the second half of that year.

It places seventh-generation Florence CPUs in 2028 and an eighth generation in 2030.

On the accelerator side, the roadmap calls for a new GPU and Helios generation each year, including Instinct MI500 in 2027 and MI600 in 2028.

ROCm.AI is intended to support that cadence with command-line tools, reusable skills and AI-assisted software optimisation across the hardware portfolio.

The remaining evidence gap is operational.

Public material establishes production status, shipment timing, product specifications and customer commitments, but it does not provide order sizes, deployed-cluster economics or independently reproduced benchmarks.

Share this article
inXf

Related articles

More
US Export-Control Shift Opens UAE AI Chip Access, DCD Reports
Chips & Semiconductors

US Export-Control Shift Opens UAE AI Chip Access, DCD Reports

Data Center Dynamics reported that the US government has moved the UAE into a lower-restriction export-control category, giving UAE companies broader access to advanced AI chips from Nvidia and AMD. The account names G42 and hyperscaler data-centre projects, but it does not list chip volumes, approved licences or specific UAE projects.

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs
Chips & Semiconductors

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs

The Register reported that Alibaba launched Qwen 3.8-Max as a downloadable open-weight model while DeepSeek refreshed V4 Flash, sharpening the price and performance challenge facing US AI labs.

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace
Chips & Semiconductors

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace

SiliconANGLE reported that Compute Exchange has launched a marketplace for used and refurbished Nvidia H100 and A100 GPUs. Compute Exchange said requests range from hundreds to tens of thousands of GPUs, but the public launch did not name suppliers, customers or pricing benchmarks.

AMD’s EPYC 9006 Reset Pushes Venice Toward Denser AI Racks
Chips & Semiconductors

AMD’s EPYC 9006 Reset Pushes Venice Toward Denser AI Racks

AMD’s forthcoming EPYC 9006 Venice family adds Zen 6 cores, higher memory bandwidth, PCIe 6, CXL 3.1 and cache-coherent CPU-GPU links for AI-oriented servers.

Oxmiq Raises $35 Million For GPU IP And AI Factory Design
Chips & Semiconductors

Oxmiq Raises $35 Million For GPU IP And AI Factory Design

Oxmiq Labs raised $35 million in Series A funding, taking total funding to $60 million, as the startup expands from GPU IP toward data-centre-scale hardware, orchestration software and AI factory design, EE Times reported. Other large customers, production contracts and independent benchmarks remain outside the public record.

Nvidia's 70% Revenue Forecast Meets Memory Supply Constraint
Chips & Semiconductors

Nvidia's 70% Revenue Forecast Meets Memory Supply Constraint

Nvidia forecast 70 per cent revenue growth for the fiscal year ending January 2028 while warning that memory shortages, component costs and China uncertainty still limit the AI compute boom.

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size
Chips & Semiconductors

Microsoft Adds AMD Helios AI Racks To Azure Without Order Size

Microsoft will deploy AMD Helios rack-scale AI accelerators for Azure AI workloads, with watts, dollars and rack counts still absent from the public terms of the commitment.

Maybank Sees AI Demand Driving Singapore Semiconductor Growth
Chips & Semiconductors

Maybank Sees AI Demand Driving Singapore Semiconductor Growth

Maybank Investment Bank said AI-related demand should remain the main growth driver for Singapore semiconductor companies, while warning that large AI infrastructure spending still needs stronger revenue proof.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.