SendTech Times
News
SUPPLY CHECK:

Qualcomm AI250 Stacks DRAM Over Compute But Leaves FLOPS Undisclosed

Newsroom brief

Qualcomm is pitching high-bandwidth compute for AI inference, with AI250 cards claiming 768 GB of memory and 133 TB/s of effective bandwidth, while the public record still lacks peak FLOPS or named customers.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: The Register
Qualcomm AI250 Stacks DRAM Over Compute But Leaves FLOPS Undisclosed
Image source: The Register

Qualcomm AI250 Stacks DRAM Above Logic For Inference

Qualcomm is using its AI250 accelerator roadmap to push a different answer to the AI inference memory bottleneck.

The company describes high-bandwidth compute, or HBC, as a 3D-stacked design that places DRAM above logic so some work can happen closer to memory.

The AI250 is due to follow the AI200 Dragonfly rack systems and is planned to begin shipping in 2027.

The Register reported that Qualcomm also outlined a second-generation HBC platform, the AI300, for 2028.

Qualcomm says the AI250 card will carry 768 GB of memory and up to 133 TB/s of effective memory bandwidth.

The company ties those claims to bandwidth-bound inference work, especially decode, where model weights are streamed from memory during token generation.

Effective Bandwidth Claims Need More Detail

The company is presenting HBC as a way to reduce data movement between memory and compute.

Qualcomm says the architecture uses LPDDR memory in a purpose-built near-memory design and differs from HBM because HBC does computing in the base logic die.

The bandwidth claims still depend on Qualcomm's definition of effective bandwidth.

For the AI200 generation, Qualcomm had cited 414 TB/s of effective memory bandwidth across 56 chips.

The AI250 marketing material says HBC gives 18x the AI200's effective bandwidth, while the AI300 would reach 54x.

Qualcomm says the AI250 can operate as a standalone AI accelerator.

It also says the part can sit in disaggregated inference systems, with GPUs or other Qualcomm parts handling prompt processing and AI250 accelerators handling memory-intensive decode.

The company declined to give peak FLOPS for AI250.

It also did not give the detailed physical bandwidth calculation behind the headline effective-bandwidth figures, even as the disclosed figures indicate that ordinary LPDDR5x bandwidth would not explain the claimed totals by itself.

Modular Deal Targets The Software Gap

Qualcomm's investor-day push also included its planned acquisition of Modular, the AI software startup behind Mojo and the Max serving platform.

Mojo is positioned as a low-level programming interface that can run across different hardware, while Max targets LLM model serving.

AI accelerator buyers are comparing more than silicon specifications.

They need serving tools, developer support and deployment paths that do not lock every workload to one vendor stack.

Qualcomm is using Modular to address that software gap while Nvidia and AMD remain the main comparison points for AI infrastructure buyers.

The plan also assumes Qualcomm can make a heterogeneous inference model attractive.

Production deployments using that design remain outside the public record.

The remaining public gaps are peak FLOPS for AI250, the detailed method behind its effective bandwidth calculation, named AI250 customers, production deployment dates beyond the 2027 target or whether regulators will clear the Modular acquisition this year, according to the source material.

Share this article
inXf

Related articles

More
Nvidia Revenue-Share Model Names 210,000 GPUs But Leaves Split Undisclosed
Chips & Semiconductors

Nvidia Revenue-Share Model Names 210,000 GPUs But Leaves Split Undisclosed

Nvidia is offering a revenue-sharing and credit-support model for AI cloud partners that use its infrastructure. Sharon AI and Firmus Technologies are named partners with up to 210,000 GPUs combined, while the public record still lacks the revenue-split percentages.

Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference
Chips & Semiconductors

Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference

Cerebras chief executive Andrew Feldman said the company plans to scale manufacturing capacity by 8x to 10x this year and claimed its systems can run inference 10, 15, 20 or 30 times faster than GPUs. The interview-led source named customers including AlphaSense, Cognition AI, OpenAI, Block and GlaxoSmithKline, The public record still lacks third-party benchmark methodology.

Qualcomm Wins Meta CPU Agreement, But Production Waits Until 2028
Chips & Semiconductors

Qualcomm Wins Meta CPU Agreement, But Production Waits Until 2028

Qualcomm Technologies said it will supply data centre CPUs for Meta under a multi-generation agreement, with the first Dragonfly C1000 production scheduled for the second half of 2028 and capacity terms still undisclosed.

Qualcomm Names Meta As First Dragonfly Data Center CPU Customer
Chips & Semiconductors

Qualcomm Names Meta As First Dragonfly Data Center CPU Customer

Qualcomm said Meta will use its Dragonfly C1000 data center CPU when production starts in 2028, while the chipmaker raised its fiscal 2029 non-handset revenue projection to $40 billion.

India AI Compute Buyers Face 36-To-52-Week GPU Lead Times
Chips & Semiconductors

India AI Compute Buyers Face 36-To-52-Week GPU Lead Times

Indian AI infrastructure buyers are still reserving GPU capacity months ahead. NeevCloud cofounder Narendra Sen said next-generation enterprise AI GPU lead times now range between 36 and 52 weeks, while named Indian startups are shifting training and inference around scarce capacity.

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push
Chips & Semiconductors

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push

Rest of World reported that Saudi Arabia and the UAE are trying to diversify AI chip supply while major projects still rely on Nvidia hardware. The report cited Humain data-centre plans, G42’s Stargate project and analyst warnings that U.S. approvals, TSMC capacity and high-bandwidth memory remain constraints.

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity
Chips & Semiconductors

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity

Tom’s Hardware put Sandisk and SK hynix’s OCP High Bandwidth Flash specification at up to 512GB per package, with bandwidth grades reaching 3.0 TB/s for AI inference systems.

SK hynix Memristor AI Chip Shows 21.3 TOPS/W But Leaves Throughput Gap
Chips & Semiconductors

SK hynix Memristor AI Chip Shows 21.3 TOPS/W But Leaves Throughput Gap

SK hynix, TetraMem and USC researchers developed a 65 nm memristor-based in-memory computing chip for edge AI. The paper listed 21.3 TOPS/W at 100 MHz, while the public record still lacks full-chip saturated throughput.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.