SendTech Times
News
DEPLOYMENT WATCH:

Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference

Newsroom brief

Cerebras chief executive Andrew Feldman said the company plans to scale manufacturing capacity by 8x to 10x this year and claimed its systems can run inference 10, 15, 20 or 30 times faster than GPUs. The interview-led source named customers including AlphaSense, Cognition AI, OpenAI, Block and GlaxoSmithKline, The public record still lacks third-party benchmark methodology.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: SiliconANGLE
Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference
Image source: SiliconANGLE

Cerebras is making inference speed the centre of its AI compute pitch, with chief executive Andrew Feldman saying the company plans to scale manufacturing capacity by 8x to 10x this year as agentic AI workloads increase demand for fast model responses.

The public record still lacks independently audited benchmark results for the performance claims.

Cerebras Plans 8x To 10x Manufacturing Capacity Growth

Feldman said Cerebras plans to scale manufacturing capacity by eight to 10 times this year.

He linked that expansion to data-centre buildout, next-generation chip and system design, and growing demand from customers using inference for agents and coding workflows.

The company is positioning inference as a hardware bottleneck as AI systems make more sequential calls and reason over longer context windows.

Feldman said faster inference can let customers search across more documents and run more reasoning iterations before returning an answer.

The article named AlphaSense as a customer using speed to search over more documents.

It also said customers include Cognition AI, OpenAI in coding flows, Block for financial agents, GlaxoSmithKline for enterprise deployments and European high-performance computing centres.

Feldman Claims 10, 15, 20, 30 Times Faster Inference

Feldman said Cerebras is "10, 15, 20, 30 times faster than GPUs" for inference.

That is a vendor performance claim from the company's chief executive; independent benchmark methodology, test configuration, model details and third-party validation remain outside the public record.

The technical argument centres on Cerebras's wafer-scale architecture.

Feldman said the design keeps model weights in on-chip SRAM, which the company presents as a way to avoid memory constraints that slow conventional GPU systems.

RAISE Summit 2026 Interview Names Customers But Not Orders

Feldman spoke with John Furrier and Dave Vellante at the RAISE Summit.

The discussion covered Cerebras's IPO milestone, inference speed, European data-centre buildout and manufacturing scale, The public record still lacks IPO timing, production-site names, wafer supply commitments or signed order totals.

The article disclosed that theCUBE was a paid media partner for the RAISE Summit event and said sponsors did not have editorial control over the content.

The public record lists the source useful for company strategy and named-customer claims, but not enough to treat the speed comparison as independently verified.

Independent benchmark methodology, exact manufacturing partners, named manufacturing sites, signed order volumes and a timetable for the IPO milestone remain outside the public record.

Share this article
inXf

Related articles

More
Imec Says AI Inference Pushes Optical I/O Closer To Chips
Chips & Semiconductors

Imec Says AI Inference Pushes Optical I/O Closer To Chips

Imec researchers told EE Times that AI inference is shifting the interconnect problem from rack-scale optics towards 2.5D and 3D optical I/O near processors. The group discussed a 250 Tb/s bandwidth projection while the public record still lacks commercial product timing, customer deployments or cooling test results.

IBM Shows Sub-1 Nanometer Chip Research, With Production Proof Still Missing
Chips & Semiconductors

IBM Shows Sub-1 Nanometer Chip Research, With Production Proof Still Missing

IBM unveiled a 0.7 nm nanostack chip technology with nearly 100 billion transistors, but the company has not announced a manufacturing partner, production date or customer design win.

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity
Chips & Semiconductors

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity

Tom’s Hardware put Sandisk and SK hynix’s OCP High Bandwidth Flash specification at up to 512GB per package, with bandwidth grades reaching 3.0 TB/s for AI inference systems.

FuriosaAI Starts RNGD Accelerator Deployment At Equinix Lisbon Datacenter
Chips & Semiconductors

FuriosaAI Starts RNGD Accelerator Deployment At Equinix Lisbon Datacenter

FuriosaAI has begun deploying its RNGD AI accelerators at Equinix’s LS2 datacenter in Lisbon as the South Korean chip startup looks for European sovereign AI demand. The company describes 48 GB of HBM3, 1.5 TB/s memory bandwidth and 512 teraFLOPS of dense FP8 performance per card, but its Broadcom-linked next-generation accelerator still depends on HBM4 and HBM4e timing.

Groq Raises $650 Million As Its AI Chip Story Turns Toward Neocloud Scale
Chips & Semiconductors

Groq Raises $650 Million As Its AI Chip Story Turns Toward Neocloud Scale

Groq announced a $650 million round after Nvidia licensed its technology and hired senior leaders, leaving the AI chip company to rebuild around a neocloud business spanning 13 data centers.

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Chips & Semiconductors

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

Firebird Plans 70,000 GPUs For Armenia AI Factory By 2027
Chips & Semiconductors

Firebird Plans 70,000 GPUs For Armenia AI Factory By 2027

NVIDIA's newsroom detailed Firebird's Armenia AI factory launch, including plans for more than 70,000 Rubin and Blackwell GPUs and 300 megawatts of AI infrastructure capacity by 2027.

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch
Chips & Semiconductors

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch

NVIDIA introduced Jetson T3000 and T2000 modules for edge AI and robotics systems, with Q1 2027 availability planned and source-listed performance, memory and emulation details still tied to company-provided figures.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.