SendTech Times
Analysis
DEPLOYMENT WATCH:

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs

Newsroom brief

The Register reported that Alibaba launched Qwen 3.8-Max as a downloadable open-weight model while DeepSeek refreshed V4 Flash, sharpening the price and performance challenge facing US AI labs.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: The Register
Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs
Image source: The Register

Chinese AI developers are turning open weights into a pricing and deployment challenge for US model labs.

The Register reported that Alibaba launched Qwen 3.8-Max, a 2.4 trillion-parameter model, while DeepSeek's V4 Flash refresh arrived with a smaller footprint and lower task cost.

The movement is not only a benchmark contest.

For enterprise buyers, the practical question is whether capable models can be downloaded, fine-tuned and run on owned infrastructure instead of remaining tied to proprietary cloud APIs.

The Register framed Alibaba, DeepSeek, Moonshot, MiniMax and Z.ai as the Chinese developers pushing hardest into that opening.

Alibaba Sets Qwen Weight Release

Alibaba had kept its strongest models behind an API.

Qwen 3.8-Max changes that by putting the company's most capable model weights on a release path for popular repositories, including Hugging Face, starting next week.

The model is large enough to keep deployment difficult.

The report said Qwen 3.8-Max has 2.4 trillion total parameters and uses 95 billion active parameters for a request.

Its hardware estimate puts customer-facing deployment at 48 to 64 Nvidia B200-class GPUs, while internal workloads would still need 8 to 16 B300 or AMD MI355X GPUs.

Alibaba also plans a 27-billion-parameter version alongside the Max model.

A smaller weight release gives enterprise teams a more realistic path for budgeting, hardware planning and fine-tuning workflows.

DeepSeek Prices V4 Flash

DeepSeek's V4 Flash refresh attacks the same market from the opposite direction.

DeepSeek V4-Flash-0731 is a 284-billion-parameter model that The report said can fit into about 142 GB of GPU memory at FP4, making large-scale local operation possible on a single system.

Artificial Analysis benchmarks, as cited by The Register, put DeepSeek V4 Flash close to OpenAI's GPT 5.6 Luna and assigned DeepSeek a 40 percent task-cost advantage.

DeepSeek API pricing was listed at $0.14 for each million input tokens and $0.28 for each million output tokens; Alibaba's QwenCloud listing for Qwen 3.8-Max was $2 for each million input tokens and $6 for each million output tokens.

The cost comparison is not limited to posted token prices.

Reasoning and agent workloads can consume different token volumes, so model efficiency matters when developers run large coding or automation jobs.

DeepSeek's integration of DSpark speculative decoding is presented as one reason for the efficiency claim, with DeepSeek saying it can deliver 57 to 85 percent more per-user speed on the same hardware.

Buyers Compare Closed APIs With Weights

US and European AI companies have warned about Chinese open models, safety standards and model provenance.

Hugging Face CEO Clément Delangue said on CNBC that Chinese developers are clearly dominating open models and could start dominating frontier models by the end of this year or next year if the pace continues.

That leaves US labs with a narrower response than broad safety rhetoric.

Enterprise customers weighing proprietary APIs against open models will compare control, auditability, data handling, price and available hardware.

If Chinese models keep closing benchmark gaps while offering downloadable weights, the competitive question becomes how much extra trust or capability customers get from closed systems.

Production use will decide how much of the launch economics survives outside charts.

Alibaba's highest-end model still requires costly infrastructure, and DeepSeek's lower-cost path depends on whether compressed or quantized deployments preserve output quality under real workloads.

Share this article
inXf

Related articles

More
AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Chips & Semiconductors

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

CXMT’s 4.13 Trillion Yuan Rally Puts AI Memory Supply In Focus
Chips & Semiconductors

CXMT’s 4.13 Trillion Yuan Rally Puts AI Memory Supply In Focus

South China Morning Post reported that CXMT’s market value reached 4.13 trillion yuan after a record close, as investors priced in tight memory supply tied to AI data-centre demand.

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace
Chips & Semiconductors

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace

SiliconANGLE reported that Compute Exchange has launched a marketplace for used and refurbished Nvidia H100 and A100 GPUs. Compute Exchange said requests range from hundreds to tens of thousands of GPUs, but the public launch did not name suppliers, customers or pricing benchmarks.

Kioxia GP1 SSD Demo Pushes PCIe Gen6 Storage To 10 Million IOPS
Chips & Semiconductors

Kioxia GP1 SSD Demo Pushes PCIe Gen6 Storage To 10 Million IOPS

ServeTheHome showed Kioxia’s GP1 SSD running just above 10 million IOPS at FMS 2026, using PCIe Gen6 and second-generation XL-FLASH for a fast server storage tier.

Oxmiq Raises $35 Million For GPU IP And AI Factory Design
Chips & Semiconductors

Oxmiq Raises $35 Million For GPU IP And AI Factory Design

Oxmiq Labs raised $35 million in Series A funding, taking total funding to $60 million, as the startup expands from GPU IP toward data-centre-scale hardware, orchestration software and AI factory design, EE Times reported. Other large customers, production contracts and independent benchmarks remain outside the public record.

T-Head V900 Pushes Alibaba AI Chips Toward Full-Stack Infrastructure
Chips & Semiconductors

T-Head V900 Pushes Alibaba AI Chips Toward Full-Stack Infrastructure

TechNode reported that Alibaba chip subsidiary T-Head unveiled the Zhenwu V900 AI chip, a supernode server design and a Yitian CPU roadmap aimed at larger AI training and inference systems.

Astera Labs Pushes Leo Memory Controllers Into DDR4 Reuse And AI Fabrics
Chips & Semiconductors

Astera Labs Pushes Leo Memory Controllers Into DDR4 Reuse And AI Fabrics

Astera Labs has launched Leo 2 and Leo X CXL memory controllers, pairing PCIe Gen6/CXL 3.2 support with DDR4 reuse options and a fabric-attached design for AI accelerator memory pools.

Nvidia's 70% Revenue Forecast Meets Memory Supply Constraint
Chips & Semiconductors

Nvidia's 70% Revenue Forecast Meets Memory Supply Constraint

Nvidia forecast 70 per cent revenue growth for the fiscal year ending January 2028 while warning that memory shortages, component costs and China uncertainty still limit the AI compute boom.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.