Analysis
CAPACITY TEST:

Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs

Newsroom brief

The Register reported that Alibaba launched Qwen 3.8-Max as a downloadable open-weight model while DeepSeek refreshed V4 Flash, sharpening the price and performance challenge facing US AI labs.

Verified against source materialEdited by SendTech Times Chips & Compute DeskSource: The Register
Alibaba Qwen Release Puts Open-Model Pressure Back On US AI Labs
Image source: The Register

Chinese AI developers are turning open weights into a pricing and deployment challenge for US model labs.

The Register reported that Alibaba launched Qwen 3.8-Max, a 2.4 trillion-parameter model, while DeepSeek's V4 Flash refresh arrived with a smaller footprint and lower task cost.

The movement is not only a benchmark contest.

For enterprise buyers, the practical question is whether capable models can be downloaded, fine-tuned and run on owned infrastructure instead of remaining tied to proprietary cloud APIs.

The Register framed Alibaba, DeepSeek, Moonshot, MiniMax and Z.ai as the Chinese developers pushing hardest into that opening.

Alibaba Sets Qwen Weight Release

Alibaba had kept its strongest models behind an API.

Qwen 3.8-Max changes that by putting the company's most capable model weights on a release path for popular repositories, including Hugging Face, starting next week.

The model is large enough to keep deployment difficult.

The report said Qwen 3.8-Max has 2.4 trillion total parameters and uses 95 billion active parameters for a request.

Its hardware estimate puts customer-facing deployment at 48 to 64 Nvidia B200-class GPUs, while internal workloads would still need 8 to 16 B300 or AMD MI355X GPUs.

Alibaba also plans a 27-billion-parameter version alongside the Max model.

A smaller weight release gives enterprise teams a more realistic path for budgeting, hardware planning and fine-tuning workflows.

DeepSeek Prices V4 Flash

DeepSeek's V4 Flash refresh attacks the same market from the opposite direction.

DeepSeek V4-Flash-0731 is a 284-billion-parameter model that The report said can fit into about 142 GB of GPU memory at FP4, making large-scale local operation possible on a single system.

Artificial Analysis benchmarks, as cited by The Register, put DeepSeek V4 Flash close to OpenAI's GPT 5.6 Luna and assigned DeepSeek a 40 percent task-cost advantage.

DeepSeek API pricing was listed at $0.14 for each million input tokens and $0.28 for each million output tokens; Alibaba's QwenCloud listing for Qwen 3.8-Max was $2 for each million input tokens and $6 for each million output tokens.

The cost comparison is not limited to posted token prices.

Reasoning and agent workloads can consume different token volumes, so model efficiency matters when developers run large coding or automation jobs.

DeepSeek's integration of DSpark speculative decoding is presented as one reason for the efficiency claim, with DeepSeek saying it can deliver 57 to 85 percent more per-user speed on the same hardware.

Buyers Compare Closed APIs With Weights

US and European AI companies have warned about Chinese open models, safety standards and model provenance.

Hugging Face CEO Clément Delangue said on CNBC that Chinese developers are clearly dominating open models and could start dominating frontier models by the end of this year or next year if the pace continues.

That leaves US labs with a narrower response than broad safety rhetoric.

Enterprise customers weighing proprietary APIs against open models will compare control, auditability, data handling, price and available hardware.

If Chinese models keep closing benchmark gaps while offering downloadable weights, the competitive question becomes how much extra trust or capability customers get from closed systems.

Production use will decide how much of the launch economics survives outside charts.

Alibaba's highest-end model still requires costly infrastructure, and DeepSeek's lower-cost path depends on whether compressed or quantized deployments preserve output quality under real workloads.

Share this article
inXf

Related articles

More
AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge
Chips & Semiconductors

AMD Helios AI Racks Bring Epyc CPUs To Nvidia Compute Challenge

Data Center Knowledge reported that AMD moved Helios into production with MI455X GPUs, Epyc processors, Pensando networking and ROCm software, while vendor performance claims still lack third-party benchmark validation.

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace
Chips & Semiconductors

Compute Exchange Opens Used Nvidia H100 And A100 GPU Marketplace

SiliconANGLE reported that Compute Exchange has launched a marketplace for used and refurbished Nvidia H100 and A100 GPUs. Compute Exchange said requests range from hundreds to tens of thousands of GPUs, but the public launch did not name suppliers, customers or pricing benchmarks.

US Export-Control Shift Opens UAE AI Chip Access, DCD Reports
Chips & Semiconductors

US Export-Control Shift Opens UAE AI Chip Access, DCD Reports

Data Center Dynamics reported that the US government has moved the UAE into a lower-restriction export-control category, giving UAE companies broader access to advanced AI chips from Nvidia and AMD. The account names G42 and hyperscaler data-centre projects, but it does not list chip volumes, approved licences or specific UAE projects.

Maybank Sees AI Demand Driving Singapore Semiconductor Growth
Chips & Semiconductors

Maybank Sees AI Demand Driving Singapore Semiconductor Growth

Maybank Investment Bank said AI-related demand should remain the main growth driver for Singapore semiconductor companies, while warning that large AI infrastructure spending still needs stronger revenue proof.

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity
Chips & Semiconductors

Sandisk And SK Hynix Open HBF Spec For AI Memory Capacity

Tom’s Hardware put Sandisk and SK hynix’s OCP High Bandwidth Flash specification at up to 512GB per package, with bandwidth grades reaching 3.0 TB/s for AI inference systems.

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push
Chips & Semiconductors

Gulf AI Plans Still Depend On Nvidia Chips Despite Supplier Push

Rest of World reported that Saudi Arabia and the UAE are trying to diversify AI chip supply while major projects still rely on Nvidia hardware. The report cited Humain data-centre plans, G42’s Stargate project and analyst warnings that U.S. approvals, TSMC capacity and high-bandwidth memory remain constraints.

Nvidia-SK AI Deal Carries $500 Billion Value With Few Build Details
Chips & Semiconductors

Nvidia-SK AI Deal Carries $500 Billion Value With Few Build Details

Tom's Hardware reported that Nvidia and SK Group signed letters of intent for a strategic relationship valued at more than $500 billion, anchored by SK Telecom's planned 2-gigawatt South Korea AI data centre and a long-term SK hynix memory supply agreement.

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch
Chips & Semiconductors

NVIDIA Sets Jetson T3000 And T2000 Modules For Q1 2027 Robotics Launch

NVIDIA introduced Jetson T3000 and T2000 modules for edge AI and robotics systems, with Q1 2027 availability planned and source-listed performance, memory and emulation details still tied to company-provided figures.

Keep Reading

More Stories

Latest
Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaCloud & Data CentersAug 8, 2026Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaData Center Dynamics reported that Indosat, Ooredoo Group, Nokia and Nvidia launched Zankore by Indosat with a plan for up to 1GW of AI data centre capacity in Indonesia.Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightAIAug 8, 2026Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightCNBC reported that Black Hat cybersecurity leaders treated the Hugging Face AI-agent breach as a turning point for governing autonomous cyber models rather than a one-off failure.Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAIAug 8, 2026Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAI News reported that Alibaba plans revenue-sharing terms for some commercial users of its next Qwen open-weight AI model, following a licensing pattern already used by Moonshot for Kimi K3.Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanCapital & PolicyAug 8, 2026Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanArs Technica reported that a New Mexico judge ordered Meta to provide $567 million for treatment, screening, awareness and prevention after finding that its platforms contributed to a public nuisance.Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationAIAug 8, 2026Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationSiliconANGLE reported that Harvey AI is seeking at least $500 million in new funding that could value the legal AI startup at $15.5 billion after annualized revenue passed $350 million.Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaScience & TechAug 7, 2026Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaTech Collective SEA wrote that Shopee’s Vietnam share fell from 61% to 53% between May 2025 and April 2026 as TikTok Shop rose from 33% to 44%, showing how social commerce is reshaping regional ecommerce infrastructure.China Opens Security Review Of Palo Alto Networks ProductsCybersecurityAug 7, 2026China Opens Security Review Of Palo Alto Networks ProductsChina's cyberspace regulator opened a security review of Palo Alto Networks products, with no named product line, technical flaw or decision timetable disclosed.AI Pioneers Split Over Risk As Compute Buildout AcceleratesAIAug 7, 2026AI Pioneers Split Over Risk As Compute Buildout AcceleratesData Center Knowledge reported that Geoffrey Hinton, Fei-Fei Li and Andrew Ng disagreed at Ai4 over AI risk, jobs, openness and regulation, leaving infrastructure investors to plan capacity amid unsettled deployment rules.SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportTelco & ConnectivityAug 7, 2026SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportLight Reading reported that SpaceX urged the FCC to sunset High-Cost rural broadband subsidies, while rural telecom and electric-cooperative groups said LEO satellite coverage cannot replace terrestrial network support.OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutAIAug 7, 2026OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutBleepingComputer reported that OpenAI is rolling out GPT-5.6 Sol for paid ChatGPT users and GPT-5.6 Luna for Free and Go users, pairing unlimited free text chats with a new reasoning control and additional safeguards for users believed to be under 18.JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsCapital & PolicyAug 7, 2026JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsData Center Dynamics reported that JLL's EMEA Mid-Year Data Centre Report 2026 put FLAPD live capacity at 3.8GW, while the Middle East had 2.6GW in development paused and 13.8GW in planning.AWS Adds Persistent Runtime Instances For Production AI AgentsCloud & Data CentersAug 7, 2026AWS Adds Persistent Runtime Instances For Production AI AgentsAWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.