News
AI SHIFT:

Thinking Machines Releases Inkling Open-Weights Model With 975 Billion Parameters

Newsroom brief

SiliconANGLE reported that Thinking Machines Lab has released Inkling, its first foundation model, with full open weights and fine-tuning through Tinker. The account cited 975 billion total parameters, about 41 billion active parameters per average prompt and training on about 45 trillion tokens, while leaving customer deployments and independent benchmark validation undisclosed.

Verified against source materialEdited by SendTech Times AI & Enterprise DeskSource: SiliconANGLE
Thinking Machines Releases Inkling Open-Weights Model With 975 Billion Parameters
Image source: SiliconANGLE

Thinking Machines Lab has released its first foundation model, Inkling, with full open weights and developer fine-tuning through Tinker, SiliconANGLE reported.

The launch gives the AI startup a public model after a year in which its funding rounds and Nvidia partnership drew most of the attention.

Inkling is not being presented as a closed chatbot.

Developers can download, adjust and run the weights, while Tinker remains the paid service for fine-tuning open-weights models.

Inkling Uses 975 Billion Parameters And Open Weights

A company blog post describes Inkling as a mixture-of-experts model with 975 billion parameters.

Thinking Machines said an average prompt draws on about 41 billion parameters to process tasks faster and keep costs low.

The training description lists about 45 trillion tokens spanning text, image, audio and video.

Inkling can reason across all four inputs, but its outputs are limited to text, including code, styled artifacts and structured data.

Those full open weights let developers inspect and adapt the model code.

Thinking Machines also outlined thinking-effort controls for trading processing speed against accuracy, and SiliconANGLE wrote that the model flags uncertainty in outputs.

Mira Murati previously served as Chief Technology Officer of OpenAI before leaving in September 2024, according to the account.

Her stated focus on accessibility, customisation and multimodal collaboration appears in the launch, with public outputs still limited to text.

Tinker Carries The Fine-Tuning Revenue Model

Developers can fine-tune the model directly on Tinker, the startup's training API that launched in October, according to SiliconANGLE.

Rather than charging for metered access to the model itself, the paid API carries the revenue plan for the release.

The training path also extends the Nvidia connection.

Thinking Machines stated that Inkling was trained on Nvidia's GB300 NVL72 system under a partnership announced in March.

In early test results cited by the company, Inkling reached comparable coding performance with Nvidia's Nemotron 3 Ultra while using two-thirds fewer tokens, the startup claimed.

Independent benchmark methodology and customer deployment results are not included in that comparison.

Bridgewater Test Gives Inkling A Finance Example

SiliconANGLE cited a collaboration with Bridgewater Associates in which researchers used Tinker to fine-tune an open model with specialised financial data.

The resulting lightweight model scored 84.7% on financial reasoning benchmarks at less than 10% of the cost of advanced proprietary alternatives, the account said.

Futurum Group analyst Mitch Ashely told the Wall Street Journal, as cited in the account, that the open-weight model ecosystem had been dominated by Chinese AI firms for the last year.

The quoted assessment described the release as a Western alternative for enterprises weighing customisation economics and infrastructure control.

The lab acknowledged that its new model is not as strong as some advanced proprietary AI systems.

The release is positioned as a base model that organisations can fine-tune and run on their own infrastructure, not as a rigid chatbot application.

Thinking Machines developed the model from scratch in less than nine months, according to the account.

Thinking Machines did not provide independent benchmark validation for Inkling.

Share this article
inXf

Related articles

More
Fireworks Raises $1.5B As AI Platform Reaches $17.5B Valuation
AI

Fireworks Raises $1.5B As AI Platform Reaches $17.5B Valuation

Fireworks AI raised a $1.5 billion Series D at a $17.5 billion valuation, SiliconANGLE reported. The company says it processes more than 40 trillion tokens a day, while contract values, total GPU capacity and customer commitments tied to the new capital remain outside the public record.

NVIDIA Lists Nemotron Enterprise AI Use Cases Without Contract Data
AI

NVIDIA Lists Nemotron Enterprise AI Use Cases Without Contract Data

NVIDIA said its Nemotron open models are being customised by enterprise and national AI builders, with examples across clinical documentation, legal work, enterprise search and Malaysian-language AI. The company cited partner benchmark and cost claims, while contract values, deployment volumes and independent benchmark audits remain outside the public account.

Nvidia Opens Medical Simulation Framework For Healthcare Robot Training
AI

Nvidia Opens Medical Simulation Framework For Healthcare Robot Training

Nvidia's open-source Medical Physics Simulation framework is designed to generate training environments for healthcare robotics, while the named adopter list does not include a patient-side deployment.

Cadence Adds AuraStack AI Agent For PCB And Advanced Packaging Design
Chips & Semiconductors

Cadence Adds AuraStack AI Agent For PCB And Advanced Packaging Design

The Register reported that Cadence Design Systems introduced AuraStack, an agentic AI system for PCB and advanced packaging workflows. Cadence cited a 15x productivity claim and named Nvidia among customers, but The Register did not include pricing, availability dates, full customer names or independent benchmark results.

IMEC Outlines CMOS 2.0 Path As AI Compute Demand Rises
AI

IMEC Outlines CMOS 2.0 Path As AI Compute Demand Rises

CommonWealth Magazine English reported that IMEC's new CEO Patrick Vandenameele outlined semiconductor roadmap work tied to AI inference demand, CMOS 2.0 stacking, memory placement and optical interconnects. CommonWealth reported that Vandenameele estimated a 150-fold workload increase as AI shifts from training to inference, while TSMC's Kevin Zhang pointed to nanosheet, CFET and 3D stacking work.

H2LooP Raises $2m For Embedded-System AI Coding Tools
AI

H2LooP Raises $2m For Embedded-System AI Coding Tools

YourStory reported that Bengaluru startup H2LooP raised $2 million to build hardware-aware AI coding tools for firmware and embedded systems, with early deployments in semiconductor and automotive teams.

AI Memory Demand Pushes India Smartphone Shipments Down 10%
Chips & Semiconductors

AI Memory Demand Pushes India Smartphone Shipments Down 10%

TechCrunch reported that India smartphone shipments fell 10% year over year in the April-June quarter as AI data centre demand pulled memory suppliers toward high-bandwidth memory. The story cited Counterpoint Research and IDC, while customer-level component contracts and supplier allocation data remain undisclosed.

TCS Opens Bengaluru Industrial AI Lab With Nvidia Infrastructure
AI

TCS Opens Bengaluru Industrial AI Lab With Nvidia Infrastructure

Tech Monitor reported that Tata Consultancy Services opened an Autonomous Engineering Lab powered by Nvidia at its Global Axis campus in Bengaluru. The lab targets industrial AI prototypes for mobility and manufacturing customers, while TCS did not name customers, budgets, deployment dates or production results.

Keep Reading

More Stories

Latest
Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaCloud & Data CentersAug 8, 2026Indosat AI Data Centre Plan Targets 1GW With Ooredoo, Nokia And NvidiaData Center Dynamics reported that Indosat, Ooredoo Group, Nokia and Nvidia launched Zankore by Indosat with a plan for up to 1GW of AI data centre capacity in Indonesia.Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightAIAug 8, 2026Hugging Face Hack Pushes AI Agents Into Cybersecurity SpotlightCNBC reported that Black Hat cybersecurity leaders treated the Hugging Face AI-agent breach as a turning point for governing autonomous cyber models rather than a one-off failure.Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAIAug 8, 2026Alibaba Tests Revenue Sharing For Commercial Qwen AI UseAI News reported that Alibaba plans revenue-sharing terms for some commercial users of its next Qwen open-weight AI model, following a licensing pattern already used by Moonshot for Kimi K3.Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanCapital & PolicyAug 8, 2026Meta Ordered To Fund $567M New Mexico Youth Mental Health PlanArs Technica reported that a New Mexico judge ordered Meta to provide $567 million for treatment, screening, awareness and prevention after finding that its platforms contributed to a public nuisance.Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationAIAug 8, 2026Harvey Funding Talks Could Lift Legal AI Startup To $15.5B ValuationSiliconANGLE reported that Harvey AI is seeking at least $500 million in new funding that could value the legal AI startup at $15.5 billion after annualized revenue passed $350 million.Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaScience & TechAug 7, 2026Vietnam Shows Shopee-TikTok Shop Race Tightening In Southeast AsiaTech Collective SEA wrote that Shopee’s Vietnam share fell from 61% to 53% between May 2025 and April 2026 as TikTok Shop rose from 33% to 44%, showing how social commerce is reshaping regional ecommerce infrastructure.China Opens Security Review Of Palo Alto Networks ProductsCybersecurityAug 7, 2026China Opens Security Review Of Palo Alto Networks ProductsChina's cyberspace regulator opened a security review of Palo Alto Networks products, with no named product line, technical flaw or decision timetable disclosed.AI Pioneers Split Over Risk As Compute Buildout AcceleratesAIAug 7, 2026AI Pioneers Split Over Risk As Compute Buildout AcceleratesData Center Knowledge reported that Geoffrey Hinton, Fei-Fei Li and Andrew Ng disagreed at Ai4 over AI risk, jobs, openness and regulation, leaving infrastructure investors to plan capacity amid unsettled deployment rules.SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportTelco & ConnectivityAug 7, 2026SpaceX Asks FCC To Wind Down $4.5bn Rural Broadband SupportLight Reading reported that SpaceX urged the FCC to sunset High-Cost rural broadband subsidies, while rural telecom and electric-cooperative groups said LEO satellite coverage cannot replace terrestrial network support.OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutAIAug 7, 2026OpenAI Expands Free ChatGPT Access In GPT-5.6 RolloutBleepingComputer reported that OpenAI is rolling out GPT-5.6 Sol for paid ChatGPT users and GPT-5.6 Luna for Free and Go users, pairing unlimited free text chats with a new reasoning control and additional safeguards for users believed to be under 18.JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsCapital & PolicyAug 7, 2026JLL Data Centre Report Shows Middle East Pipeline Pause As FLAPD GrowsData Center Dynamics reported that JLL's EMEA Mid-Year Data Centre Report 2026 put FLAPD live capacity at 3.8GW, while the Middle East had 2.6GW in development paused and 13.8GW in planning.AWS Adds Persistent Runtime Instances For Production AI AgentsCloud & Data CentersAug 7, 2026AWS Adds Persistent Runtime Instances For Production AI AgentsAWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.