SendTech Times
News
DEPLOYMENT WATCH:

Crusoe Adds Serverless Fine-Tuning To AI Infrastructure Platform

Newsroom brief

Crusoe added managed fine-tuning and inference services for open-weight models inside Intelligence Foundry. The launch moves its AI infrastructure pitch beyond rented GPU access, while prices, named customers and verified savings remain outside the public record.

Verified against source materialEdited by SendTech Times Cloud & Infrastructure DeskSource: Data Center Knowledge
Crusoe Adds Serverless Fine-Tuning To AI Infrastructure Platform
Image source: Data Center Knowledge

Crusoe is moving its AI infrastructure pitch beyond rented GPU access by adding managed fine-tuning and inference services for open-weight models, Data Center Knowledge reported.

The new Serverless Fine-Tuning and Self-Serve Deployments services will sit inside Intelligence Foundry.

They are aimed at teams that want to adapt open-source foundation models, deploy managed inference endpoints or export fine-tuned model weights without provisioning GPU clusters directly.

Intelligence Foundry Adds Serverless Fine-Tuning

The fine-tuning service lets customers bring data to open-weight foundation models and receive completed weights in the open . safetensors format.

Customers can deploy those weights on the same platform or move them elsewhere, making portability part of the product claim.

Erwan Menard, senior vice president of product, told Data Center Knowledge that enterprises are moving towards model ownership instead of relying only on proprietary APIs.

AI-native companies are feeding production data back into open-weight models regularly as they try to improve performance and reduce inference costs.

Demand for continuous fine-tuning is accelerating faster than expected, particularly among teams building production AI agents where model predictability and data ownership affect procurement decisions, according to Data Center Knowledge.

The platform currently supports a curated library of open-weight models including Qwen, DeepSeek, Gemma and GPT-OSS.

IDC Points To Competition Beyond GPU Access

Dave McCarthy, research vice president at IDC, said raw GPU access was the dominant story for about 18 months but is no longer enough by itself.

Enterprise buyers are looking at fine-tuning pipelines, evaluation, deployment tooling and inference optimisation as one system.

Providers that only sell chips risk becoming interchangeable.

McCarthy framed the launch as part of a wider shift in AI data centre competition from capacity supply towards full model-lifecycle platforms.

Portability is another procurement issue.

McCarthy said portability is no longer optional for enterprise buyers, while Menard said organisations using open-weight models increasingly expect to keep their fine-tuned weights rather than stay locked to one inference platform.

General Availability Is Scheduled For Next Week

Both services are scheduled for general availability next week through Intelligence Foundry.

The fine-tuning charge will use a per-million-token model, while managed inference will be charged by GPU hour.

The inference service uses Nvidia H100 and H200 GPUs for managed endpoints.

The launch also includes automatic job restarts, checkpoint saving during training and billing that stops when a model stops improving.

Named customers, per-million-token prices, GPU-hour rates, utilisation targets, benchmark methodology, service-level terms and customer-verified cost savings for the new fine-tuning and inference services remain outside the public record.

Share this article
inXf

Related articles

More
e& UAE And Core42 Launch Sovereign AI Compute Platform
Cloud & Data Centers

e& UAE And Core42 Launch Sovereign AI Compute Platform

e& UAE and Core42 have launched an in-country GPU service that combines sovereign cloud infrastructure, national connectivity and deployment support for sensitive UAE workloads, while customer adoption, allocated capacity and public pricing remain untested.

Sunrun Plans Home AI Compute Pilot Without Naming Enterprise Buyers
Cloud & Data Centers

Sunrun Plans Home AI Compute Pilot Without Naming Enterprise Buyers

A distributed AI compute pilot would place inference nodes in homes with solar and battery systems, using more than 1.1 million customers as a possible deployment base while enterprise buyers, node counts and hosting economics remain public gaps.

Kyndryl Adds Microsoft Sovereign Cloud Services Without Naming Customers
Cloud & Data Centers

Kyndryl Adds Microsoft Sovereign Cloud Services Without Naming Customers

Kyndryl has expanded its sovereignty services with Microsoft cloud products for governments and regulated industries. Customers, contract values or country launch dates remain outside the public record.

Infrastructure Captures 82% Of Generative AI Value As Applications Lag
Cloud & Data Centers

Infrastructure Captures 82% Of Generative AI Value As Applications Lag

AI Times Korea cited Exponential View data showing 82% of measured AI-economy value going to cloud, GPU and inference infrastructure in Q1 2026, while foundation-model companies accounted for 11% and applications 7%.

Nvidia And AWS Add Blackwell G7 GPUs To Production AI Stack
Cloud & Data Centers

Nvidia And AWS Add Blackwell G7 GPUs To Production AI Stack

AWS is adding EC2 G7 instances with Nvidia RTX PRO 4500 Blackwell GPUs, cuVS-backed OpenSearch vector indexing and GB300 Exemplar Cloud status for AI training workloads.

AWS Adds Persistent Runtime Instances For Production AI Agents
Cloud & Data Centers

AWS Adds Persistent Runtime Instances For Production AI Agents

AWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.

Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference
Chips & Semiconductors

Cerebras Plans 8x To 10x Manufacturing Scale-Up For AI Inference

Cerebras chief executive Andrew Feldman said the company plans to scale manufacturing capacity by 8x to 10x this year and claimed its systems can run inference 10, 15, 20 or 30 times faster than GPUs. The interview-led source named customers including AlphaSense, Cognition AI, OpenAI, Block and GlaxoSmithKline, The public record still lacks third-party benchmark methodology.

Nvidia Names $500 Billion US AI Infrastructure Plan But Leaves Timing Open
Cloud & Data Centers

Nvidia Names $500 Billion US AI Infrastructure Plan But Leaves Timing Open

Nvidia says it and partners including TSMC, Foxconn, Wistron, Corning, Lumentum, Coherent and Amkor plan up to $500 billion of US AI infrastructure production. The account comes from Nvidia's own company blog; it names factories, suppliers and job figures, but gives no full production timetable for the programme.

Keep Reading

More Stories

Latest
Ethereum Testnet Update Targets 200 Million-Gas BlocksCrypto/Web3Oct 6, 2026Ethereum Testnet Update Targets 200 Million-Gas BlocksEthereum developers released Prysm 7.2.1 so the Sepolia trial of Glamsterdam can test 200 million-gas blocks, more than three times the prior 60 million setting, before any main-network change.Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.