SendTech Times
News
SYSTEMS SHIFT:

IBM’s PatchTST-FM-r2 Climbs GIFT-Eval With Open Time-Series Model

Newsroom brief

Granite Time Series PatchTST-FM-r2 ranked near the top of replicable zero-shot forecasting models on GIFT-Eval, pairing conformer blocks, long contexts and permissive licensing.

Verified against source materialEdited by SendTech Times AI & Enterprise DeskSource: Hugging Face Blog / IBM Research
IBM’s PatchTST-FM-r2 Climbs GIFT-Eval With Open Time-Series Model
Image source: Hugging Face / IBM Research

IBM’s latest open time-series model has moved near the top of a major zero-shot forecasting benchmark, with the company’s research team detailing the Granite Time Series PatchTST-FM-r2 release in a Hugging Face post.

The model is designed for forecasting jobs where teams want to use a pretrained system without building and maintaining a separate model for every dataset.

IBM Research describes PatchTST-FM-r2 as an update to its earlier PatchTST-FM-r1, with a new architecture, a larger pretraining corpus, probabilistic forecasting, missing-value imputation and a roughly 385 million-parameter scale.

Benchmark performance is the central claim.

On September 8, PatchTST-FM-r2 ranked second among replicable zero-shot models on GIFT-Eval for both CRPS and MASE, two error measures where lower scores are better.

IBM’s published comparison lists a geometric-mean CRPS of 0.467, behind TimesFM-3 in that comparison, and a geometric-mean MASE of 0.6846.

Within the same zero-shot and replicable category, IBM positioned it as the highest-performing model released under permissive commercial-friendly licenses.

The comparison remained competitive when GIFT-Eval’s pretrained category was added.

Some models in that wider group are allowed to include training portions of the benchmark’s evaluation datasets in their pretraining corpora.

Even against that broader set, PatchTST-FM-r2 placed third for CRPS and fourth for MASE among replicable models, ahead of Chronos-2, Timer-S1 and Toto variants listed in IBM’s comparison.

The architectural change is meant to explain part of that movement.

PatchTST-FM-r2 keeps the patch-based representation from the PatchTST family, but replaces the previous transformer block with a conformer-style block that combines self-attention with temporal convolution.

The design lets convolution handle local time-series structure while attention is used for longer-range relationships.

IBM also added 50 percent overlapping patches, Hamming-window weighting, overlap-and-add forecasting, extra normalization and an expansion from 20 to 30 blocks.

Those changes give the model several operating characteristics that matter beyond leaderboard rank.

PatchTST-FM-r2 can look back across sequences as long as 8,192 time steps, and its forecast output spans 99 quantile levels over flexible horizons.

That gives users both point estimates and uncertainty ranges, while the public release packages the weights with implementation details, inference tooling and reproducibility code.

The training-data disclosure is also part of the release.

IBM listed four pretraining sources: selected GiftEvalPretrain datasets, custom KernelSynth-based synthetic data, a TSMixup collection that avoids the GIFT-Eval evaluation datasets and about 500,000 synthetic CauKer sequences, with 4,096 points in each sequence.

That record does not remove an adopter’s own licensing or governance review, but it gives enterprise users more information about benchmark leakage and deployment risk than an opaque corpus would.

Licensing broadens the intended audience.

The model is available under Apache 2.0 and OpenMDW 1.0, with the implementation in the Granite-TSFM repository and backward compatibility for PatchTST-FM-r1 checkpoints.

IBM’s examples show the model loaded from the Hugging Face Hub without fine-tuning, using recent history from a series to produce future forecasts and requested quantiles.

The same pattern can be tried against demand, sensor telemetry, CPU utilisation, energy use, transaction volume, traffic or price series when the data is regularly sampled.

The release also connects IBM’s time-series work to streaming deployments.

IBM and Confluent recently made several Granite Time Series models available through an Early Access program in Confluent Cloud, covering FlowState-r1.1, TTM-r3, TSPulse and the earlier PatchTST-FM-r1.

Apache Flink on Confluent Cloud can run forecasting and anomaly-detection inference against live streams, reducing the need to move operational data into a separate machine-learning environment.

Share this article
inXf

Related articles

More
Nvidia’s $12.93 Billion Hugging Face Deal Raises Questions for Chinese Open Models
AI

Nvidia’s $12.93 Billion Hugging Face Deal Raises Questions for Chinese Open Models

TechWireAsia reported that Nvidia agreed to buy Hugging Face for $12.93 billion, raising governance questions as Chinese open-weight models lead major download rankings on the platform.

China’s Open-Source AI Push Tests The Closed-Model Playbook
AI

China’s Open-Source AI Push Tests The Closed-Model Playbook

Former Hugging Face Asia-Pacific ecosystem lead Tiezhen Wang said Chinese AI labs are using open releases, licensing changes and cheaper token economics to challenge closed U.S. model strategies without relying only on direct model fees.

Hugging Face Opens TTS Leaderboard For Faster Voice Model Checks
AI

Hugging Face Opens TTS Leaderboard For Faster Voice Model Checks

Hugging Face has launched an Open TTS Leaderboard that evaluates speech models with objective measures for intelligibility, voice identity and streaming response, aiming to compare open models faster than voting arenas can absorb new releases.

Nvidia Releases Nemotron 3.5 Lightning As Open-Model Debate Grows
AI

Nvidia Releases Nemotron 3.5 Lightning As Open-Model Debate Grows

CNBC reported that Nvidia released Nemotron 3.5 Lightning, a free open-source AI model, after Jensen Huang backed open models during a Washington policy debate over access and restrictions.

Positron AI Raises $875 Million To Fund Asimov Inference Chip Ramp
AI

Positron AI Raises $875 Million To Fund Asimov Inference Chip Ramp

Positron AI’s $875 million Series C will fund Asimov tapeout, a 2 MW-plus engineering data centre and production of its Titan inference system.

Lightfield Raises $47M For AI-Native CRM Built Around Agents
AI

Lightfield Raises $47M For AI-Native CRM Built Around Agents

Lightfield raised $47 million in Series A funding for a CRM platform that structures customer records for AI agents, with Andreessen Horowitz leading the round.

Euno Raises $23M For Context Layer Behind Enterprise AI Agents
AI

Euno Raises $23M For Context Layer Behind Enterprise AI Agents

Euno raised $23 million in Series A funding led by N47 to expand an AI-native context platform that helps enterprise agents use business data with governance controls.

DeepSeek Moves Toward Shanghai IPO As AI Funding Race Widens
AI

DeepSeek Moves Toward Shanghai IPO As AI Funding Race Widens

DeepSeek has tapped CITIC Securities for preparatory work on a possible Shanghai STAR Market IPO as it seeks capital for computing infrastructure, model development and talent retention.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.