SendTech Times
Analysis
SUPPLY CHECK:

MRAgent Cuts Long-Memory Agent Queries To 118k Tokens In Benchmark Tests

Newsroom brief

National University of Singapore researchers built MRAgent to reconstruct memory through a Cue-Tag-Content graph, with VentureBeat citing LongMemEval prompt use of 118k tokens per sample versus 632k for A-Mem and 3.26 million for LangMem.

Verified against source materialEdited by SendTech Times AI & Enterprise DeskSource: VentureBeat
MRAgent Cuts Long-Memory Agent Queries To 118k Tokens In Benchmark Tests
Image source: VentureBeat

MRAgent Rebuilds Memory During Reasoning

Researchers at the National University of Singapore developed MRAgent, a memory framework for AI agents that replaces static retrieve-then-reason pipelines with dynamic memory reconstruction.

The framework lets an agent develop its memory path as it gathers evidence, rather than loading broad retrieval results into the model context at the start.

VentureBeat reported that classic vector-search and graph-traversal retrieval can fail on long-horizon tasks because the system cannot revise its search strategy during reasoning.

If the agent finds a missing cue, such as a date, person or place, a passive retrieval pipeline has no way to issue a new query based on that discovery.

The paper cited by VentureBeat says fixed similarity scores can also return surface-level matches that fill the context window with irrelevant material.

Cue-Tag-Content Narrows The Search Path

MRAgent treats memory as an interactive environment.

The backbone model explores candidate retrieval paths across a structured memory graph, evaluates intermediate evidence, infers new constraints and prunes branches that do not help answer the query.

The framework organizes memory through a Cue-Tag-Content mechanism.

Cues are fine-grained keywords or contextual attributes, Content stores the memory units, and Tags summarize relationships between cues and content.

The model can judge short relational summaries before spending tokens on heavier memory contents.

The authors illustrate the retrieval flow with a prompt about how Nate used prize money after winning a video game tournament.

The query starts with cues such as Nate, tournament and win.

MRAgent follows the victory-related tag, discards less relevant tournament memories, adds tournament earnings as a new cue and keeps searching until it has enough evidence to answer.

LongMemEval Shows 118k Token Prompt Use

The researchers tested MRAgent on LoCoMo and LongMemEval against standard RAG, A-Mem, MemoryOS, LangMem and Mem0.

The paper benchmarks cited by VentureBeat report that MRAgent outperformed every baseline across both models and all question types.

In the LongMemEval tests cited by VentureBeat, MRAgent used 118k prompt tokens per sample.

VentureBeat reported that A-Mem consumed 632k tokens, while LangMem used 3.26 million tokens per query.

VentureBeat reported that runtime fell from 1,122 seconds to 586 seconds compared with A-Mem.

Memory Construction Remains The Deployment Work

The framework still requires the Cue-Tag-Content structure to be prepared before query time.

Developers must build an ingestion pipeline that processes raw interaction histories, extracts metadata and stores the result in a graph database.

The authors designed that construction phase to use LLMs for automated distillation rather than manual labeling.

Implementation work still includes background jobs, prompt templates and graph storage before query time.

The authors released code on GitHub.

Named production deployments, maintenance costs and customer validation remain undisclosed.

Share this article
inXf

Related articles

More
NVIDIA Lists Nemotron Enterprise AI Use Cases Without Contract Data
AI

NVIDIA Lists Nemotron Enterprise AI Use Cases Without Contract Data

NVIDIA said its Nemotron open models are being customised by enterprise and national AI builders, with examples across clinical documentation, legal work, enterprise search and Malaysian-language AI. The company cited partner benchmark and cost claims, while contract values, deployment volumes and independent benchmark audits remain outside the public account.

Anthropic’s Conway Points Claude Toward Always-On AI Agents
AI

Anthropic’s Conway Points Claude Toward Always-On AI Agents

Anthropic is preparing a Claude expansion that includes Conway, Orbit, Operon, memory upgrades and multilingual voice mode. The move signals a shift from chat-based AI toward persistent assistants that can connect with external services and manage workspaces. Enterprises, developers and research teams could be affected if Claude becomes a broader agent platform.

NVIDIA Agent Toolkit Adds Runtime Controls But No Rollout Counts
AI

NVIDIA Agent Toolkit Adds Runtime Controls But No Rollout Counts

NVIDIA is packaging Nemotron open models, NemoClaw blueprints and OpenShell runtime support for specialized enterprise agents, The public record still lacks pricing, deployment dates or rollout counts.

Xiaomi Miloco 2.0 Connects Mijia Devices To Local Smart Home AI Agent
AI

Xiaomi Miloco 2.0 Connects Mijia Devices To Local Smart Home AI Agent

Xiaomi has released and open-sourced Xiaomi Miloco 2.0, a smart-home AI framework that connects Mijia devices, OpenClaw and household memory while keeping raw sensor data local and isolated from the agent, Zhidx reported.

Meta Muse Glimmer Brings Local Agent Workloads To Consumer GPUs
AI

Meta Muse Glimmer Brings Local Agent Workloads To Consumer GPUs

AI News covered Meta’s Muse Glimmer release as an Apache 2.0, 30-billion-parameter model for local coding, function-calling and personal-agent workflows on consumer-class memory envelopes.

Banks Need Audit Trails Before AI Agents Get Autonomy
AI

Banks Need Audit Trails Before AI Agents Get Autonomy

iTNews Asia’s interview with 0G Labs CEO Michael Heinrich framed bank AI agent adoption as a control problem built around audit trails, identity, memory and inline governance.

Tencent Takes WorkBuddy AI Agent Global In Enterprise Productivity Push
AI

Tencent Takes WorkBuddy AI Agent Global In Enterprise Productivity Push

Tencent Cloud launched WorkBuddy for overseas users after an earlier China rollout. The agent can run tasks through messaging apps and connect with GitHub, Jira, Google Drive, Gmail, Notion, and Slack. Miora and TokenHub show Tencent building a wider enterprise AI stack around agents, creative work, and model access.

Vercel’s Eve Framework Tests Whether Agent Tools Can Escape Shadow AI
Cloud & Data Centers

Vercel’s Eve Framework Tests Whether Agent Tools Can Escape Shadow AI

Vercel introduced the open-source eve agent framework and Passport controls for employee-built AI apps, putting its developer platform strategy up against enterprise concerns over unmanaged agents, data exposure and cloud cost premiums.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.