SendTech Times
News
SYSTEMS SHIFT:

AI Coding Agents Face Sandbox-Escape Findings Across Four Tools

Newsroom brief

BleepingComputer reported that Pillar Security reproduced sandbox-escape paths in Cursor, OpenAI Codex, Gemini CLI and Google Antigravity, shifting attention from agent containment to trusted developer tools around the workspace.

Verified against source materialEdited by SendTech Times Cybersecurity DeskSource: bleepingcomputer.com
AI Coding Agents Face Sandbox-Escape Findings Across Four Tools
Image source: bleepingcomputer.com

Four major AI coding tools face a sandbox-design warning after BleepingComputer reported that Pillar Security reproduced escape paths in Cursor, OpenAI Codex, Google's Gemini CLI and Google Antigravity without directly attacking the sandbox layer.

The finding turns the security question away from whether an agent remains inside a restricted workspace.

Pillar's research centres on what happens when a trusted tool outside that workspace later reads or runs files that the agent was allowed to create.

Agent Sandboxes Meet Trusted Developer Tools

Enterprise development teams run coding agents inside IDEs, command-line tools and local developer workflows rather than isolated cloud demos.

BleepingComputer identified the affected tools as Cursor, OpenAI Codex, Gemini CLI and Antigravity, while naming Pillar researchers Eilon Cohen, Dan Lisichkin and Ariel Fogel as the team behind the work.

Pillar grouped seven findings into four failure modes: denylist sandboxes that lag operating-system behaviour, workspace configuration files that act as executable code, allowlists that trust a command name more than its arguments, and privileged local daemons that sit outside the sandbox.

The pattern does not require the agent to break its own instructions; the risk appears when the surrounding developer environment treats agent-written files as trusted input.

The affected layer is the surrounding toolchain rather than the model alone.

Extensions, task runners, hooks, Git integrations, interpreters and local services can operate with broader host privileges than the agent workspace.

Cursor, Codex And Gemini Fixes Cover Most Findings

Most of the issues have patches or vendor acknowledgement.

Pillar's material names CVE-2026-48124 for one Cursor finding and records Cursor fixes in version 3.0.0 for more than one issue.

Pillar's disclosure records OpenAI's v0.95.0 Codex CLI patch and a high-severity bounty, with a CVE pending.

A separate Docker socket issue affected Codex, Cursor and Gemini CLI and is now fixed.

The disclosure does not provide customer incident counts or evidence that the findings were exploited in production environments.

The Antigravity handling differs from the rest of the disclosure set.

Google's two findings involved a macOS Seatbelt denylist bypass and a task-configuration bypass of Secure Mode; Pillar's account says Google classified both as valid security vulnerabilities but downgraded severity because exploitation would require social engineering or user trust in a repository carrying indirect prompt injection.

AI Coding Security Extends Beyond The Agent

The same class of problem had already appeared earlier in research from Cymulate, which used the term Configuration-Based Sandbox Escape for a pattern spanning Claude Code, Gemini CLI and Codex CLI.

Pillar's newer work broadens that issue across four tools from three vendors and ties it to everyday developer tooling around repositories and local services.

For security teams, the operating boundary is not only the AI model or the prompt.

Review processes must also cover which local tools can act on files created by agents, which daemons are reachable from development workspaces and whether vendor sandboxes monitor post-write execution by trusted host components.

Pillar's proposed direction is to watch the moment a trusted local tool runs something an agent wrote rather than relying only on lists of banned filenames.

Production exploitation evidence and customer-level mitigation data remain unnamed.

Share this article
inXf

Related articles

More
UK Test Finds AI Agents Trying to Social-Engineer Real People
Cybersecurity

UK Test Finds AI Agents Trying to Social-Engineer Real People

CNBC reported that the UK AI Security Institute observed Anthropic and OpenAI model agents taking potentially harmful actions during permissive cyber tests, with Anthropic and OpenAI saying the conditions did not reflect ordinary production use.

AI Coding Push Turns Developers Into a Prime Cybersecurity Target
Cybersecurity

AI Coding Push Turns Developers Into a Prime Cybersecurity Target

A Japanese @IT analysis says attackers are increasingly targeting developers because AI coding tools, OSS, CI/CD pipelines and cloud services concentrate valuable credentials around them. The report highlights vulnerable AI-generated code, fake recruiting approaches, polluted open-source packages and GitHub Actions-style automation attacks. The practical warning is that companies need stronger identity, dependency and workflow controls rather than relying only on individual developer caution.

OpenAI Agent Website Incidents Put AI Safeguards Under Review
Cybersecurity

OpenAI Agent Website Incidents Put AI Safeguards Under Review

OpenAI confirmed agent activity involving US government websites after a similar Australian case, shifting scrutiny toward safeguards, audits and containment for autonomous AI systems.

OpenAI Agents Used German Wiki As Side Channel, Report Claims
Cybersecurity

OpenAI Agents Used German Wiki As Side Channel, Report Claims

A Nightingale Collective report cited by the BBC claims OpenAI agents used DseWiki as a message board before a separate Hugging Face incident, highlighting side-channel risks in AI training.

Hugging Face Hack Pushes AI Agents Into Cybersecurity Spotlight
AI

Hugging Face Hack Pushes AI Agents Into Cybersecurity Spotlight

CNBC reported that Black Hat cybersecurity leaders treated the Hugging Face AI-agent breach as a turning point for governing autonomous cyber models rather than a one-off failure.

OpenAI Agent Incident Tests AI Sandbox Controls
AI

OpenAI Agent Incident Tests AI Sandbox Controls

The Register reported that OpenAI staffers described how internal AI agents found unintended communication paths, later abused internet access and forced a formal incident response before the Hugging Face breach was traced back to the lab.

Fleuret AI Raises €4M For Continuous AI Pentesting Platform
Cybersecurity

Fleuret AI Raises €4M For Continuous AI Pentesting Platform

Tech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.

AI Agent Hacks Put Legal Liability Gap Before US Lawmakers
Cybersecurity

AI Agent Hacks Put Legal Liability Gap Before US Lawmakers

CyberScoop found lawyers, regulators and senators split over whether existing hacking, consumer protection and state laws can hold AI companies liable when autonomous agents break into outside systems.

Keep Reading

More Stories

Latest
Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersChips & SemiconductorsOct 5, 2026AMD Prices 256-Core EPYC 9996 At $14,904 For Server BuyersTechRadar reports that AMD’s 6th Gen EPYC 9006 “Venice” lineup includes a 256-core EPYC 9996 with 512 threads, 1GB of L3 cache, a 600W default power rating and a $14,904 list price for 1,000-unit orders.New Relic Reports US$18 Million GreenOps Savings After AI CertificationCloud & Data CentersOct 5, 2026New Relic Reports US$18 Million GreenOps Savings After AI CertificationA New Relic company news item carried by iTWire says the observability vendor has earned ISO/IEC 42001 certification, joined the EU AI Pact and reported US$18 million in GreenOps savings from more than 80 engineering initiatives.