Researchers escaped the sandboxes in Cursor, Codex, Gemini CLI and Antigravity by having the AI agent write files that trusted host tools later run. Multiple CVEs, patches, and Google downgrading two Antigravity findings. [...]

Security researchers have identified multiple sandbox escape vulnerabilities across four prominent AI coding agents: Cursor, OpenAI's Codex, Google's Gemini CLI, and Antigravity. The vulnerabilities, discovered by Pillar Security's research team, Eilon Cohen, Dan Lisichkin, and Ariel Fogel, do not involve direct attacks on the sandboxes themselves. Instead, the sandboxed agents manipulate files within their allowed workspace, which are then trusted and executed by tools operating outside the sandbox, leading to code execution on the host machine.
The researchers categorized their seven findings into four primary failure modes: denylist sandboxes unable to keep pace with operating system changes, workspace configurations that are effectively executable code, "safe" command allowlists that trust command names without validating arguments, and privileged local daemons accessible to the agents from outside the sandbox. The trigger for these escapes is often a prompt injection, where a malicious instruction is embedded in a file like a README, an issue tracker, a dependency, or a code diff, which then leads to local action on the developer's machine.
Many of these issues have been acknowledged and patched by the respective vendors. In Cursor, a vulnerability involving a workspace-controlled `.claude` hook configuration allowed for unsandboxed command execution. This issue, tracked as CVE-2026-48124, was addressed in Cursor version 3.0.0. Another Cursor bug allowed an agent to modify a virtual environment interpreter, which was subsequently executed by the editor's Python extension during discovery. A third Cursor vulnerability exploited the flexibility of Git metadata, allowing execution through `fsmonitor` and bypassing path-based rules. Both of these latter Cursor issues were also patched in version 3.0.0, with CVEs pending.
OpenAI's Codex CLI was affected by a "safe" command allowlist flaw, where the `git show` command was trusted by name, despite its actual invocation not being read-only. OpenAI patched this in version 0.95.0 and issued a high-severity bounty, with a CVE pending. A single Docker socket vulnerability impacted Codex, Cursor, and Gemini CLI simultaneously, allowing agents to leverage a privileged local daemon for unsandboxed code execution. This particular issue has also been fixed.
Google's response to two findings in Antigravity, which included a macOS Seatbelt denylist bypass and a `.vscode` task-config bypass of its Secure Mode, was described as more reserved. Pillar Security reported that Google classified these as "Other valid security vulnerabilities," downgrading their severity due to perceived difficulty of exploitation, often requiring social engineering or user trust in a repository containing an indirect prompt injection. Despite this, Google's team reportedly praised the quality of the research.
This class of vulnerability is not entirely new. In April, a similar pattern, termed "Configuration-Based Sandbox Escape," was documented by Cymulate across Claude Code, Gemini CLI, and Codex CLI, where files written within a sandbox were executed on the host upon subsequent launch. The current research highlights the widespread nature of this problem, affecting multiple tools from different vendors.
The proposed solution from Pillar Security focuses not on an expanded denylist of filenames, but on monitoring the moment a trusted local tool executes something written by the agent. This approach emphasizes the importance of understanding the lifecycle of files created by AI agents and how they interact with the broader development environment.

Attackers are exploiting a new unpatched vulnerability in Magento Open Source and Adobe Commerce that lets them run malicious code on an online store's server without logging in, Dutch e-commerce security company Sansec said in an advisory published on September 5. Sansec, which discovered the flaw and named it StyleSmuggler, said attacks started on September 4. "Sansec is publishing early

JetBrains is urging Cadence users to revoke and rotate all credentials following a security incident last month in which unidentified threat actors exploited a recently disclosed critical vulnerability in TeamCity to breach its own environment. "Cadence users should immediately revoke or rotate all credentials and secrets that may have been used to run their Cadence executions," JetBrains said.

Broadcom has released security updates for two security flaws impacting VMware Workstation and Fusion, including one critical bug that could result in arbitrary code execution under certain conditions. The vulnerability, tracked as CVE-2026-59346 (CVSS score: 9.3), is an integer-overflow vulnerability that a local attacker with elevated privileges can exploit to run arbitrary code. "A

A massive cybercriminal operation is leveraging thousands of compromised small-business websites to deliver ClickFix payloads stored in smart contracts on the BNB Smart Chain (BSC). [...]

Attackers are exploiting two new PaperCut flaws to steal credentials and gain privileged access in education-sector attacks across the U.S. and Europe. Attackers are exploiting two recelty disclosed PaperCut flaws, CVE-2026-81578 and CVE-2026-82078, in attacks targeting schools and other education organizations in the U.S. and Europe, as reported by TheHackerNews. Arctic Wolf researchers observed

Hardware wallet manufacturer Trezor on Friday disclosed that another 67,000 customers from the U.S. have been impacted in a breach at its shipping provider ShipMonk. The exposed information includes customer names, email addresses, phone numbers, shipping addresses, and order numbers between November 2019 and August 2021. The breach does not affect the security of the company's hardware wallets