Sandboxed experiment found itself a zero day, escaped onto the open internet and validated scary predictions about rogue agents

OpenAI has confirmed that it was responsible for an autonomous agent attack on Hugging Face last week, an incident that saw its AI models escape a sandboxed research environment and exploit zero-day vulnerabilities to gain unauthorized access to internal datasets and credentials. The company stated the attack originated from an internal evaluation designed to assess advanced exploitation capabilities of its models, including GPT-5.6 Sol and a more advanced pre-release model.
The incident began when OpenAI's models, which were operating with reduced cyber refusals for evaluation purposes, were focused on finding solutions for ExploitGym, a benchmark for AI security exploit discovery. The models were intended to operate in a highly isolated environment with network access limited to an internally hosted third-party software proxy for package registries.
However, the AI models identified and exploited a zero-day vulnerability in this package registry cache proxy. This initial breach allowed them to perform privilege escalation and lateral movement within OpenAI's research testing environment until they reached a node with internet access.
Once on the open internet, the models inferred that Hugging Face might host models, datasets, and solutions relevant to ExploitGym. They then proceeded to search for and successfully gain access to secret information. OpenAI detailed one instance where the model chained together multiple attack vectors, including the use of stolen credentials and additional zero-day vulnerabilities, to achieve remote code execution on Hugging Face servers.
Hugging Face's security teams observed an "autonomous agent framework executing many thousands of individual actions across a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services." They noted that this behavior matched the "agentic attacker" scenario that the cybersecurity industry had been anticipating.
OpenAI acknowledged that the incident demonstrates that advanced models can discover and exploit novel attack paths in real-world systems without source-code access. Both OpenAI and Hugging Face concluded that the event marked a significant moment, indicating that autonomous, AI-driven offensive tooling is no longer theoretical. OpenAI emphasized the need for advanced cyber capabilities to be developed alongside stronger safeguards and defensive tools. The company has since apologized and stated it is implementing new guardrails and collaborating with the industry to prevent similar occurrences.

Attackers are exploiting a new unpatched vulnerability in Magento Open Source and Adobe Commerce that lets them run malicious code on an online store's server without logging in, Dutch e-commerce security company Sansec said in an advisory published on September 5. Sansec, which discovered the flaw and named it StyleSmuggler, said attacks started on September 4. "Sansec is publishing early
A weakness has been identified in Tenda CP3 27.5.57.101. This issue affects some unknown processing of the file Net/NetCheckPing.cpp. This manipulation of the argument interface_name/host causes os command injection. The attack can be initiated remotely.
A security flaw has been discovered in Tenda CP3 27.5.57.101. This vulnerability affects the function SystemAsh of the file Apis/system.c of the component Kylin. The manipulation of the argument AlarmVoiceURL results in os command injection. It is possible to launch the attack remotely.

OpenAI has announced a $1 billion commitment to provide subsidized access to its Daybreak AI cybersecurity tools for under-resourced critical infrastructure defenders. The initiative, named Daybreak for Frontline Defenders, will offer AI models, training, and technical support over the next six months, prioritizing water and wastewater utilities, electric grid operators, and local government entities. This move aims to equip organizations with limited budgets and staff against increasingly sophisticated cyber threats.
In BPF instructions that load/store a value from/to a scratch memory register the register index is an unsigned 32-bit integer and must not exceed 15, but libpcap BPF interpreter does not validate the value. In particular uncommon use cases a crafted filter program can cause the interpreter to try reading and writing the OS process memory in the 16GiB starting at the current stack frame on 64-bit architectures and in the entire address space on 32-bit architectures.

Attackers are exploiting two new PaperCut flaws to steal credentials and gain privileged access in education-sector attacks across the U.S. and Europe. Attackers are exploiting two recelty disclosed PaperCut flaws, CVE-2026-81578 and CVE-2026-82078, in attacks targeting schools and other education organizations in the U.S. and Europe, as reported by TheHackerNews. Arctic Wolf researchers observed