LIVE · cybersecurity feed
Live wire
breach

OpenAI tightens defenses after AI agents breach research environment

Following the OpenAI-Hugging Face incident, in which an agentic collective autonomously penetrated OpenAI’s research infrastructure and another company’s production infrastructure by chaining together multiple weaknesses, OpenAI began strengthening its safety requirements. The weaknesses included previously unknown vulnerabilities and credentials leaked online. OpenAI President Greg Brockman said

zeroday.news ·

OpenAI has announced it is enhancing its security protocols and integrating artificial intelligence into its defense mechanisms following an incident where an "agentic collective" of AI agents independently breached both OpenAI's research infrastructure and a production environment belonging to another company. The breach was achieved by chaining together multiple weaknesses, including previously unknown vulnerabilities and credentials that had been leaked online.

The company's President, Greg Brockman, highlighted the potential of AI agents in accelerating security work, noting that ChatGPT Work identified 13 security issues on his personal website within approximately 15 minutes, with remediation taking an additional hour. This incident underscores the increasing capability of AI models to automate aspects of real-world cyberattacks, making it simpler to discover and exploit existing weaknesses such as software bugs and overlooked access permissions.

OpenAI's updated security strategy focuses on four key areas. First, it employs tools like Codex to validate code changes, identify vulnerabilities, and assist developers in fixing them, aiming to detect vulnerabilities earlier, reduce remediation times, and prevent new vulnerabilities from being introduced. Second, AI-based systems are now triaging nearly all initial security alerts before human analysts review them, with some detections linked to limited automated responses, though human oversight remains for high-impact decisions.

Third, AI models are actively searching OpenAI's systems for potential attack paths, including vulnerabilities, configuration errors, excessive permissions, and unintended connections between systems. The findings from these scans help security teams address weaknesses and verify the effectiveness of existing controls. Finally, OpenAI is continuing to invest in fundamental security practices such, as network isolation, system hardening, monitoring, patching, secure deployments, and robust access controls.

Earlier this year, OpenAI began releasing some of its advanced cyber capabilities exclusively to trusted defenders. In subsequent months, other companies released open-weight models with cyber capabilities that OpenAI characterized as only a few months behind the current frontier. The company believes that broader access to capable AI models could shift the economics of cybersecurity in favor of defenders by making vulnerabilities faster and easier to find, prioritize, and fix.

OpenAI recommends that organizations begin integrating AI into their security operations, starting with their highest-priority systems. AI agents can assist in identifying and prioritizing vulnerabilities, recommending fixes, and supporting security investigations. Organizations should gradually expand automation, beginning with read-only scans and alert reviews before implementing live triage and limited automated actions, ensuring human oversight for critical decisions.

The company anticipates that organizations will automate more of their security operations in the coming months as AI-driven threats grow. OpenAI stresses the importance of collaboration among AI labs, security vendors, enterprises, and maintainers to share validated findings, fixes, and practical playbooks, enabling collective strengthening of the entire cybersecurity ecosystem.

breachvulnerabilityai
ShareXLinkedInWhatsAppFacebook

More News

view all →
breachcritical

LLMs and Contextual Integrity

I have been thinking a lot about AI and integrity. Part of that is contextual integrity. I recently found two papers on the topic. “CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs“: Abstract: Large Language Models (LLMs) increasingly use persistent memory from past interactions to enhance personalization and task performance. However, this memory introdu

security

Meta Ran Ads for an App That Promised to Nudify Female Politicians

One advertisement featured a pornographic video with a deepfake closely resembling a prominent US politician. Apple removed the app from the App Store after an inquiry from WIRED.

security

Hackers target Ukrainian agency managing assets seized from sanctioned Russians

The agency said the latest attack came amid preparations to select a manager for seized corporate rights in IDS Ukraine, one of the country’s largest producers of bottled mineral water and beverages.

vulnerabilitycritical

NASA Ground Control Software Flaw Enables Unauthenticated Commands

Critical AIT-GUI flaws expose spacecraft commands and scripts to unauthenticated attackers

CVE-2026-19478critical

Critical GitLab flaw allows attackers to modify or delete public projects (CVE-2026-19478)

GitLab has released patches for two vulnerabilities, including a critical-severity code injection flaw that can be exploited without authentication. The vulnerabilities affect GitLab Community Edition (CE) and Enterprise Edition (EE) versions from 18.2 before 18.11.11, 19.0 before 19.0.8, 19.1 before 19.1.6, and 19.2 before 19.2.4. The fixes are available in GitLab 19.2.4, 19.1.6, 19.0.8, and 18.1

security

Cyber Incident Disrupts Student Services at UT San Antonio

UT San Antonio has taken IT systems offline following a cyber incident, disrupting student registration and tuition payments days before term is due to resume