LIVE · cybersecurity feed
Live wire
OpenAI locks down Astra over potential critical cyber capabilitiesCritical Flaws Discovered in Belgian eID Software Used by 2 Million PeopleSecurity Affairs newsletter Round 589 by Pierluigi Paganini – INTERNATIONAL EDITIONWebmail CSS Attacks Expose a New Risk for AI-Powered Email ToolsMetabase Zero-Day Exploited in the Wild, Exposing Admin Access and Sensitive DataCritical One-Click Vulnerability in Atlassian’s Rovo AI Exposed Enterprise DataCVE-2026-8037 · CISA Adds Progress LoadMaster Command Injection Flaw to KEV CatalogSensitive Info Goes Into ‘No Reply’ Emails Constantly. This Guy Sees It AllAtlassian Rovo Can Be Tricked Into Sending Jira and Confluence Data to AttackersNew CSS Attacks Can Break Webmail Defenses to Steal Passwords and Tokens
aicritical

OpenAI locks down Astra over potential critical cyber capabilities

OpenAI’s internal evaluation of its upcoming model, Astra, found significant advances in agentic coding and cybersecurity, leading the company to conclude that it cannot rule out the model reaching the critical capability level for cybersecurity under its Preparedness Framework. The Preparedness Framework, first published in December 2023, outlines how OpenAI evaluates frontier AI risks and determ

zeroday.news ·

OpenAI has initiated a lockdown of its forthcoming Astra model after internal evaluations revealed significant advancements in its agentic coding and cybersecurity capabilities. The company stated it cannot definitively rule out Astra achieving a "critical capability" level for cybersecurity under its Preparedness Framework, which outlines risk assessment and safeguard protocols for advanced AI models.

The Preparedness Framework, established in December 2023, categorizes high-risk areas such as cybersecurity, biological and chemical threats, harmful persuasion, and AI self-improvement. Models undergo safety evaluations prior to deployment, and those deemed to pose unacceptable risks face delayed deployment or the implementation of additional safeguards.

Under this framework, an AI model is classified as possessing critical cybersecurity capabilities if it can autonomously identify previously unknown software vulnerabilities in secure systems or plan and execute sophisticated cyberattacks against well-protected targets with minimal human intervention. OpenAI noted it is applying a similar cautious approach to Astra as it did in 2025 when its AI models began demonstrating advanced biological capabilities.

OpenAI clarified that this is a preliminary assessment and Astra has not yet been formally classified as a critical cybersecurity model. The company is continuing its evaluations to make a final determination. It also explicitly stated that Astra is an upcoming model and was not implicated in any exploitation of Hugging Face.

In response to these findings, OpenAI has bolstered its safeguards and security controls for models exhibiting such advanced capabilities. During Astra's development, stricter security measures were implemented, including isolated testing environments, restricted network and tool access, enhanced protection and encryption of model weights, additional monitoring and detection systems, and sandboxed execution.

Activities involving Astra that do not meet these heightened security requirements have been paused. The company has also introduced comprehensive monitoring to detect risky actions and signs of misalignment within agentic applications.

Prior to Astra's deployment, OpenAI intends to engage government agencies and independent AI safety organizations to evaluate the model's cybersecurity capabilities. External testing partners will also receive guidance and security measures to facilitate the safe evaluation of the model's higher-risk functionalities.

ai
ShareXLinkedInWhatsAppFacebook

More News

view all →
ai

OpenAI's Next AI Model Astra Shows Cyber Performance Strong Enough to Trigger Pause

OpenAI has announced that it's pausing some "internal activities" involving its upcoming artificial intelligence (AI) model Astra after an internal evaluation found it had made significant advancements in agentic coding and cybersecurity. In response to the discovery, the AI upstart said it's implementing security controls for higher-capability models and associated activities, such as isolated

nation-state

How to report an AI Act violation in the EU

The EU’s fight to regulate AI models entered a new chapter on 2 August 2026, when the European Commission’s AI Office and national authorities began enforcing the AI Act. The AI Act is the EU’s law regulating AI, the first broad legal framework of its kind. It creates a common set of rules for AI systems used or sold in the EU, with the goal of encouraging innovation while protecting people’s safe

security

Solidity Pro VS Code Extensions Steal Crypto Wallets, API Keys, and Credentials

Cybersecurity researchers have flagged a malicious Microsoft Visual Studio Code (VS Code) extension named Solidity Pro ("solidity-pro") that has been observed delivering a browser wallet and credential stealer. The names of the extensions are below - helper-beeps.solidity-pro web3devtoolsx.solidity-pro Although neither of the extensions is now available on Open VSX, the GitHub repository

malware

GitHub Dependabot malware alerts now cover eight ecosystems

GitHub has flagged npm malware since March 2026. Anyone pulling in a bad PyPI, Maven, RubyGems, NuGet, Go, crates.io, or PHP Composer package has had no such warning, because GitHub’s malware detection only ever watched one ecosystem. That changed this month. GitHub’s Advisory Database now ingests malware reports from OpenSSF’s malicious-packages repository, a public feed in OSV format that launch

security

Chainloop: Open-source evidence store and policy engine for the software supply chain

Chainloop is an open source evidence store for the software supply chain. A command line tool runs inside a GitHub Actions, GitLab, Jenkins, or Dagger pipeline, picks up what the build produced, uploads those files to content-addressable storage, and references each one in a signed in-toto attestation. in-toto is a specification for recording who ran which step of a build, so the record can be che

cloud

Product showcase: Enpass Password Manager breaks away from the proprietary cloud model

Enpass is a password manager that stores passwords, passkeys, payment cards, identities, secure notes, software licenses, and other sensitive information in encrypted vaults. Vaults remain on the device or in a cloud storage service selected by the user. Users who work across multiple devices can install Enpass on Windows, macOS, Linux, Android, and iOS. Browser extensions are available for Chrome