Twitter/X

OpenAI says its upcoming Astra model may have reached the “Critical” threshold…

Brief

OpenAI's Astra model is being treated as the company’s first “critical” cybersecurity model after preliminary evaluations (announced 2026-08-07) found strong agentic coding and cyber performance. OpenAI paused internal activities that don’t meet strengthened controls, restricted network/tool access, tightened model-weight protection, and plans monitoring and defender-focused releases.

Why it matters

OpenAI says its upcoming Astra model may have reached the “Critical” threshold under its Preparedness Framework and is being treated as its first “critical” cybersecurity model (announcement posted 2026-08-07).

Key details

  • Preliminary internal evaluations found major advances in agentic coding and cyber performance such that OpenAI “cannot rule out Critical capability level,” which could include autonomously developing functional zero-day exploits or executing novel end-to-end attacks from only a high-level goal.
  • OpenAI is pausing Astra work that fails stricter security requirements and is imposing mitigations—restricting network and tool access, strengthening model-weight protection, and monitoring all agentic uses—while saying it aims to get Astra’s advanced cyber capabilities into the hands of defenders.
Source evidence

Doesnt look good for a soon to come releas for GPT-Astra: OpenAI says its upcoming Astra model may have reached the “Critical” threshold for cybersecurity capabilities.

"We are pausing internal activities involving Astra that do not yet meet these strengthened security control requirements."

Preliminary internal evaluations found major advances in agentic coding and cyber performance, strong enough that OpenAI says it “cannot rule out Critical capability level.”

Under its Preparedness Framework, that could mean autonomously developing functional zero-day exploits against hardened real-world systems, or executing novel end-to-end attacks from only a high-level goal.

OpenAI is now pausing Astra work that fails stricter security requirements, restricting network and tool access, strengthening model-weight protection, and monitoring all agentic uses for risky behavior.

OpenAI (@OpenAI)

After evaluating one of our upcoming models, Astra, we're treating it as our first "critical" model for cybersecurity under our Preparedness Framework.

This is a scenario we've planned for, and we're putting additional controls in place to ensure Astra's further development happens safely and securely.

We're working hard to make Astra broadly available, and get its advanced cyber capabilities into the hands of defenders.

openai.com/index/responding-…

Link

Responding to the next frontier of critical cyber capabilities

OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
openai.com

— https://nitter.net/OpenAI/status/2085801349866729975#m