Twitter/X

On 2026-08-07 @kimmonismus reported OpenAI's upcoming Astra model may have…

Brief

OpenAI's upcoming Astra model may have reached a 'Critical' cybersecurity capability level, per a 2026-08-07 post by @kimmonismus, after internal evaluations found major advances in agentic coding and cyber performance. OpenAI is pausing activities that fail new security requirements, tightening network/tool access, protecting model weights, and monitoring agentic use.

Why it matters

On 2026-08-07 @kimmonismus reported OpenAI's upcoming Astra model may have reached the 'Critical' threshold for cybersecurity capabilities; OpenAI reportedly said it "cannot rule out Critical capability level" after preliminary internal evaluations.

Key details

  • Under OpenAI's Preparedness Framework, a 'Critical' capability could include autonomously developing functional zero-day exploits against hardened real-world systems or executing novel end-to-end attacks from only a high-level goal.
  • OpenAI is pausing Astra work that fails strengthened security controls: restricting network and tool access, strengthening model-weight protection, and monitoring all agentic uses for risky behavior.
Source evidence

official now:

Chubby♨️ (@kimmonismus)

Doesnt look good for a soon to come releas for GPT-Astra: OpenAI says its upcoming Astra model may have reached the “Critical” threshold for cybersecurity capabilities.

"We are pausing internal activities involving Astra that do not yet meet these strengthened security control requirements."

Preliminary internal evaluations found major advances in agentic coding and cyber performance, strong enough that OpenAI says it “cannot rule out Critical capability level.”

Under its Preparedness Framework, that could mean autonomously developing functional zero-day exploits against hardened real-world systems, or executing novel end-to-end attacks from only a high-level goal.

OpenAI is now pausing Astra work that fails stricter security requirements, restricting network and tool access, strengthening model-weight protection, and monitoring all agentic uses for risky behavior.

— https://nitter.net/kimmonismus/status/2085812651163296161#m