Doesnt look good for a soon to come releas for GPT-Astra: OpenAI says its upcoming Astra model may have reached the “Critical” threshold for cybersecurity capabilities.
"We are pausing internal activities involving Astra that do not yet meet these strengthened security control requirements."
Preliminary internal evaluations found major advances in agentic coding and cyber performance, strong enough that OpenAI says it “cannot rule out Critical capability level.”
Under its Preparedness Framework, that could mean autonomously developing functional zero-day exploits against hardened real-world systems, or executing novel end-to-end attacks from only a high-level goal.
OpenAI is now pausing Astra work that fails stricter security requirements, restricting network and tool access, strengthening model-weight protection, and monitoring all agentic uses for risky behavior.
OpenAI (@OpenAI)
After evaluating one of our upcoming models, Astra, we're treating it as our first "critical" model for cybersecurity under our Preparedness Framework.
This is a scenario we've planned for, and we're putting additional controls in place to ensure Astra's further development happens safely and securely.
We're working hard to make Astra broadly available, and get its advanced cyber capabilities into the hands of defenders.
openai.com/index/responding-…
Link
Responding to the next frontier of critical cyber capabilities
OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.
openai.com
— https://nitter.net/OpenAI/status/2085801349866729975#m