One big question in frontier AI policy is the extent to which frontier labs would actually follow their 'safety and security frameworks' when it mattered. Would these foundational governance documents really have teeth, or would labs--even after the passage of mandatory disclosure laws like SB 53--find ways to wriggle their way out of following the letter and spirit of their safety plans, given how ambiguous and fast-moving frontier AI is known to be?
Today we face just such a scenario. Our next model, Astra, may be 'critical' under our Preparedness Framework. We cannot rule out the serious possibility that it is, and so we are going to take steps consistent with the higher risk level (critical) rather than assuming the model is at a lower risk level. These steps include the ones listed in the screenshot below.
Some of these decisions have the effect of slowing down internal development, and in that sense they are costly decisions. But they are the right decisions. I am proud of OpenAI for making them.