Path to Astra: critical capabilities and frontier safeguards
Astra is OpenAI's first model to hit Critical cybersecurity threshold under Preparedness Framework
“Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with stronger safeguards for release.”— OpenAI
OpenAI's Astra model has become the first to reach the 'Critical' cybersecurity capability threshold defined in the company's Preparedness Framework, triggering a new tier of required safeguards before release. This matters because it confirms OpenAI's internal capability evaluations are now reaching their highest risk thresholds, forcing new deployment constraints. It sets a precedent for how frontier labs will handle increasingly dangerous capability levels going forward.