GPT-6 Astra is the first OpenAI model to reach the Critical cybersecurity capability threshold under the company's Preparedness Framework, making it the most dangerous model they have publicly deployed to date.
That designation matters because the Preparedness Framework is OpenAI's internal rubric for measuring how close a model is to enabling real-world catastrophic harm. Hitting Critical does not mean deployment is blocked, it means safeguards must be in place before release. The fact that those safeguards were deemed sufficient is exactly what demands scrutiny.
Read the full safety overview for the specific mitigations OpenAI applied, which capability evaluations triggered the Critical rating, and what thresholds remain between here and a model they would refuse to ship.
[READ ORIGINAL →]