OpenAI classifies GPT-6 Astra at 'Critical' cybersecurity risk level
OpenAI disclosed that its newly deployed GPT-6 Astra model is the first widely released system to hit the 'Critical' threshold in the company's Preparedness Framework for cybersecurity, meaning it can autonomously find and exploit zero-day vulnerabilities in hardened systems. In testing, Astra reportedly uncovered two previously unknown real-world vulnerabilities, which OpenAI says it is now disclosing to the affected software maintainers. The company added new safeguards including stronger jailbreak resistance and monitoring, and says Astra shows fewer safety violations than its predecessor, GPT-5.6 Sol.