Tech News
← Home  ·  All topics

Preparedness Framework

3 GoKawiil briefs on this topic

OpenAI classifies GPT-6 Astra at 'Critical' cybersecurity risk level

OpenAI disclosed that its newly deployed GPT-6 Astra model is the first widely released system to hit the 'Critical' threshold in the company's Preparedness Framework for cybersecurity, meaning it can autonomously find and exploit zero-day vulnerabilities in hardened systems. In testing, Astra reportedly uncovered two previously unknown real-world vulnerabilities, which OpenAI says it is now disclosing to the affected software maintainers. The company added new safeguards including stronger jailbreak resistance and monitoring, and says Astra shows fewer safety violations than its predecessor, GPT-5.6 Sol.

OpenAI classifies upcoming Astra model as 'Critical' risk under cybersecurity framework

OpenAI announced that its forthcoming Astra AI model is the first to cross the company's 'Critical' cybersecurity capability threshold, meaning it can discover unknown vulnerabilities and exploit them autonomously without human step-by-step direction. The company says Astra will still launch soon, but access to its most sensitive cybersecurity abilities will be restricted, with further safety details to come in a system card at release.

AI lab designates 'Astra' model as first to cross critical cybersecurity risk threshold

The company behind the Astra model says new testing shows it has crossed a 'Critical' cybersecurity capability threshold under its Preparedness Framework, meaning it could independently discover unknown vulnerabilities and craft exploits against well-defended systems without step-by-step human guidance. This is the first model the company has classified at that severity level, prompting delays to strengthen safeguards before release.