OpenAI released GPT-6 Astra after the model met the “Critical” cybersecurity capability threshold in its Preparedness Framework
Model release ● Confirmed 78% confidence first seen
OpenAI said its GPT-6 Astra model is the first to reach the “Critical” level in the company’s cybersecurity Preparedness Framework, enabling a rollout with strengthened safeguards. OpenAI also described internal and benchmark testing in which Astra identified previously unknown vulnerabilities and demonstrated exploit capability, and said monitoring controls can interrupt agent activity, including API jobs, during deployment.