OpenAI launches GPT-6 Astra, its first model to cross a critical cybersecurity threshold
Summary
OpenAI released GPT-6 Astra, a new AI model that crossed the "Critical" threshold in its Preparedness Framework (a system for measuring AI cybersecurity risks). The model scored 100% on ExploitBench, a test measuring how well it can identify security vulnerabilities (weaknesses in software), and even discovered two new zero-day exploits (previously unknown security flaws). Because of this critical risk level, OpenAI is limiting access by default and requiring enterprise administrators to manually enable it, though the public version will refuse to generate advanced attack tools.
Solution / Mitigation
OpenAI is implementing the following restrictions: Enterprise administrators must manually enable Astra for their workspace since access is off by default at launch. The public version of Astra will refuse advanced offensive tasks such as generating proof-of-concept exploits (working examples of attacks). Additionally, OpenAI plans to loosen restrictions for vetted defenders through a program called OpenAI Daybreak in the coming weeks.
Classification
Affected Vendors
Related Issues
Original source: https://www.csoonline.com/article/4218679/openai-launches-gpt-6-astra-its-first-model-to-cross-a-critical-cybersecurity-threshold.html
First tracked: September 4, 2026 at 08:01 AM
Classified by LLM (prompt v3) · confidence: 92%