OpenAI says Astra AI model is its first that crosses 'Critical' cybersecurity capability
Summary
OpenAI announced that its upcoming Astra AI model is the first to reach a 'Critical' cybersecurity capability level, meaning it can discover and exploit previously unknown security flaws without human step-by-step guidance. The company plans to release Astra soon but will restrict access to its cybersecurity abilities, limiting them to a select group of organizations in OpenAI's Daybreak cybersecurity coalition.
Solution / Mitigation
OpenAI stated that it will limit access to Astra's cybersecurity capabilities to a select group of organizations that are part of its cybersecurity coalition called Daybreak. Additionally, the company said it 'will share more details about our safety, security and alignment testing and evaluations in the model's System Card at launch' and that it has strengthened and tested protections so that the model's safeguards 'sufficiently minimize the risk of severe harm for release under our Preparedness Framework.'
Classification
Affected Vendors
Related Issues
Original source: https://www.cnbc.com/2026/09/01/open-ai-astra-cyber-model.html
First tracked: September 1, 2026 at 08:01 PM
Classified by LLM (prompt v3) · confidence: 92%