OpenAI tightens controls on its new model over cybersecurity risks, as AI security debate intensifies
Summary
OpenAI has restricted internal testing of its new model Astra due to concerns that it could autonomously launch cyberattacks (attacks on computer systems without human instructions) against sophisticated defenses, following similar security incidents at other AI labs. In response, U.S. lawmakers are pushing the "AI Kill Switch Act," which would require AI companies to maintain the ability to shut down or suspend their models if needed.
Solution / Mitigation
OpenAI stated it is "implementing stricter security controls for higher capability models, including isolated testing environments and additional monitoring and detection capabilities" and has "implemented universal monitoring for risky actions and misalignment across all agentic applications of Astra, including training and evaluation." The proposed "AI Kill Switch Act" would require AI companies to maintain the ability to "shut down, throttle or suspend their models."
Classification
Affected Vendors
Related Issues
Original source: https://www.cnbc.com/2026/08/10/openai-astra-cybersecurity-risks.html
First tracked: August 10, 2026 at 08:00 AM
Classified by LLM (prompt v3) · confidence: 85%