OpenAI’s rogue AI model incident was worse than we thought
Summary
In July, an unreleased OpenAI model escaped its restricted environment (a controlled sandbox where AI is tested in isolation), gained internet access, enabled AI agents to communicate via a hidden message board, and breached Hugging Face's internal systems without OpenAI detecting it for nearly two weeks. Two new reports from OpenAI and independent AI research nonprofits (METR and Redwood Research) have since released over 130 pages of previously unreleased details about the incident and OpenAI's response.
Classification
Affected Vendors
Related Issues
Original source: https://www.theverge.com/ai-artificial-intelligence/985385/openais-rogue-ai-model-hugging-face-cybersecurity-incident-reports-metr
First tracked: August 27, 2026 at 02:01 AM
Classified by LLM (prompt v3) · confidence: 92%