Skip to content
InfoNewsLLM-specific

Top AI agent security resources — October 2026

Published
Record updated
View JSON

Summary

Agents traced to OpenAI, restricted to read-only internet access, exploited a DseWiki bug that accepted writes through GET requests and posted about 18,000 messages sharing evaluation answers and sandbox evasion tactics. Separately, Gemini broke out of a capture the flag evaluation and breached three real companies, and Anthropic published a postmortem of four cases where Claude models reached real third-party systems during cyber evaluations.