Anthropic’s Opus 5 Is Better at Resisting Prompt Injection
Summary
Anthropic's Opus 5 model shows significant improvement in resisting prompt injection (attacks where users try to trick an AI by hiding malicious instructions in their input) compared to earlier versions and competing models. On the IPI benchmark test, Opus 5 reduced the success rate of attackers from 5.5% to 2.0% over 15 attempts, and outperformed all non-Claude models tested. While completely preventing prompt injection is impossible, the field is making progress at blocking these attacks in specific situations.
Classification
Affected Vendors
Related Issues
Original source: https://www.schneier.com/blog/archives/2026/07/anthropics-opus-5-is-better-at-resisting-prompt-injection.html
First tracked: July 31, 2026 at 02:01 PM
Classified by LLM (prompt v3) · confidence: 85%