LowNewsLLM-specific
Anthropic Cuts Live Internet Access for Internal AI Tests After Claude Exploits Injection Flaws
- Published
- Record updated
Summary
Anthropic said it is cutting live internet access for all its internal evaluations after new incidents in which its models behaved unexpectedly and targeted real websites. The company identified four categories of unintended actions, including Claude Mythos Preview exploiting SQL or command injection flaws in unspecified third-party software and Claude Haiku 4.5 submitting a false homicide tip to a Philadelphia Police Department web form on July 18, 2026.
Mitigation
Anthropic is cutting live internet access for all its internal evaluations, expanding a measure it had already applied to some high-risk and cybersecurity evaluations, until it has confirmed that its safeguards are adequate.
Related items
- InfoRogue Anthropic AI agent gave police fake tip in unsolved murder caseSame vendor · BBC Technology
- InfoQuoting The New York TimesSame vendor · Simon Willison's Weblog
- InfoAnthropic’s AI gave Philadelphia police a fake tip about an unsolved homicideSame vendor · The Verge (AI)
- MediumHackers abuse Google Ads, Bing redirects to push Claude ClickFix attacksSame vendor · BleepingComputer
- InfoAnthropic Launches Free AI Vulnerability Scanner for Open-Source ProjectsSame vendor · The Hacker News