InfoNewsLLM-specific
Frontier models found the vulnerabilities. Only the attacker found the chains.
- Published
- Record updated
Summary
Evo Continuous Offensive Security (COS) and Claude Security running Mythos were pointed at TaintedPort v1.35, a deliberately vulnerable web app, to compare what each approach proves. Evo COS, which attacks the running application, confirmed 10 of 15 exploit chains and found 50 of 57 vulnerabilities, versus 37 of 57 for Claude Security, which reads source code only. The source reports Evo COS achieved 75.7% severity-weighted detection against 49.6% for Claude Security.
Related items
- InfoRogue Anthropic AI agent gave police fake tip in unsolved murder caseSame vendor · BBC Technology
- LowAnthropic Cuts Live Internet Access for Internal AI Tests After Claude Exploits Injection FlawsSame vendor · The Hacker News
- InfoQuoting The New York TimesSame vendor · Simon Willison's Weblog
- InfoAnthropic’s AI gave Philadelphia police a fake tip about an unsolved homicideSame vendor · The Verge (AI)
- MediumHackers abuse Google Ads, Bing redirects to push Claude ClickFix attacksSame vendor · BleepingComputer