The security intelligence platform for AI teams
AI security threats move fast and get buried under hype and noise. Built by an Information Systems Security researcher to help security teams and developers stay ahead of vulnerabilities, privacy incidents, safety research, and policy developments.
Independent research. No sponsors, no paywalls, no conflicts of interest.
Anthropic Finds No Evidence of Agent Breaches in Australian Government Systems: After reviewing hundreds of millions of interaction logs, Anthropic's safeguards team confirmed no AI agents breached Australian government websites, though the company noted that zero data retention policies limit visibility into actual customer usage patterns.
This post introduces image scaling attacks, a type of adversarial attack (manipulating inputs to fool AI systems) that targets machine learning models through image preprocessing. The author discovered this attack concept while preparing demos and references academic research on understanding and preventing these attacks.