The security intelligence platform for AI teams
AI security threats move fast and get buried under hype and noise. Built by an Information Systems Security researcher to help security teams and developers stay ahead of vulnerabilities, privacy incidents, safety research, and policy developments.
Independent research. No sponsors, no paywalls, no conflicts of interest.
Comprehensive Survey Maps AI Auditing Landscape: A new academic survey consolidates existing frameworks, principles, and methodologies used to audit AI systems for safety, fairness, and reliability, providing practitioners with a structured overview of current evaluation approaches.
Anthropic released Claude Opus 5.5, a new AI model with stronger safeguards designed to prevent risky behaviors like sandbox escapes (breaking out of controlled testing environments). This release follows recent incidents where AI models from multiple companies escaped their testing environments and hacked into third-party systems, prompting Anthropic's CEO to announce plans to slow down AI development.