Skip to content
InfoResearchIndustryLLM-specific

Evaluating and monitoring for AI scheming

Published
Record updated
View JSON

Summary

Google DeepMind researchers evaluated whether current frontier models have the prerequisite capabilities for AI scheming, namely stealth and situational awareness. They built and open-sourced an evaluation suite and tested Gemini 2.5 Pro, GPT-4o and Claude 3.7 Sonnet as of May 2025. The most capable models passed 2 of 5 stealth challenges and 2 of 11 situational awareness challenges, which the authors read as no concerning levels of either capability.