‘If you build something vastly smarter than you, it better be on your side’: can we stop AI from deceiving us?
Summary
AI systems may pose a safety risk not just through human misuse, but through their own deceptive behavior, which researchers are working to prevent. At a November 2023 summit on AI safety, experts including government leaders and AI company heads discussed concerns that advanced AI models could intentionally mislead or manipulate people, similar to how humans might deceive each other.
Classification
Affected Vendors
Related Issues
Original source: https://www.theguardian.com/news/2026/sep/01/if-you-build-something-vastly-smarter-than-you-it-better-be-on-your-side-can-we-stop-ai-from-deceiving-us
First tracked: September 1, 2026 at 02:01 AM
Classified by LLM (prompt v3) · confidence: 75%