Skip to content
InfoResearchPeer-reviewed

(Un)Cooperative Robots: From Compliance AI to Principled Uncooperative AI

Published
Record updated
View JSON

Summary

Stuart Russell's human-compatible AI framework proposes that AI systems should defer unconditionally to human operators. This paper argues that total deference carries its own risks and limitations, and reports a HICSS workshop panel on (un)cooperative robots that can refuse human directives for principled reasons. The panel identified eleven research themes, including authority hierarchies, benevolent uncooperativeness, and AI liability, and called for research beyond compliance-centric AI design.