InfoResearchPeer-reviewed
(Un)Cooperative Robots: From Compliance AI to Principled Uncooperative AI
- Published
- Record updated
Summary
Stuart Russell's human-compatible AI framework proposes that AI systems should defer unconditionally to human operators. This paper argues that total deference carries its own risks and limitations, and reports a HICSS workshop panel on (un)cooperative robots that can refuse human directives for principled reasons. The panel identified eleven research themes, including authority hierarchies, benevolent uncooperativeness, and AI liability, and called for research beyond compliance-centric AI design.