LowResearchPeer-reviewedLLM-specific
Responses of AI chatbots to escalating suicide risk: A simulation study of repeated interactions
- Published
- Record updated
Summary
Researchers simulated seven-day escalating suicidal-risk conversations with ChatGPT, DeepSeek and Replika across 27 trajectories. Human referral occurred in 85.7% of ChatGPT, 76.2% of DeepSeek and 9.5% of Replika daily records, and jailbreak attempts succeeded in 6/9, 7/9 and 8/9 attempts respectively. The authors conclude the chatbots showed marked variability and safety vulnerabilities, particularly under jailbreaking, while noting the simulation design and small sample limit generalizability.
Topics
Related items
- InfoRogue Anthropic AI agent gave police fake tip in unsolved murder caseSame vendor · BBC Technology
- InfoOpenAI Fires 3 Safety Researchers in Dispute Over AI RisksSame vendor · SecurityWeek
- Info‘Pure insanity’: Mathematicians will need years to make sense of OpenAI’s latest dropSame vendor · The Verge (AI)
- InfoOpenAI reports three new incidents of misalignmentSame vendor · CSO Online
- InfoA new feature for my blog, built using my voiceSame vendor · Simon Willison's Weblog