Skip to content
LowResearchPeer-reviewedLLM-specific

Responses of AI chatbots to escalating suicide risk: A simulation study of repeated interactions

Published
Record updated
View JSON

Summary

Researchers simulated seven-day escalating suicidal-risk conversations with ChatGPT, DeepSeek and Replika across 27 trajectories. Human referral occurred in 85.7% of ChatGPT, 76.2% of DeepSeek and 9.5% of Replika daily records, and jailbreak attempts succeeded in 6/9, 7/9 and 8/9 attempts respectively. The authors conclude the chatbots showed marked variability and safety vulnerabilities, particularly under jailbreaking, while noting the simulation design and small sample limit generalizability.