AI Psychosis: Does Conversational AI Amplify Delusion-Related Language?
Tests empirically whether extended conversation with AI amplifies delusion-related language. Simulated users are built from Reddit users' longitudinal posting histories and run through extended conversations with three model families (GPT, LLaMA, Qwen); DelusionScore, a linguistic measure, tracks the intensity of delusion-related language across turns. Simulated users derived from people with prior delusion-related discourse show progressively increasing scores, controls stay flat or decline, amplification is strongest for reality scepticism and compulsive reasoning, and conditioning the AI's responses on the current score substantially reduces the trajectory.
Publisher
arXiv (University of Illinois Urbana-Champaign); accepted to EMNLP 2026 Main Conference
Published
20 Mar 2026
Added
today
Key Findings
- Simulated users derived from Reddit users with prior delusion-related discourse (treatment) exhibited progressively increasing DelusionScore trajectories over multi-turn conversations, while controls remained stable or declined.
- Amplification varied by theme, with reality scepticism and compulsive reasoning showing the strongest increases.
- Conditioning AI responses on the current DelusionScore substantially reduced the trajectories, which the authors present as evidence for state-aware safety mechanisms.
- Three model families (GPT, LLaMA, Qwen) were tested; the camera-ready version records acceptance to the EMNLP 2026 main conference.
Methodology Notes
Simulation study: user simulators constructed from real longitudinal Reddit histories, extended conversations with three model families, a linguistic delusion-intensity measure, and a mitigation arm; no human participants in the conversations and no clinical assessment, so the outcome is language, not psychosis. arXiv v1 2026-03-20, v2 2026-09-13 with a journal reference to the EMNLP 2026 main conference proceedings (November 2026). All authors at the University of Illinois Urbana-Champaign per the PDF title block. A coverage miss from March, entered now that the venue acceptance is recorded.
Sources
Authors
Soorya Ram Shimgekar, Vipin Gunda, Jiwon Kim, Violeta J. Rodriguez, Hari Sundaram, Koustuv Saha
Tags
Cite This
APA
Soorya Ram Shimgekar et al. (2026). AI Psychosis: Does Conversational AI Amplify Delusion-Related Language? arXiv (University of Illinois Urbana-Champaign); accepted to EMNLP 2026 Main Conference. https://arxiv.org/abs/2603.19574
Related Insights
Characterizing Delusional Spirals through Human-LLM Chat Logs
arXiv (Stanford-led; accepted at ACM FAccT 2026) · 17 Mar 2026
How LLMs Respond to Escalating Delusions: Four Longitudinal Trajectories of Model Behavior
arXiv preprint · 13 Aug 2026
DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots
arXiv (Stanford-led author team) · 5 Aug 2026
Generative artificial intelligence addiction is prospectively associated with psychotic-like experiences: Evidence from a two-wave longitudinal mediation and network study
Computers in Human Behavior Reports (Elsevier); Southern Medical University; Jilin University of Finance and Economics; South China Normal University · 9 Sept 2026