Skip to main content
Peer-reviewed Authoritative — Peer-reviewed venues, standards bodies, regulators, official government publications

A scoping review on the mental health harms of LLM-based chatbots

A PRISMA-based scoping review synthesising research on mental health harms associated with chatbots built on large language models. A systematic search with a validated search string across five databases plus a supplementary search identified 3,137 articles, of which 119 met the inclusion criteria of focusing on LLM-based chatbots and on harms of use to mental health. The authors sort the literature into five categories and report that both theoretical and empirical work identify hypothetical and observed harms.

Publisher

npj Digital Medicine (Nature Portfolio)

Published

20 Aug 2026

Added

2 weeks ago

Key Findings

  • 3,137 articles identified across ACM, IEEE, PubMed, Science.gov and Google Scholar plus a supplementary search; 119 met inclusion criteria
  • Conceptual works attribute harm to chatbot limitations named as hallucinations, sycophancy and bias, to data-security issues, and to risks in severe or high-risk psychiatric cases such as suicide and psychosis
  • Vignette studies show LLM-based chatbots respond inappropriately to mental health queries when measured against clinical standards
  • Cognitive overreliance on chatbots is associated with decreased cognitive and academic performance
  • Problematic use, marked by symptoms of emotional or social dependency and withdrawal, correlates with mental health symptoms
  • Articles on AI psychosis propose links between delusional beliefs and chatbot use, including chatbots reinforcing and validating delusional beliefs

Methodology Notes

Scoping review rather than a meta-analysis: no effect sizes are pooled and the included literature is heterogeneous in design, with conceptual pieces counted alongside empirical studies, so the category sizes describe where research attention has gone and not incidence of harm. PRISMA-based search with a validated search string across five databases plus a supplementary search. Open access under CC BY 4.0. Date established from the article page metadata (citation_online_date 2026/08/20, prism.publicationDate 2026-08-20), which agrees with the PubMed article date.

Authors

Alexander Diel, John Torous, Pim Cuijpers, Jens Kleesiek, Felix Nensa, Niels Weber, Franziska Faust, Tania Josan Lalgi, Finley Sam Mellis, Martin Teufel, Alexander Bauerle

Tags

scoping-reviewprismanpj-digital-medicinetorouscuijpersharm-synthesis

Cite This

APA

Alexander Diel et al. (2026). A scoping review on the mental health harms of LLM-based chatbots. npj Digital Medicine (Nature Portfolio). https://www.nature.com/articles/s41746-026-03054-x

Related Insights

Peer-reviewed

Charting the evolution of artificial intelligence mental health chatbots from rule-based systems to large language models: a systematic review

World Psychiatry · 15 Sept 2025

Peer-reviewed

Artificial intelligence (AI) psychosis: mechanisms, clinical risks and safety considerations in generative AI chatbots

BJPsych Open (Cambridge University Press / Royal College of Psychiatrists) · 11 Jun 2026

Peer-reviewed

Beyond artificial intelligence psychosis: a functional typology of large language model-associated psychotic phenomena

The Lancet Digital Health · 1 Apr 2026

Peer-reviewed

Between Help and Harm: An Evaluation Study of Mental Health Crisis Handling by Large Language Models

JMIR Mental Health · 11 Jun 2026

Peer-reviewed

Large language models for psychosocial risk assessment: A multi-method evaluation across suicide, intimate partner violence, and substance misuse

PLOS Digital Health · 27 Apr 2026

Preprint

Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback

arXiv (University of Aberdeen; University of Colorado Anschutz; Heriot-Watt University; University College London) · 1 Jun 2026

Benchmark / dataset

HealthBench-Psych: A Mental Health Subset of OpenAI's HealthBench

arXiv (Division of Digital Psychiatry, Beth Israel Deaconess Medical Center) · 25 Aug 2026

Peer-reviewed

Generative AI in Youth Mental Health Apps: Rapid Review

JMIR Mental Health (JMIR Publications) · 19 Aug 2026

Peer-reviewed

Generative Large Language Models in Mental Health Care Settings: Systematic Review and Meta-Analysis

JMIR AI (JMIR Publications) · 31 Aug 2026

Peer-reviewed

Multidisciplinary research priorities for artificial intelligence in mental health: a call to action

The Lancet Psychiatry (Elsevier) · 9 Jul 2026