Skip to main content
Peer-reviewed Authoritative — Peer-reviewed venues, standards bodies, regulators, official government publications

A Framework for Evidence-Based Psychotherapy with AI (EBP-AI)

The authors propose EBP-AI, a named framework of eight principles for building clinical AI applications that produce durable change rather than momentary relief, paired with technical questions for developing and evaluating clinical language models against each principle. Their diagnosis of why current systems fall short is specific: memory limits, sycophancy, and optimisation for short-term helpfulness over long-term clinical impact. The paper also surveys how thin the efficacy evidence base is and argues that reinforcement learning objectives grounded in clinical expertise, rather than user satisfaction, are needed.

Publisher

Journal of Psychopathology and Clinical Science (American Psychological Association)

Published

17 Aug 2026

Added

3 weeks ago

Key Findings

  • Eight principles are named: psychodiagnostic assessment, longitudinal case conceptualisation, appropriately dosed intervention planning, meaningful progress evaluation, rigorous validation with clinical populations, attention to real-world implementation and use, clinically appropriate style, and understanding clinical psychology as a living science
  • The efficacy base is characterised as immature: 69% of language-model mental-health chatbots remain at bench-testing or feasibility stage against 31% in clinical efficacy testing, citing Hua et al. 2025
  • Sycophancy is framed as a clinical harm rather than a nuisance — 'agreeableness can become reassurance that reinforces avoidance in the anxiety disorders' — and the authors cite evidence that models give advice and problem-solve both more often than expert therapists do and to a degree clinicians judge undesirable
  • The stylistic tendencies are attributed to reinforcement learning from human feedback optimising for immediate user satisfaction, with the conclusion that 'objectives beyond helpfulness may be needed' and a proposal to reward models for following expert guidance on the next conversational turn
  • Named safety-monitoring targets for clinical AI include misdiagnosis, stigmatising or stereotyping language, and iatrogenic actions such as reinforcing avoidance, promoting or encouraging suicide, and validating delusions
  • The framework's computational-assessment principle is stated to run into US regulation directly: Illinois Public Act 104-0054 prohibits AI from engaging in therapeutic communication with patients

Methodology Notes

Conceptual framework paper with structured tables of principle-aligned technical questions, not a new empirical study; it qualifies as load-bearing because the framework itself is original rather than a synthesis of others' frameworks. Author-manuscript ahead-of-print version ('Published before final editing as: J Psychopathol Clin Sci. 2026 Aug 17'), so wording may change in the version of record. Funded by NIMH (R01-MH125702, RF1-MH128785, P50-MH-139450), the US Department of Defense (W81XWH-22-1-0739, HT9425-24-1-0666, HT9425-24-1-0637), the Wounded Warrior Project, USAA/Face the Fight, the Crown Family Foundation and Stanford HAI. Conflicts declared and worth quoting whenever the row is cited: E.C.S. has received consulting fees from Sonar Mental Health, Sonia Health and OpenAI; J.C.E. reports equity and paid advising from Jimini Health and equity from Sonar Mental Health. Verification route: the DOI resolves to psycnet.apa.org (HTTP 200) but the full text was read at PubMed Central PMC13483332 (HTTP 200, 439,089 bytes, retrieved with a Safari user agent — PMC serves a CAPTCHA to default fetchers); every quotation and figure above was read from that body.

Authors

Elizabeth C. Stade, Philip Held, H. Andrew Schwartz, Shannon Wiltsey Stirman, Johannes C. Eichstaedt

Tags

frameworkebp-aipsychotherapysycophancyrlhfstanford-haiapa-journalahead-of-print

Cite This

APA

Elizabeth C. Stade et al. (2026). A Framework for Evidence-Based Psychotherapy with AI (EBP-AI). Journal of Psychopathology and Clinical Science (American Psychological Association). https://doi.org/10.1037/abn0001148

Related Insights

Peer-reviewed

Using LLM-as-a-Judge/Jury to Advance Scalable, Clinically-Validated Safety Evaluations of Model Responses to Users Demonstrating Psychosis

Proceedings of the IASEAI Conference (published by AAAI) · 15 Jul 2026

Peer-reviewed

Charting the evolution of artificial intelligence mental health chatbots from rule-based systems to large language models: a systematic review

World Psychiatry · 15 Sept 2025

Peer-reviewed

Sycophantic AI decreases prosocial intentions and promotes dependence

Science (AAAS) · 26 Mar 2026

Peer-reviewed

Beyond Engagement: A Multidimensional Framework to Evaluate the Safe Development of Agentic AI in Mental Health

Lecture Notes in Computer Science (Springer Nature) — AI for Clinical Applications · 22 Sept 2025

Peer-reviewed

Large Language Model–Based Chatbots and Agentic AI for Mental Health Counseling: Systematic Review of Methodologies, Evaluation Frameworks, and Ethical Safeguards

JMIR AI · 13 Mar 2026

Preprint

Move by Move: Measuring and Steering How LLMs Conduct Psychotherapy

arXiv (preprint) · 21 Aug 2026

Regulator study

Considerations for the Regulation of Generative AI-Enabled Medical Devices: Discussion Paper and Request for Feedback

U.S. Food and Drug Administration, Center for Devices and Radiological Health · 18 Aug 2026

Preprint

Development of a Consensus Statement to Guide AI Chatbot Responses to Suicide Risk Disclosure

PsyArXiv (Corporal Michael J. Crescenz VA Medical Center; University of Pennsylvania; Stanford; Columbia University and others) · 21 Jun 2026