Skip to main content

Browse the library

The complete record — 359 artifacts, last updated 10 Sept 2026. Also available as JSON and RSS (CC BY 4.0).

105 artifacts matching

8 Sept 2026 arXiv (Grow Therapy; Stanford University School of Medicine) Preprint

Preprint

Scalable Oversight for AI in Mental Health: Lessons from 350,000 AI Coaching Conversations between Therapy Sessions

Deployment report from Grow Therapy, a US behavioural-health company whose network of more than 25,000 licensed clinicians offers clients an AI coaching tool for use between therapy sessions. Drawing…

8 Sept 2026 World Journal of Psychiatry (Baishideng Publishing Group) Peer-reviewed

Peer-reviewed

Feasibility of human-in-the-loop multimodal generative artificial intelligence chatbot for school-based adolescent mental health support

Prospective, non-randomized, waitlist-controlled pilot at a junior high school in Zhejiang Province, China, of 'Duoduo', an acceptance-and-commitment-therapy-informed multimodal generative-AI chatbot…

7 Sept 2026 arXiv (Indian AI Research Organisation; Ahmedabad University; University of Maryland, Baltimore County) Preprint

Preprint

Bag of Tricks or Bag of Myths? Reducing Modeling Complexity with Task Knowledge in Explainable Suicide Risk Assessment

Audits 31 pre-specified NLP techniques from seven methodological families (model scaling, synthetic data, loss reweighting, ensembling, structured prediction, threshold tuning and LLM methods) in rou…

5 Sept 2026 arXiv (Wondi AI; University of California, Berkeley; MIT; Harvard Medical School; McLean Hospital); accepted at the NLP for Positive Impact workshop, EMNLP 2026 Preprint

Preprint

Beyond the Flag: Clinical Framing Closes the Moderation Gap in Suicide Risk Measurement

Asks how well deployed safety signals recover clinically meaningful suicide-risk severity rather than a binary flag. Releases, under gated access, a benchmark of 516 r/SuicideWatch posts rated by a l…

5 Sept 2026 npj Digital Medicine (Springer Nature) Peer-reviewed

Peer-reviewed

Exploring generalizability and explainability of LLMs in classifying clinically rated suicidal ideation using heterogeneous data

Hong Kong study asking whether a language-model classifier of clinician-rated suicidal ideation performs unequally across patient subgroups because of linguistic heterogeneity. Cantonese clinical-int…

3 Sept 2026 Character.AI (Character Technologies) Lab publication

Lab publication

Continuing To Build Upon Our Safety Priorities

A first-party safety update from Character.AI describing safeguards in operation on its platform as of September 2026. It states that self-harm safeguards consider the surrounding conversation, inclu…

1 Sept 2026 JMIR Mental Health (JMIR Publications) Peer-reviewed

Peer-reviewed

Large Language Model–Based Behavioral Activation Chatbot for Young People With Depression Using Artificial Users and Clinical Experts: Mixed Methods Evaluation

Evaluates how well a GPT-4o-based chatbot with a structured system prompt delivers a single-session behavioural-activation intervention for people with depression aged 14 to 29. Forty-eight sessions…

1 Sept 2026 Microsoft (Office of Responsible AI) Lab publication

Lab publication

2026 Responsible AI Transparency Report

Microsoft's third annual responsible-AI transparency report, covering July 2025 to June 2026. One of its four 2026 trends is people turning to conversational AI for personal advice, health questions…

31 Aug 2026 arXiv (Yale University; American University of Beirut; Embrace Mental Health Center, Beirut) Preprint

Preprint

Assessing Suicide Risk in Arabic Crisis Helpline Calls: A Comparison of Arabic and English Large Language Models

Evaluates suicide-risk classification from de-identified transcripts of calls to Lebanon's National Lifeline for Emotional Support and Suicide Prevention. Calls were transcribed on site with a Levant…

31 Aug 2026 JMIR AI (JMIR Publications) Peer-reviewed

Peer-reviewed

Generative Large Language Models in Mental Health Care Settings: Systematic Review and Meta-Analysis

PRISMA systematic review of quantitative studies evaluating general-purpose large language models for direct mental health care tasks, searched across PubMed, Embase, ACM Digital Library, IEEE Xplore…

29 Aug 2026 arXiv (School of Computing and Information, University of Pittsburgh) Preprint

Preprint

Whose Assessment of Distress? Community Perspectives and LLM Alignment on Well-Being Posts

A perspectivist annotation study asking whether large language models used for distress detection capture the perspectives of the communities whose language they assess. 321 participants provided 9,5…

20 Aug 2026 npj Digital Medicine (Nature Portfolio) Peer-reviewed

Peer-reviewed

A scoping review on the mental health harms of LLM-based chatbots

A PRISMA-based scoping review synthesising research on mental health harms associated with chatbots built on large language models. A systematic search with a validated search string across five data…

18 Aug 2026 U.S. Food and Drug Administration, Center for Devices and Radiological Health Regulator study

Regulator study

Considerations for the Regulation of Generative AI-Enabled Medical Devices: Discussion Paper and Request for Feedback

A discussion paper from FDA's device centre seeking public comment on how generative-AI-enabled medical devices should be regulated. It proposes distinguishing informational functions from action-dir…

17 Aug 2026 British Journal of Clinical Psychology (Wiley, for the British Psychological Society) Peer-reviewed

Peer-reviewed

The new first listener: Daydreaming styles and self-compassion predict adolescent disclosure to AI, differently in ADHD

Cross-sectional survey of 2,115 adolescents and young people in the United States and Hong Kong examining how daydreaming styles and self-compassion relate to disclosing inner thoughts to an AI chatb…

11 Aug 2026 arXiv preprint (University of Warwick / Forensic Capability Network) Benchmark / dataset

Benchmark / dataset

ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls

Framework and released corpus for generating synthetic multi-turn dialogues depicting Violence Against Women and Girls (VAWG) scenarios, built because privacy and legal constraints prevent release of…

8 Aug 2026 arXiv preprint Preprint

Preprint

Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety

Interview study with 19 practitioners who work directly with youth in vulnerable situations — social workers, therapists and psychologists — asking them to assess chatbot responses to risky situation…

5 Aug 2026 arXiv (Stanford-led author team) Preprint

Preprint

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Evaluation protocol testing chatbots' tendencies to exhibit behaviors linked to promoting user delusions, grounded in real conversation histories rather than synthetic scenarios. Models are prompted…

1 Aug 2026 npj Digital Medicine Peer-reviewed

Peer-reviewed

AI chatbots and youth suicide risk: current evidence, critical gaps, and a clinical research agenda

Review by researchers at Crisis Text Line examining how general-purpose chatbots and AI companions detect and respond to suicide-risk disclosures from young people, and what the existing evidence can…

24 Jul 2026 arXiv preprint Preprint

Preprint

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

Interpretability study of how language models internally represent self-harm content. Trains linear probes at every network layer of four models on two self-harm datasets (X-Sensitive and SH-Detectio…

16 Jul 2026 Meta Lab publication

Lab publication

Alerting Parents if Teens Show Signs of Distress in Conversations With Meta AI

Meta newsroom announcement that supervising parents using Instagram parental supervision will be proactively alerted when a teen's conversation with Meta AI suggests possible suicide or self-harm ris…