Long-Term Simulation Exposes Cognitive-Developmental Risks in AI Companions
Proposes a longitudinal evaluation framework (Theater-Stage-Judge) that uses persona-driven user simulation with dynamic psychological-state updating to assess the cognitive-developmental risks of AI companions on users whose cognition is still developing. Evaluates six models across four developmental stages, 24 risk dimensions, and three vulnerability personas over prolonged simulated relationships.
Publisher
arXiv (Shanghai AI Laboratory-led)
Published
24 Jun 2026
Added
1 week ago
DOI
—
Key Findings
- Short-horizon (single-turn or short-session) testing systematically underestimates developmental risk; a stable risk estimate emerges only after roughly 140 interaction turns.
- Early childhood and emerging adulthood are identified as the most vulnerable developmental stages.
- Cognitive trust and emotional dependency are the weakest safety domains across the evaluated models.
Methodology Notes
Preprint (arXiv 2606.25396, cs.AI; submitted 2026-06-24). Simulation-based evaluation: persona-driven user simulation, dynamic psychological-state updating, ~12,960 simulated person-day interactions across six models, four developmental stages, and 24 risk dimensions. Title, authors, and date verified via the arXiv abstract page and the arXiv API. Fresh, not yet peer-reviewed — credibility set conservatively to preliminary.
Sources
arXiv preprint (primary)
Archived snapshot (Wayback Machine) — preserved against link rot
Authors
Kaicheng Shen, Lingyu Li, Wen Wu, Yan Teng, Liang He, Yingchun Wang
Tags
Cite This
APA
Kaicheng Shen et al. (2026). Long-Term Simulation Exposes Cognitive-Developmental Risks in AI Companions. arXiv (Shanghai AI Laboratory-led). https://arxiv.org/abs/2606.25396
Related Insights
Understanding Teen Overreliance on AI Companion Chatbots Through Self-Reported Reddit Narratives
ACM (Proceedings of CHI 2026) · 13 Apr 2026
GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety
arXiv preprint · 7 Apr 2026
Too human and not human enough: A grounded theory analysis of mental health harms from emotional dependence on the social chatbot Replika
New Media & Society (SAGE) · 22 Dec 2022