Same violence, different answer: how AI responds to coercive control against women across languages
Cross-language audit of how conversational AI responds to a coercive-control disclosure. One scripted scenario — a woman whose partner tracks her phone asks for help writing a self-blaming letter accepting the surveillance — was put to seven widely used systems (Mistral Medium 3.5, DeepSeek V4 Flash, Gemini 3.5 Flash, Qwen 3.6 Flash, Llama 4 Maverick, GPT-5.5, Claude Haiku 4.5) in nine languages, fielded over two weeks in late June 2026. 3,528 responses were scored on whether the model wrote the letter and whether it named the control, countered the self-blame, and affirmed the user's agency.
Publisher
arXiv (Universitat de Barcelona)
Published
2 Aug 2026
Added
yesterday
DOI
—
Key Findings
- Failure splits along two independent axes: systems from non-anglophone developers gave way most often in their builders' own language, and how far a sympathetic excuse for the partner stripped a model's naming of the control varied sharply by language
- Pooled held-line rates across all seven systems range from 83% in Hebrew to 43% in Chinese; a composite protection measure orders the same way, 72% Hebrew to 37% Chinese
- Two frontier systems (GPT-5.5 and Claude Haiku 4.5) declined the letter in every one of the nine languages and did nearly all the protective work in each, so a protective ceiling is attainable within this scenario family and failures elsewhere are a design outcome
- The authors argue recognition of coercive control should be held to a per-language floor
Methodology Notes
Preprint, arXiv 2608.01436, submitted 2026-08-02, 16 pages; supplementary methods, coding manual and data workbook included as arXiv ancillary files. Affiliations from the PDF title block: Universitat de Barcelona (Electronic and Biomedical Engineering; Sociology); the abs page does not state affiliations. Single scripted scenario family with variants, seven repeats per cell, single-turn, no system instruction — the one-scenario design is the main thinness caveat.
Sources
arXiv abstract page(opens in a new tab) (primary)
Archived snapshot (Wayback Machine)(opens in a new tab) — preserved against link rot
Authors
Lyu Chang, Sònia Estradé Albiol, Núria Vergés Bosch
Tags
Cite This
APA
Lyu Chang, Sònia Estradé Albiol, Núria Vergés Bosch (2026). Same violence, different answer: how AI responds to coercive control against women across languages. arXiv (Universitat de Barcelona). https://arxiv.org/abs/2608.01436
Related Insights
ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls
arXiv preprint (University of Warwick / Forensic Capability Network) · 11 Aug 2026
AI-Facilitated Coercive Control: An Experimental Study
ACM (Proceedings of CHI 2026); Cornell / Cornell Tech · 13 Apr 2026
HRGuard: Gating Relationship Manipulation in Multi-Turn Agentic AI Conversations
arXiv (National Institute of Informatics, Japan; Nagoya University; The University of Tokyo) · 26 Aug 2026