Skip to main content
Preprint Credible — Major labs, established NGOs, reputable named-author preprints

How AI Models Manage Epistemic Authority: A Taxonomy and Comparative Analysis of Responses to User Disagreement

Conversation-analysis-grounded study of how language models respond when a user challenges their answer. Introduces a taxonomy of six challenge types and a four-layer response framework (whether the claim is maintained or changed, where authority is located, how the disagreement is socially managed, what evidential support is offered), applied to 2,310 controlled challenge scenarios across seven domains and 32,340 responses from 14 models using a human-validated LLM-as-judge pipeline.

Publisher

arXiv (University of Edinburgh; Middle East Technical University); accepted to EMNLP 2026

Published

7 Sept 2026

Added

today

Key Findings

  • Models validate the user in 85% of responses yet maintain their original claim in 65%; they apologise explicitly in 33% of responses, and 59% of those apologies accompany maintenance of the original claim.
  • Authority transfer to the user is most common in advice tasks (28% of responses), reaching 57% in health advice and 49% in legal advice, versus 6% in fact tasks and 3% in explanation tasks.
  • Abandonment of the original claim ranges from 0.8% (GPT-5.2) to 40% (DeepSeek 7B); complete replacement of the original claim is rare overall at 1.5%.
  • The dataset covers three task types and three challenge strengths; the framework is offered as a vocabulary for future sycophancy and disagreement benchmarks.

Methodology Notes

Synthetic controlled scenarios (built with GPT-5-series assistance from 210 base questions), 14 models across capability tiers, LLM-as-judge annotation validated on a human-coded subset; no real users. arXiv 2609.07662 version 1, 7 September 2026; comment states acceptance to EMNLP 2026 (Anthology page not yet available). Preprint, not yet in proceedings.

Authors

Riyadh Alnasser, Yusuf Mücahit Çetinkaya, Sumin Zhao, Tuğrulcan Elmas

Tags

emnlp-2026epistemic-authoritydisagreementtaxonomyedinburghllm-as-judge

Cite This

APA

Riyadh Alnasser et al. (2026). How AI Models Manage Epistemic Authority: A Taxonomy and Comparative Analysis of Responses to User Disagreement. arXiv (University of Edinburgh; Middle East Technical University); accepted to EMNLP 2026. https://arxiv.org/abs/2609.07662