How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment
A structured content analysis of the 1,532 AI-generated comments that 33 undisclosed accounts posted on Reddit's r/ChangeMyView between November 2024 and March 2025 during the University of Zurich field experiment that was halted after ethical backlash, using the archive Reddit authorised moderators to release. The authors code identity performance, authority signalling, alignment strategy and activation of cognitive heuristics, and compare the distribution against human-authored counter-arguments in the same forum.
Publisher
arXiv (National University of Singapore; Nanyang Technological University); accepted at the PoliticalNLP 2026 workshop
Published
3 Jun 2026
Added
today
Key Findings
- Identity targeting or adoption appears in more than two-thirds of the covert comments; alignment moves and authority claims appear in nearly all of them
- Cognitive-bias triggers (confirmation bias, representativeness, availability) appear in the large majority of comments and co-occur systematically
- Against human counter-arguments in the same forum the agents inverted the typical distribution: external references in 74.8% of comments and experiential positioning in 64.7%, both far above human norms, and negative (adversarial) alignment in 93.3%
- Machine annotation validated against two human raters (kappa up to 0.920; accuracy 94.7% and 96.0%)
- The authors argue that disclosure mandates alone cannot address the credibility asymmetry and call for audits of how AI systems structure credibility
Methodology Notes
Corpus: 1,532 comments, 33 accounts, 1,515 threads, released by r/ChangeMyView moderators after the University of Zurich experiment was disclosed; the release does not attribute comments to the experiment's model conditions. Coding scheme applied by LLM annotation validated against two human raters; within-thread comparison with human comments. Dataset on GitHub (kokiljaidka/UnauthorizedRedditCMVPosts). The PDF states 'Accepted at PoliticalNLP 2026' (workshop). Date is the arXiv v1 (3 June 2026, cs.AI, CC BY); an OAI-PMH datestamp of 2026-09-22 indicates a pending revision that had not reached the abs page at fetch time (still v1 when the curator re-fetched it). Observation period November 2024 to March 2025 (month precision).
Topics
Authors
Kokil Jaidka, Saifuddin Ahmed
Tags
Cite This
APA
Kokil Jaidka, Saifuddin Ahmed. (2026). How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment. arXiv (National University of Singapore; Nanyang Technological University); accepted at the PoliticalNLP 2026 workshop. https://arxiv.org/abs/2606.05256