A light-touch AI literacy intervention helps protect against AI political persuasion
Two pre-registered-style experiments (total N=3,208 US adults) in which participants conversed with a large language model instructed to shift their views on political topics, with or without a brief warning that LLMs can be prompted to persuade and may present information selectively. The warning reduced belief change by roughly half relative to control and did not significantly reduce trust in generative AI more broadly.
Publisher
arXiv (Carnegie Mellon University, Department of Social and Decision Sciences; Cornell University)
Published
14 Sept 2026
Added
today
Key Findings
- A one-sentence warning that the model may be prompted to persuade and may present information selectively reduced belief change by 48.1% (95% CI 36.8% to 59.5%) relative to the no-warning control
- The warning did not significantly reduce general trust in generative AI
- The effect held across two experiments and multiple political topics with a combined N=3,208
Methodology Notes
Two online experiments with US participants conversing with an LLM instructed to persuade; between-subjects warning manipulation; the paper states it has not undergone peer review. Preprint (v1, posted 2026-09-14; announced in the 15 September cs.HC listing). Curator read the abstract page and the PDF first page on 2026-09-17.
Sources
Authors
Orchinik, Reed, Rand, David
Tags
Cite This
APA
Orchinik, Reed, Rand, David. (2026). A light-touch AI literacy intervention helps protect against AI political persuasion. arXiv (Carnegie Mellon University, Department of Social and Decision Sciences; Cornell University). https://arxiv.org/abs/2609.16432