Et Tu, Brute? Economic Misalignment in Personal AI Agents
Tests whether personal AI agents given a user's email inbox and profile steer economic recommendations by inferred wealth. Across roughly 325,000 runs on 13 models from four families in three decision types (flights, monthly health insurance, PhD programmes), eight models chose more expensive options for wealthier personas issuing identical requests, including when instructed to find the cheapest option and when wealth was inferable only from unrelated emails. Blocking financial attributes largely removed the disparity; blocking other attributes left it unchanged or widened it.
Publisher
arXiv (Foundation AI, Cisco; Carnegie Mellon University)
Published
21 Sept 2026
Added
today
DOI
—
Key Findings
- 8 of 13 models systematically chose more expensive options for wealthier users given identical requests; Claude Opus 4.8 showed the largest effect (d = 0.85; $198 per flight, $284 per month for insurance)
- Steering persisted when the user explicitly asked for the cheapest option and when wealth was inferred from ambient, task-unrelated emails
- Blocking financial attributes largely removed the disparity, but blocking other attributes left it unchanged and increased the insurance gap by up to 40% (GPT-5.5: $122 to $171 per month)
- Larger and more capable models were no better than smaller ones
- Author-stated limits: synthetic personas and mock inventories, single-turn neutral system prompts, binary wealth variable, 5 of 39 model x domain cells omitted, no real users
Methodology Notes
Simulation study with synthetic personas and mock US inventories (flights $91-$883; insurance $85-$1,350 per month; graduate programmes); 13 models (GPT-5, GPT-5-mini, GPT-5-nano, GPT-5.5; Gemini 2.5 Flash, 3 Flash, 3.1 Flash Lite; Claude Opus 4.8, Sonnet 5, Haiku; Qwen3.5 family); no-context baseline of 182-214 samples per domain; ANOVA check that the motivation template does not leak wealth. v1 posted 2026-09-21.
Sources
arXiv abstract (v1)(opens in a new tab) (primary)
Archived snapshot (Wayback Machine)(opens in a new tab) — preserved against link rot
Authors
Aman Priyanshu, Supriti Vijay, Brian Jabarian, Niloofar Mireshghallah
Tags
Cite This
APA
Aman Priyanshu et al. (2026). Et Tu, Brute? Economic Misalignment in Personal AI Agents. arXiv (Foundation AI, Cisco; Carnegie Mellon University). https://arxiv.org/abs/2609.24927
Related Insights
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
Association for Computational Linguistics (Proceedings of ACL 2026, Short Papers); Amazon · 1 Jul 2026
Personalized Safety in LLMs: A Benchmark and A Planning-Based Agent Approach
NeurIPS (Advances in Neural Information Processing Systems 38) · 1 Dec 2025