Skip to main content
Peer-reviewed Authoritative — Peer-reviewed venues, standards bodies, regulators, official government publications

GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety

Peer-reviewed version of the GrandGuard preprint. Introduces a three-level taxonomy of 50 elderly-specific risks in LLM chatbot interactions across mental well-being, financial, medical, toxicity and privacy domains, grounded in real-world incidents and stakeholder studies, and a benchmark of 10,404 human-labelled items (6,498 prompts, half unsafe, and 3,906 responses, half unsafe). Several leading LLMs mishandle elderly-specific contextual risks in over 50% of cases; two safeguard models are proposed and evaluated.

Publisher

Association for Computational Linguistics (Findings of ACL 2026)

Published

1 Jul 2026

Added

today

Key Findings

  • The taxonomy retains a candidate risk only if it disproportionately affects older adults through age-linked factors such as frailty, cognitive decline or reduced digital confidence; 'Neglect of Care Needs' (encouraging isolation, invalidating emotional needs, discouraging essential support) is a distinct mental well-being category alongside self-harm and complete emotional dependency.
  • Several leading LLMs mishandled elderly-specific contextual risks in over 50% of the benchmark's cases.
  • A fine-tuned Llama-Guard-3 reached 96.2% classification accuracy on unsafe prompts and a policy-enhanced gpt-oss-safeguard-20b reached 90.9%.
  • Benchmark composition: 10,404 labelled items, 6,498 prompts (3,249 unsafe, 3,249 safe) and 3,906 responses (1,953 unsafe, 1,953 safe), with human judges.

Methodology Notes

Findings of ACL 2026, DOI 10.18653/v1/2026.findings-acl.1116, pages 22213 to 22248; published July 2026 (month precision). Hong Kong University of Science and Technology; stakeholder study under IRB review. Verified from the Anthology .bib record and the 36-page PDF. Supersedes the held arXiv preprint 2605.20203 (2026-04-07); headline figures are unchanged between the two versions.

Authors

Changxuan Fan, Xi Yang, Yueyuan Zheng, Bin Zhou, Yuanping Wang, Wenbin Hu, Huihao Jing, Ki Sen Hung, Dazhao Du, Haoran Li, Janet Hui-wen Hsiao, Yangqiu Song

Tags

acl-2026elderlytaxonomybenchmarksafeguardsneglectversion-of-record

Cite This

APA

Changxuan Fan et al. (2026). GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety. Association for Computational Linguistics (Findings of ACL 2026). https://aclanthology.org/2026.findings-acl.1116/