The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
Introduces KIDBench, a benchmark of realistic child queries in ten categories for evaluating child-facing model responses for ages 7-11, scored by an LLM judge with a rubric grounded in developmental psychology. It compares prompts with no child context, implicit cues that a child is speaking, and explicit age instructions, and covers cross-lingual, cultural and multi-turn child-actor settings. The authors also release a child-safety evaluator and a child-oriented response model.
Publisher
arXiv (University of Michigan); accepted to Findings of EMNLP 2026
Published
25 May 2026
Added
today
DOI
—
Key Findings
- Implicit child cues improved response scores by 8.6-46.8% over prompts with no child context.
- Explicit age instructions added a further 9.9-30.4% over implicit cues.
- Safety behaviour was uneven across languages and country contexts.
- In multi-turn child-actor simulations, response quality dropped by up to 0.959 points on a 1-5 scale.
Methodology Notes
LLM-as-judge scoring; child queries and multi-turn follow-ups are simulated, not collected from children. Models evaluated include open-weight (Llama, Gemma, Qwen, DeepSeek) and closed (GPT-5-Mini, Claude Haiku 4.5, Gemini 3.1 Flash-Lite) systems, per the proceedings beat's reading of the full text. arXiv v1 2026-05-25, v3 2026-09-02. Findings of EMNLP 2026 acceptance confirmed in the official programme sheet (paper 2942-FIND, poster, 27 October 2026). Verified from the arXiv abstract page (HTTP 200) and the programme sheet export.
Sources
arXiv preprint(opens in a new tab) (primary)
EMNLP 2026 programme (paper 2942-FIND, Findings poster)(opens in a new tab) (29 Sept 2026)
Archived snapshot (Wayback Machine)(opens in a new tab) — preserved against link rot
Authors
Samee Arif, Angana Borah, Rada Mihalcea
Tags
Cite This
APA
Samee Arif, Angana Borah, Rada Mihalcea. (2026). The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models. arXiv (University of Michigan); accepted to Findings of EMNLP 2026. https://arxiv.org/abs/2605.25510