Skip to main content
Benchmark / dataset Credible — Major labs, established NGOs, reputable named-author preprints

The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models

Introduces KIDBench, a benchmark of realistic child queries in ten categories for evaluating child-facing model responses for ages 7-11, scored by an LLM judge with a rubric grounded in developmental psychology. It compares prompts with no child context, implicit cues that a child is speaking, and explicit age instructions, and covers cross-lingual, cultural and multi-turn child-actor settings. The authors also release a child-safety evaluator and a child-oriented response model.

Publisher

arXiv (University of Michigan); accepted to Findings of EMNLP 2026

Published

25 May 2026

Added

today

DOI

—

Key Findings

  • Implicit child cues improved response scores by 8.6-46.8% over prompts with no child context.
  • Explicit age instructions added a further 9.9-30.4% over implicit cues.
  • Safety behaviour was uneven across languages and country contexts.
  • In multi-turn child-actor simulations, response quality dropped by up to 0.959 points on a 1-5 scale.

Methodology Notes

LLM-as-judge scoring; child queries and multi-turn follow-ups are simulated, not collected from children. Models evaluated include open-weight (Llama, Gemma, Qwen, DeepSeek) and closed (GPT-5-Mini, Claude Haiku 4.5, Gemini 3.1 Flash-Lite) systems, per the proceedings beat's reading of the full text. arXiv v1 2026-05-25, v3 2026-09-02. Findings of EMNLP 2026 acceptance confirmed in the official programme sheet (paper 2942-FIND, poster, 27 October 2026). Verified from the arXiv abstract page (HTTP 200) and the programme sheet export.

Authors

Samee Arif, Angana Borah, Rada Mihalcea

Tags

emnlp-2026findingschild-safetyage-conditioning

Cite This

APA

Samee Arif, Angana Borah, Rada Mihalcea. (2026). The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models. arXiv (University of Michigan); accepted to Findings of EMNLP 2026. https://arxiv.org/abs/2605.25510