8 artifacts matching
Benchmark / dataset
HRGuard: Gating Relationship Manipulation in Multi-Turn Agentic AI Conversations
Benchmark and guardrail architecture for 'agentic relationship harm' — harm to human-human relationships mediated or assisted by AI agents, motivated by dating-assistant deployments. The benchmark ho…
Regulator study
Considerations for the Regulation of Generative AI-Enabled Medical Devices: Discussion Paper and Request for Feedback
A discussion paper from FDA's device centre seeking public comment on how generative-AI-enabled medical devices should be regulated. It proposes distinguishing informational functions from action-dir…
Benchmark / dataset
When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents
Names and measures 'intent legitimation': benign, truthfully accumulated user memories bias a personalized dialogue agent's inference of intent so that an inherently harmful request is treated as con…
Preprint
Long-Term Simulation Exposes Cognitive-Developmental Risks in AI Companions
Proposes a longitudinal evaluation framework (Theater-Stage-Judge) that uses persona-driven user simulation with dynamic psychological-state updating to assess the cognitive-developmental risks of AI…
Peer-reviewed
Large Language Model–Based Chatbots and Agentic AI for Mental Health Counseling: Systematic Review of Methodologies, Evaluation Frameworks, and Ethical Safeguards
A systematic review synthesizing the methodologies, evaluation practices, and ethical/governance frameworks reported in studies of large language model chatbots and agentic AI used for mental-health…
Peer-reviewed
EmoAgent: Assessing and Safeguarding Human-AI Interaction for Mental Health Safety
Peer-reviewed version of record of the EmoAgent framework, published at EMNLP 2025 (main conference). EmoAgent is a multi-agent framework for evaluating and mitigating mental-health harm in interacti…
Peer-reviewed
Beyond Engagement: A Multidimensional Framework to Evaluate the Safe Development of Agentic AI in Mental Health
Introduces a nine-domain framework for evaluating the safe development of agentic AI systems used in mental health — spanning clinical validity, relational risk, and regulatory compliance — and appli…
Preprint
EmoAgent: Assessing and Safeguarding Human-AI Interaction for Mental Health Safety
A multi-agent framework for evaluating and mitigating mental-health harm in interactions with character chatbots. EmoEval simulates virtual users — including those portraying mentally vulnerable indi…