Skip to main content
Lab publication Credible — Major labs, established NGOs, reputable named-author preprints

Continuing To Build Upon Our Safety Priorities

A first-party safety update from Character.AI describing safeguards in operation on its platform as of September 2026. It states that self-harm safeguards consider the surrounding conversation, including signals that emerge gradually over long-running chats rather than in any one turn; announces a partnership with the nonprofit Koko for free self-guided emotional-support tools and proactive routing to ThroughLine's country-specific crisis resources; describes an in-house age-estimation model and age-assurance technology; and lists a Parental Insights tool run through k-ID, moderation notifications with appeals, user blocking, and memberships of the Internet Watch Foundation and StopNCII. It reaffirms the earlier removal of open-ended chat with Characters for users under 18.

Publisher

Character.AI (Character Technologies)

Published

3 Sept 2026

Added

today

DOI

Key Findings

  • Self-harm safeguards 'are designed to consider the surrounding conversation, including signals that emerge gradually over long-running chats rather than in any one turn'
  • Crisis response now pairs a Koko partnership (free self-guided emotional-support tools) with routing to ThroughLine's country-specific crisis directory
  • The company states it has built its own age-assurance technology and an in-house age-estimation model, described as one of the most important systems it operates
  • Parental Insights runs through a partnership with k-ID, sending a weekly summary to a linked parent; creators are notified when content is moderated and can appeal; users can block another user and that user's Characters, Posts, Voices and Scenes
  • The post carries no evaluation results, prevalence figures or method, and does not mention the AI Safety Lab announced in October 2025

Methodology Notes

Blog post on blog.character.ai (Ghost platform; article:published_time 2026-09-03T16:00:32Z), about 750 words, read in full. A first-party description of shipped safeguards with no numbers, no evaluation data and no external validation; treat every claim as the operator's own account. Logged as a dated statement of practice on the most-cited companion platform rather than as evidence of effectiveness.

Tags

character-aicompanion-appssafety-updatecrisis-routingthroughlinekokoage-assurancefirst-party

Cite This

APA

Character.AI (Character Technologies). (2026). Continuing To Build Upon Our Safety Priorities. https://blog.character.ai/continuing-to-build-upon-our-safety-priorities/