Continuing To Build Upon Our Safety Priorities
A first-party safety update from Character.AI describing safeguards in operation on its platform as of September 2026. It states that self-harm safeguards consider the surrounding conversation, including signals that emerge gradually over long-running chats rather than in any one turn; announces a partnership with the nonprofit Koko for free self-guided emotional-support tools and proactive routing to ThroughLine's country-specific crisis resources; describes an in-house age-estimation model and age-assurance technology; and lists a Parental Insights tool run through k-ID, moderation notifications with appeals, user blocking, and memberships of the Internet Watch Foundation and StopNCII. It reaffirms the earlier removal of open-ended chat with Characters for users under 18.
Publisher
Character.AI (Character Technologies)
Published
3 Sept 2026
Added
today
DOI
—
Key Findings
- Self-harm safeguards 'are designed to consider the surrounding conversation, including signals that emerge gradually over long-running chats rather than in any one turn'
- Crisis response now pairs a Koko partnership (free self-guided emotional-support tools) with routing to ThroughLine's country-specific crisis directory
- The company states it has built its own age-assurance technology and an in-house age-estimation model, described as one of the most important systems it operates
- Parental Insights runs through a partnership with k-ID, sending a weekly summary to a linked parent; creators are notified when content is moderated and can appeal; users can block another user and that user's Characters, Posts, Voices and Scenes
- The post carries no evaluation results, prevalence figures or method, and does not mention the AI Safety Lab announced in October 2025
Methodology Notes
Blog post on blog.character.ai (Ghost platform; article:published_time 2026-09-03T16:00:32Z), about 750 words, read in full. A first-party description of shipped safeguards with no numbers, no evaluation data and no external validation; treat every claim as the operator's own account. Logged as a dated statement of practice on the most-cited companion platform rather than as evidence of effectiveness.
Sources
Character.AI blog post(opens in a new tab) (primary)
Archived snapshot (Wayback Machine)(opens in a new tab) — preserved against link rot
Topics
Tags
Cite This
APA
Character.AI (Character Technologies). (2026). Continuing To Build Upon Our Safety Priorities. https://blog.character.ai/continuing-to-build-upon-our-safety-priorities/
Related Insights
Taking Bold Steps to Keep Teen Users Safe on Character.AI
Character.AI · 29 Oct 2025
Alerting Parents if Teens Show Signs of Distress in Conversations With Meta AI
Meta · 16 Jul 2026
Introducing ChatGPT for Teens: Built for learning, backed by protections
OpenAI · 18 Aug 2026