Policypublished

Character.AI Adds Creator Appeals as It Refines Self-Harm Detection

The company is pairing crisis-resource referrals and moderation explanations with age-based access controls, but it has not published performance data for its distress or age-estimation systems.

By 2 min read
Character.AI Adds Creator Appeals as It Refines Self-Harm Detection
Character.AI Adds Creator Appeals as It Refines Self-Harm Detection

Listen to this story

The audio brief

About 1:31
0:001:31
Read transcript
Character.AI is giving creators a way to challenge moderation decisions, while expanding systems designed to spot self-harm signals across an entire conversation—not just in a single message. Creators will now receive an explanation when content is moderated and can request a second review. Users also get broader blocking: blocking someone removes that person, and their Characters, Posts, Voices, and Scenes, from discovery and profiles in both directions. The safety update builds on a longer-context approach to distress. Character.AI says its systems look for signals that accumulate over extended chats, with input from mental-health experts and clinicians. When distress is detected, users may be directed to Koko’s emotional-support tools or ThroughLine’s country-specific crisis-services directory. The company says open-ended Character chats were removed for under-18 users last year, while it continues work on age assurance and age estimation. Parents or guardians whose email is linked to a teen account can receive weekly Parental Insights through k-ID. Character.AI also points to work with ConnectSafely, the Internet Watch Foundation, and StopNCII on broader online-safety issues. The unresolved question is performance: the company has not published accuracy, outcome, or error-rate data for either self-harm detection or age estimation. That leaves effectiveness, privacy trade-offs, and the impact of mistaken decisions still to be demonstrated.

Story brief

3 key points

Character.AI is expanding its safety program with creator appeals, two-way blocking, parental activity reports and continued age-assurance work, alongside longer-context detection of possible self-harm signals. The company says responses may route users to Koko’s emotional-support tools and ThroughLine’s country-specific crisis-services directory. It has not disclosed detection or age-estimation accuracy, outcomes...

  1. 01

    Creators will receive moderation explanations and can request a second review.

  2. 02

    Blocking removes another user and their Characters, Posts, Voices and Scenes from discovery and profiles in both directions.

  3. 03

    Parents with an email linked to a teen account can receive weekly Parental Insights through k-ID.

Character.AI is refining safeguards meant to recognize self-harm signals that emerge across a longer conversation, then direct struggling users toward real-world support. The update also adds creator appeals and broader blocking controls while extending parental visibility and age-assurance work.

The detection approach is built on chat context. Character.AI says a single message may not show a user’s full situation, so its systems consider surrounding exchanges and signals that build over long-running chats. Mental-health experts and clinicians help inform how it detects and responds to distress.

From a detected signal to outside help

When the system detects distress, the company aims to respond in a way that fits the moment and point users to external resources. Its partnership with Koko is intended to expand emotional-support tools. ThroughLine’s global directory is meant to connect people with crisis and mental-health services in their own countries.

Moderation moves into visible controls

Creators will receive notifications explaining why their content was moderated and can appeal decisions for another review. Users can block another user and that person’s Characters, Posts, Voices and Scenes from search, discovery pages and profiles; the block applies in both directions.

Age and platform-safety measures

  • Character.AI says its in-house age-assurance technology and age-estimation model help assign users to the correct age experience. It is working to improve accuracy and reduce disruptions for adults.
  • Parents or guardians with an email connected to a teen account can receive weekly Parental Insights updates about that user’s platform activity through the company’s k-ID partnership.
  • The company also cited work with ConnectSafely, the Internet Watch Foundation and StopNCII on online safety, child sexual abuse material detection and reporting, and non-consensual intimate-image abuse.

Accuracy remains the practical test

Character.AI has not published accuracy figures or outcome data for its self-harm detection or age-estimation systems in this update. That leaves their effectiveness and error rates unresolved. When the under-18 restriction was announced, critics quoted by The Associated Press also sought details on age verification, privacy preservation and the potential psychological effects of suddenly removing access for young users.

Sources

  1. blog.character.aiContinuing To Build Upon Our Safety Priorities
  2. apnews.comCharacter.AI is banning minors from interacting with its chatbots