Character.AI Launches Image Models for Consistent Fan Stories

The new CAI-Image family is built around a stubborn creative problem: keeping a character recognizable while style, camera, setting and cast change from one image to the next.

By 2 min read
Character.AI Launches Image Models for Consistent Fan Stories
Character.AI Launches Image Models for Consistent Fan Stories

Listen to this story

The audio brief

About 1:34
0:001:34
Read transcript
Character.AI is adding a new family of image models designed to keep the same character recognizable as a fan story moves through different poses, styles, locations, and comic panels. Called CAI-Image, the models now power the app’s Comics, Imagine, and Imagine Message experiences. They are post-trained versions of the open-source Qwen-Image model, but the important distinction is the workflow: users can combine a character reference with pose, style, and location references in one generation. That makes the target less about producing one impressive picture and more about maintaining identity across a sequence. Character.AI says CAI-Image can interpret 245 natural-language camera positions. Those are organized across five distances, plus seven horizontal and seven vertical angles—enough to describe everything from a distant side view to a close-up from above or below. There is also a dedicated manga model, tuned for comic pages with panel structure, dialogue text, and a cast that remains recognizable from panel to panel. The models run on CAI-MM-Studio, Character.AI’s system for data preparation, multi-node training, evaluation, and serving. The company claims that this setup improves speed and cost, though there are no independent comparisons in the release. The boundary is clear: CAI-Image V2 is slated to address moving-camera environments, subtler emotion, and more reliable manga layouts and text. For now, character continuity is the foundation; full scene continuity remains the harder test.

Story brief

3 key points

Character.AI is adding CAI-Image, a family of post-trained Qwen-Image models, to its app’s Comics, Imagine, and Imagine Message experiences. The release targets repeatable characters across styles, poses, camera setups, and manga panels—a product strategy built around fan-driven, character-led creation rather than one-off image quality. The system supports 245 natural-language camera positions and a dedicated manga...

  1. 01

    CAI-Image combines character, pose, style, and location references in one generation.

  2. 02

    Its camera-control system covers 245 positions across five distances and seven horizontal and seven vertical angles.

  3. 03

    CAI-MM-Studio handles data, multi-node training, evaluation, and serving for the image models.

Character.AI has put a new family of image models into its app, aiming to make fan stories visually coherent rather than merely attractive. CAI-Image powers (c.ai) Comics, Imagine and Imagine Message, with the company emphasizing recognizable Characters across changing scenes, styles and comic pages.

Character.AI describes CAI-Image as a set of post-trained versions of the open-source Qwen-Image model. The company says it began with the demands of its own character-led creation flow: users can generate images from chats and premises, where the system assembles prompts and brings the relevant Characters into a scene.

That framing explains the release’s focus. Character.AI says the models were tuned for style transfer, multi-Character scenes and pose control, camera control, and manga and comics generation. It also says a generation can combine Character images with pose, style and location references, rather than asking users to resolve those inputs separately.

A camera-control claim
245Natural-language camera positions

Character.AI says CAI-Image understands 245 camera positions, organized across five distances and seven horizontal and seven vertical angles.

A single comic page concentrates the problem: panel structure, dialogue, emotion, camera variation and visual continuity all need to work together. Character.AI has included a dedicated manga model that it says is trained for those pages, including accurate dialogue text and a cast that remains recognizable from panel to panel.

Example comic page generated with Character.AI’s CAI-Image manga model.
Character.AI presents its manga model as a tool for comic pages that require panel structure, dialogue text and Character continuity. Source: blog.character.ai.

The company says the models run on CAI-MM-Studio, its multimodal system for data collection and cleaning, multi-node training, automated evaluation and serving. Character.AI argues that controlling those stages lets it tailor models to its app’s workflow and operate them quickly and cost-effectively; those speed and cost advantages are company claims, not independently reported comparisons.

Character.AI’s stated next step is CAI-Image V2, which it says will work on consistent environments as the camera moves, more nuanced emotion between Characters, and better manga layouts and text. That is a useful boundary on today’s release: the company is presenting Character continuity as the foundation, while scene continuity and more reliable comic construction remain work in progress.

Sources

  1. blog.character.aiPost-training image models for fandom

Loading discussion...