← Back to Blog
GPT Image 2.0AI Character AppsImage GenerationAI CompanionConversion

GPT Image 2.0 Case Study: How Better Character Portraits Lifted Conversion on an AI Companion App

Tendera TeamUpdated September 14, 20264 min read

What We Changed

On April 23, 2026, we replaced the four static character portraits used across Tendera (Sophia, Mia, Elena, and Jade) with new ones generated using GPT Image 2.0. The workflow was deliberately low-tech. The prompts that drive each character's image were run through ChatGPT. The resulting PNGs were downloaded. The existing portrait files were swapped out by hand.

At the time of this portrait refresh, Tendera did not have an in-chat image generator. This specific change was not a pipeline change; it was four file uploads. Tendera later added requested character photos on Web and iPhone.

Everything else stayed the same. The character system prompts that drive each persona's voice and behavior were untouched. The UI didn't change. The chat backend didn't change.

What We Measured

Three days after the swap, two numbers had moved:

  • • Visitor-to-signup conversion: up approximately 5%
  • • Visitor-to-chat conversion (counting both guest preview chats and post-signup chats): up approximately 8%
  • These are two different metrics measuring two different events. We're not stacking them up against each other. They're parallel data points pointing the same direction.

    The reason both numbers are reported in this case study is that the second one moving was the part we didn't expect.

    What We Assumed Going In

    Before the swap, the working assumption was that better character art would mostly help acquisition. Prettier card on the landing page, more clicks, more signups. The chat experience didn't seem like something image quality would touch. By the time a user is sitting in front of a chat input, the visual selling job feels mostly done.

    That assumption was wrong, or at least incomplete.

    What Actually Changed in the Images

    The character topology is identical. Same four characters, same wardrobes. The poses are roughly the same too. What's different is how legible each character is now.

    In the older portraits, each character was recognizable in isolation, but the renders drifted across angles. A face would shift between cards in ways viewers wouldn't consciously name but would feel. A slightly different woman in a similar outfit. Hands and small details broke in ways AI image models were known to break circa 2024-25.

    GPT Image 2.0's renders look more boring in some ways. Less stylized. The model feels less like it's interpreting the prompt and more like it's executing it. But the character holds across angles. Same person across multiple shots. No drift.

    The other thing the new model nails is dimensionality. The old renders were clean but flat. They read as illustrations. The new ones have physical depth. Light hitting the side of a face. A jacket folding the way fabric actually folds. None of it is photoreal. The dimensionality just reads.

    Why We Think the Chat Number Moved

    Here's the framing that fits the data without overclaiming.

    When a visitor lands on the marketing page, the question they're answering is whether the surface signal looks decent enough to click in. Image quality affects this, but a serviceable card will still get the click.

    Once they're past the door, sitting in front of an actual character profile or already in chat with the character header in view, the question gets sharper. Now they're evaluating whether this person feels real enough to talk to. The image is the only non-text signal in the room. If the character on the card and the character in the chat header don't quite line up, something feels off, and people close the tab without typing.

    Most users wouldn't describe this consciously. We're inferring what their gut is doing. But chat-side conversion moving with prompts and copy unchanged points at the visual layer doing some work past the landing page.

    What We Want to Test Next

    The obvious next experiment is whether the same model can produce reliable expression variants for the chat header. Right now each character has one default portrait. If the same character could subtly shift expression based on conversation tone (a softer face during something quieter, a smirk during banter), chat-side recognition could go up another step.

    That was a harder problem: consistency within a session had to sit on top of consistency between angles. Tendera later added explicit, on-request character-photo generation to the chat experience on Web and iPhone; dynamic expression changes in the chat header remain unshipped.

    If we tested it on one character first, it would be Jade, the character users tend to spend the most time with. The voice on her side is already doing most of the work. The image is the one input that hasn't caught up.

    Caveats

    A few things worth flagging for anyone reading this case study:

  • • This is three to four days of data on a small app. Effects could compress as the sample grows.
  • • The change was on the portraits, not the character system prompts. If a similar AI character app has its bottleneck on the writing side (voice, dialogue cadence), regenerating images won't help.
  • • There was no clean A/B with old vs new portraits served to different cohorts. The whole site flipped over April 23. A pre-existing upward trend coinciding with the swap could absorb part of the lift.
  • • Visitor-to-signup and visitor-to-chat are different metrics measuring different events. We report both because both moved, not because one is bigger than the other.
  • • This case study covers a manual asset swap. The new portraits were generated in ChatGPT and uploaded by hand; Tendera's later requested-photo feature was a separate product project.
  • What This Might Mean for Other Character Apps

    If you're building a product where a user is supposed to form a relationship with a fictional persona (a character app, an NPC system, an AI tutor with an avatar, a virtual host), your image generator might be doing more work than acquisition-side metrics suggest.

    Worth regenerating your portraits and watching what moves.

    If you'd like to see how the new portraits hold up in actual conversation, you can try Tendera with five preview messages, no signup required.

    Ready to meet your AI companion?

    Four unique personalities. Each one remembers you. Free to start.

    Meet Your Match

    Frequently asked questions

    What was changed exactly in the swap?

    Four static character portraits on Tendera (Sophia, Mia, Elena, Jade) were regenerated using GPT Image 2.0 in ChatGPT, downloaded as PNG files, and manually uploaded to replace the older portrait files. No image-generation pipeline was added to the product. Character system prompts, UI, and chat backend were all unchanged.

    Did both signup and chat conversion really move?

    Yes. Over the three days following the April 23 swap, visitor-to-signup conversion lifted approximately 5% and visitor-to-chat conversion (which counts both guest preview chats and post-signup chats) lifted approximately 8%. These are different metrics measuring different events. We report them as parallel data points, not as a magnitude comparison.

    Why would image quality affect chat engagement, not just landing page clicks?

    Our working hypothesis: at the landing page a visitor is evaluating whether the surface signal looks decent enough to click in. The bar is fairly low. Once they're past the door and looking at a character profile or chat header, the question gets sharper. They're now evaluating whether this person feels real enough to talk to. The image is the only non-text signal in the room. Subtle inconsistencies in the older portraits (faces drifting between angles, flat illustrative quality) likely register as 'something is off' even when users can't articulate it.

    Was this a clean A/B test?

    No. The whole site flipped over on April 23 with no holdback group. A pre-existing upward trend coinciding with the swap could absorb part of the lift. The directional finding (both metrics moved up) is more trustworthy than the absolute deltas.

    What does GPT Image 2.0 do better than older image models for character consistency?

    Two things stood out. The same character holds across multiple angles and shots without facial drift, which the older generator we'd been using struggled with. And the renders have physical depth and dimensionality, light hitting the side of a face, fabric folding the way fabric actually folds, rather than reading as flat illustration. Less stylized overall, but more legible as a consistent person.