
DISCOVER
Start with our strongest curated characters.
MEL · CONVERSATIONAL AI
MEL is a live AI companion experience where conversation, expression, and emotion evolve in real time. I designed the end-to-end interaction model—from character creation and voice selection to emotional feedback, gifting, and dynamic video—so users could shape a companion that feels responsive, expressive, and uniquely their own.

03PRODUCT INSIGHT
WHAT MADE THE CHARACTER FEEL ALIVE
Conversation, voice, expression, and visual context had to respond together—consistently and in real time—for the character to feel truly alive.
04CHALLENGE 01 · VOICE CONVERSATION UX
TURN-TAKING CLARITY
A live exchange can become ambiguous quickly. The interface had to make speaking, listening, waiting, and response timing legible without interrupting the conversation.
Voice, text, and visual response competed for attention.
One dominant state at a time; transition before decoration.
Speaking, listening, waiting, and responding became distinct states.
05CHALLENGE 02 · CREATE YOUR OWN MEL
THREE WAYS TO BEGIN
Some users wanted a polished character immediately. Others had a specific face in mind, while some wanted control over traits such as age, personality, and skin tone. I designed three creation routes—Discover, Photo, and Custom—so users could begin with the level of intent and effort that felt natural to them.
A single creation flow assumed every user already knew exactly what they wanted.
Match the starting route to the user’s level of intent.
Three routes from inspiration, visual reference, or detailed control to a personalized MEL.


Start with our strongest curated characters.

Create a character from a visual reference.

Define the character through individual traits.
THE MOST NATURAL STARTING POINT
Discover became the most frequently entered creation route. Curated presets required the least upfront effort and gave users confidence in the final quality before asking them to define individual traits.


WHY AN OVAL WHEEL?
MEL’s visual identity was built around circular forms. I extended that language into a vertical oval to better frame portrait-oriented character imagery within a mobile screen. The centered character remained visually dominant, while partial characters at both edges signaled that more options could be explored.
Compared with a grid, the wheel made fast comparison less efficient, but gave each character more visual presence and made browsing feel closer to meeting characters one at a time.
The wheel changed browsing from comparison to discovery.
06CHALLENGE 03 · GENERATIVE VISUAL PRODUCTION
FROM POSSIBILITY TO PRODUCTION
At the time, generative models could produce striking single images, but character identity, anatomy, composition, and scene continuity were still unreliable.
I worked closely with the ML team to turn visual judgment into a repeatable ComfyUI-based production workflow. I defined the emotional brief, scene hierarchy, composition, continuity, and product constraints, then reviewed and refined the outputs until they could function as real date and location scenes—not just technical demonstrations.
The technology showed what was possible.Our workflow made it usable.
Mood, relationship context, location, camera direction, and product constraints.
ComfyUI nodes, prompts, references, seeds, and controllable inputs.
Character identity, anatomy, composition, artifacts, scenario clarity, and UI legibility.
Date scenes, locations, crops, safe areas, and platform-specific formats.

07IMPACT & REFLECTION
REPORTED BETA SIGNALS
Reported average chat duration
Reported onboarding exits
Beta participants
Metrics are retained from the original beta materials and presented as reported directional signals rather than causal proof.
A believable AI companion is not created by model capability alone. It emerges when conversation, voice, character creation, and generated scenes work together as one understandable experience—giving users a reason to stay and continue the relationship.
After reducing hesitation in an AI creation flow, I explored a different question: can immediate sound, motion, and haptics make interaction itself worth repeating?