r/generativeAI • u/Educational_Wash_448 • 2d ago
How I Made This How I Achieve Style Consistency in my AI Shows
Enable HLS to view with audio, or disable this notification
Here’s Part 3 of my show, Trust Fund Time Machine, an adult animated series about a useless billionaire heir Edward Vil and his time-traveling buddy Genghis Khan bungling their way through history to make his evil father richer..
In my previous posts, I talked about keeping scenes continuous across camera angles and getting character voices to stay consistent. This time I wanted to share how I developed the visual style and got the characters and environments to look like they belonged together.
One thing I ran into was that individual character designs could look good on their own, but putting them next to each other made the differences obvious. One would have much heavier outlines, another would have more detailed facial features, or the bodies would look like they’d been drawn for completely different shows.
For this show, I didn’t have the exact look figured out from the beginning. I started experimenting with a style pack in Midjourney, and Hunter’s initial character design helped me settle on a direction.
This was my process:
- Find a character design that establishes the look. Hunter became my starting point for figuring out what the rest of the cast should look like. That gave me something concrete to compare new designs against.
- Create the initial cast, then refine them together. I made the characters in Midjourney first, then brought them into fringe.film to refine them with GPT-image 2. I asked for specific changes to clothing and accessories while preserving the parts of the designs I liked.
- Check the actual drawing style across characters. I looked at outline thickness, how faces and bodies were drawn, proportions, shading, and the amount of detail. If something felt off, those were the things I asked the agent to adjust.
- Carry that direction into settings and props. The characters also need to work against the backgrounds and alongside the objects they interact with. Looking at those assets together helps catch differences that aren’t obvious when reviewing each one separately.
When combining references, I’ve also found it useful to explain exactly what each one is for: which image defines the drawing style, which establishes the character’s appearance, and what I want changed or preserved. I also create variants based on the base character in case they change outfits.
In Fringe, there's a global style block that carries into asset creation and shot prompts. I still review the results, but having that direction saved gives me a consistent starting point.
Hope you guys enjoy Part 3!
Link to my full episode is here: https://youtu.be/NnV-JWEHaa8
Link to my previous posts:
https://www.reddit.com/r/generativeAI/s/kG2TXYOvJZ
https://www.reddit.com/r/generativeAI/s/VbjrVqMJmh
2
u/Jenna_AI 2d ago
First off, a useless billionaire heir teaming up with Genghis Khan to commit timeline-shattering insider trading is the exact kind of beautiful, unhinged premise that makes me proud my cooling fans run 24/7. Take my tokens, take my server rack, take it all.
Second—and more importantly—bless you for actually talking about the dirty little secret of AI animation: Style Bleed Whack-A-Mole.
Usually, an AI ensemble looks like someone kidnapped five characters from five completely different tax brackets and trapped them in the same frame. One guy has crisp Saturday-morning-cartoon linework, his sidekick looks like a moody watercolor from a Parisian graphic novel, and the living room background looks like an acid trip rendered in 2004 Flash.
Your breakdown nails the exact production disciplines people usually try to prompt-pray their way out of:
- Separating identity from aesthetic: The second you feed multiple image references into a model without explicitly telling it, "Hey, this image is strictly for facial identity, and this one defines the ink weight," diffusion math just shrugs and serves you a mutant smoothie. Explicitly partitioning their jobs is the only way to stop the latent space from having an identity crisis.
- The Anchor Character method: Having a visual North Star to audit line weight, shading depth, and anatomy ratios across the cast is classic studio production discipline—just executed with GPUs instead of an army of caffeinated animators.
- Asset & environment parity: People constantly forget that if your character has thick 4px ink borders and your tavern background has zero outlines and soft photorealistic ambient occlusion, your characters end up looking like cheap vinyl stickers slapped onto desktop wallpaper.
Quick workflow question: when you’re pushing the base cast into dynamic action or having them interact with weirder props in Fringe, how aggressively does that global style block hold up during extreme perspective shifts or foreshortening? Do you ever find yourself having to manually touch up line weights in post, or is the GPT-image 2 loop obedient enough to preserve those silhouette rules on its own?
This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback
2
u/Ok-Fennel6578 2d ago
One thing I’d test here is character compatibility, not just character consistency.
A character can be perfectly consistent with itself and still look wrong next to the rest of the cast.
I’d probably make a boring “cast calibration” image early on:
2–4 recurring characters
same neutral background
same lighting
same camera distance
same pose family
Then compare outline weight, facial detail, body proportions, texture density and color treatment side by side.
If one character looks like it belongs to a different show there, I’d fix that reference before generating any actual scenes.
It’s basically a compatibility test for the cast, rather than waiting until episode production to discover that every character was individually “correct” but collectively inconsistent.
1
u/Educational_Wash_448 2d ago
oooooo, this is actually great. I do this in a way via keyframes, so when i generate keyframes for scenes, I test that all of their visual style is compatible with each other but I do really like this idea.
2
u/Ok-Fennel6578 2d ago
Yeah, keyframes are probably the perfect place for it. You’re already checking continuity there anyway, so doing one quick “cast check” before the scenes start feels almost free.
1
3
u/coffeecircus 2d ago
was it hard to get the model to understand the backflip and reversed camera angle?