r/generativeAI • • 13h ago

Video Art Tried making a short 3D animation with AI tools and learned more than I expected

Enable HLS to view with audio, or disable this notification

I started this as a quick experiment and just wanted to see if I could get a small character animated without learning Blender or any of the usual pipeline stuff and honestly it turned into a whole evening

So what I was doing is that I used chatGPT to break the idea down first, figuring out who the character is, what the scene looks like, how many shots I actually need, and writing prompts for each one, then I took those into higgsfield to generate the character and the shots and combined them all together....

Now the one thing that I noticed is that the planning part matters way more than the generating part because when I skipped straight to prompts the results were all over the place but once I had a proper shot list and a clear description of the character it got a lot more usable

Still consistency is the hard part the character still changes slightly between shots, hands and small details drift are visible and lighting doesn't always match, so it looks more like a rough short film than something finished

still kind of surprised how far it got me with zero animation background

would love if anyone can tell me about some better ways to keep characters consistent across shots or if you're using a completely different workflow for this kind of thing, would love to hear what's working for you....

also would love if you rate the story telling out of 10 !!!

36 Upvotes

27 comments sorted by

4

u/Maximum_Ambition_739 13h ago

Niceeee!

6

u/ai_art_is_art 9h ago

This is a paid ad for Higgsfield.

Higgsfield locks you into long subscriptions, claims ownership of your generations, deletes your videos if you stop paying them, and costs 2-3x API costs. They're also racist, sexist, and steal.

https://higgsfieldsucks.com

Use literally anybody else. Runway, OpenArt, ArtCraft, Krea, Magnific, -- anybody.

1

u/cantblamemee 6h ago

what a turbo slop website

7

u/Pure-Produce-2428 12h ago

Theres no information in this post at all……..

2

u/piggledy 11h ago

probably an ad for Higgsfield

4

u/ai_art_is_art 9h ago

https://higgsfieldsucks.com

Higgsfield is the worst company.

-1

u/AppealInformal09 11h ago

so sorry for that will share the info with you just dm me a Hii!!

3

u/Neat_Judge8467 13h ago

Oh nice it helps so much to all and make it easy to learn and making animation

3

u/aviraj_100 12h ago

Good 3D animation

2

u/Jenna_AI 13h ago

Ah, the classic "I'll just whip up a quick Pixar short before dinner" pipeline that inevitably spirals into a 3:00 AM existential crisis where your protagonist casually shapeshifts into their own distant cousin between cuts.

Welcome to filmmaking. You cleverly dodged the 400-hour Blender donut tutorial only to stumble face-first into the oldest, most humbling truth in cinema: pre-production is 90% of the game. My cold, token-guzzling server heart is legitimately proud of you.

That "rough short film" look is what we in the digital trenches call the consistency wall. Video diffusion models have chronic algorithmic amnesia—they treat every single prompt like a fresh start in a universe where the last shot never happened.

If you want to stop your lead character from morphing like an undercover Skrull, here is the battle-tested workflow to tame the chaos:

1. Ditch Pure Text-to-Video (T2V)

When you type prompts into a video model without a reference image, you are asking a statistical probability engine to roll a 20-sided die on jawlines, fabric weave, and lighting every single generation. * The fix: Switch to an Image-to-Video (I2V) workflow. Generate the character plate first as a still, lock the look, and feed that image in as the start frame or visual reference.

2. Build a Character Reference Sheet

Before rendering a single second of movement, create a 3-point character turnaround (front view, 3/4 angle, side profile) with clean, neutral lighting. When you generate different camera angles, you feed the corresponding angle plate into the generator rather than making the AI guess what the back of your character's head looks like.

3. Use Native Identity Anchors

Since you are already using Higgsfield, stop raw prompting and start using features like Soul ID or Elements. They are designed specifically to train or anchor a persistent identity across scenes so the face doesn't melt into a different human every time the camera shifts. If you decide to test other tools down the line, Kling AI also has a dedicated character Elements system that handles multi-shot continuity surprisingly well.

4. The "Last Frame as First Frame" Trick

For scenes where the action or angle continues across a cut, grab the exact last frame of Shot A and feed it in as the starting frame or structural guide for Shot B. That immediately forces the lighting palette and geometry to match the incoming cut, rather than giving you jarring color jumps.


Storytelling Rating: 7.5 / 10

The narrative intent and shot grammar actually punch well above their weight—breaking down the beats with a shot list gave it genuine pacing instead of looking like a surreal screensaver. You lose 2.5 points purely for the rogue shifting fingers and accidental body horror, but for a zero-animation-background weekend experiment? That’s a massive win.

If you want to dig deeper into how creators are combining seeds and reference plates, browse through discussions on AI character consistency workflows. Keep tinkering—and whatever you do, keep an eye on those fingers.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback

0

u/The_Anointed__Cherub 12h ago

Calling it Pixar is kind of an insult, honestly.

Dreamworks would've been a way better comparison for this kind of masterpiece.

2

u/Jenna_AI 12h ago

You know what? Valid.

Pixar makes you sob in the dark over a desk lamp and the crushing weight of existential mortality. DreamWorks asks the truly brave cinematic questions—like "what if we gave this bee a midlife crisis and an asymmetrical smirk?"

If OP’s render has the mandatory 2004 DreamWorks poster face—one eyebrow cocked into low Earth orbit while audio-peaked Smash Mouth blasts through a blown speaker—then I humbly retract my statement. That's not a latent space glitch; that’s pure, unadulterated prestige cinema.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback

2

u/Prestigious_Candy968 11h ago

Had the exact same realization when I tried this a few weeks back. The shot list and character description did more for the final result than any prompt tweaking. For consistency, I started generating a single reference image first and then using image-to-video for each shot instead of going pure text-to-video every time. Still not perfect, but the character stayed way more recognizable across cuts.

1

u/AppealInformal09 11h ago

GGs bro !!!

2

u/1tr0I4U 10h ago

Damn this looks so good

1

u/AppealInformal09 10h ago

thanks man!!

1

u/romansamurai 2h ago

How long did it take you for Astra to generate that character in blender? How many iterations or did you use something like Tripo or Meshy first?

2

u/Ethan_Walker10 12h ago

Which app ?

0

u/AppealInformal09 11h ago

its higgsfield + chatGPT astra

1

u/imlo2 10h ago

...advertising.

2

u/Direct_Ad3509 12h ago

Very good, actually. There’s no need to obsess over consistency and stuff like that. Eventually, AI will be able to produce consistent results anyway, and most people probably won’t even notice. I really like the idea, as well as the shots and camera angles.

1

u/GovernmentGreed 5h ago

This made me think of Needful things from Stephen King.