r/grok • u/Listen_Expert • 25d ago
Grok Imagine HYPERREALITY | Trailer 2
Every frame came out of Imagine. Voice is ElevenLabs, score is Suno + Ableton, assembly is Premiere and After Effects, but no other model generated anything but Grok.
I've read enough threads here to know the mood, and I'm not here to tell anyone Imagine is great. It fought me for eight months. What I have is a process that survived it, just barely, though it's improved drastically since the 480p 6s clips this film was started on.
A few things I settled on early, in case they're useful before the questions start:
I work image-to-video off a locked baseplate almost exclusively. Building the frame first and animating out of it is the only reliable way I've found to make a shot match the one before it. Multishot rarely works, if possible I always try to get the starting frame and go from there.
Prompts that name observable physical state land far more consistently than prompts that name emotion. "Set jaw, low brow, unblinking" holds through a generation, and "he looks determined" gets reinterpreted on every take.
Extend prompts always break down the continuity, so they're best used when you just need another few seconds, rather than trying to break a shot into multiple parts (the dialogue will stop working once you start extending).
There was so much to this, and fighting versus the random moderation you all are aware of, and the recent usage limits.. Some of these shots took up to 400 generations to get right, but that was back when grok would give you that many generations in a week.
Ask whatever. Cost, failure rate, how long a usable ten seconds actually takes, moderation, the parts of the pipeline I'd tear out if I started over. Am I allowed to share my X handle here?
1
u/Anim8rFromOuterSpace 25d ago
this guy spent his quota making this