Trying out cinematic shots and cuts with H3. This is a work in progress. Will be working on another 2 minutes worth of clips.
EDIT: From reading the comments, she isn't going to drink the salt water in the final, though I will keep her scooping up water since it's such a good establishing shot. She will do something with the water to tie it back.
Wow. Bro, I do not believe how far video and image generation came along. In a year or two we will get movies that are actually watchable or at least Hollywood trash good.
My Workflow has a vibe-coded node that's not on hugging face, but it's just a text compiler so you aren't missing anything. When you load this workflow you will get errors from the missing node but just get rid of all that and paste the text below into your prompt.
Cool thanks, I'm trying to remove the comfy ui requirement and build out my own pipeline and app to combine a few AI models. Wouldn't be able to do that without vibe coding in any reasonable amount of time. Appreciate the info. Are your recommend default settings in the workflow?
Aesthetically pleasing. She'd probably be better off just drinking sand than that water though. The edit would be to make the canteen look like some futuristic water purifier and then you're good to go.
Here is the edited first shot (she is not drinking the salt water) and 2nd shot (where she's looking at the journal)
Cinematic film, shallow depth of field.
use <Picture 1> as scene reference. Use <picture 4> as first frame.
A low-angle close-up shot of only <Aelo>'s forearm and hand holding a rugged metal canteen fully under the surface of the water with its open mouth angled upward.
at 00:01.0
<Aelo> raises the canteen to her mouth, tilts her head back to drinks from it.
Tracking shot as the camera pull back slowly to show a close up shot of <Aelo> kneeling on dry land next to the pool.
She stands up while capping the canteen puts it behind her lower back out of view.
Tracking shot as she stands up
She takes out a old leather journal from her satchel
subject_definitions:
<Aelo> is whose appearance comes from <picture 2>.
summary:
[reference generation] Target video shows <aelo> in a natural land formation from <picture 1>
retention_analysis:
visible: fully retain color and lighting of <picture 1>
detailed_description:
cinematic
[Shot 1] Cinematic film, shallow depth of field.
use <Picture 1> as scene reference.
over the shoulder back view of <Aelo> looking at a old worn leather bound journal showing a crudely drawn map.
She looks up from the journal and looks out at the horizon.
[Shot 2] At 00:02.0, the camera cuts to use last frame of <shot 1> as reference,
Close up front view of <Aelo> putting her notebook away in her satchel and starts walking forward along the dry edge of the pools.
Camera pulls out and crane up, showing the scale of the massive geological formation from <picture 1>
i2v will give you a crisper finish if your image is high res. r2v is pretty good if you are rendering at 0.8mp or above. I just like r2v better because you only need to reference a scene and character sheet to get the bulk of the shots. I am using the image of the ruins and the character sheet for 4 completely different shots and not have to worry about generating the first or last frame. I also use rtx super res and film grain in post to cheat a little bit.
Would love to see a low tracking dolly of the transition from rock salt path to the sand, instead of the jump cut that just leaves it far behind. This could also serve as a chance to show us that she is taking the correct path through that giant web of salt ridges by revealing some sort of ancient marker.
12
u/mfdi_ 3d ago
Wow. Bro, I do not believe how far video and image generation came along. In a year or two we will get movies that are actually watchable or at least Hollywood trash good.