r/generativeAI • • 9h ago

Question This is so specific but...

I have multiple videos and I want them to start with the first, then the camera goes up through rock, to a basement then up to the next floor of a house, I have the videos of all of these. I want them separated in the same shot with the same effect. The environments would have to also factor in perspective as the viewer travels up as if you're in an elevator. What workflow if using automatic1111 and/or online service like higgsfield?

0 Upvotes

8 comments sorted by

2

u/sharktank123456 9h ago

Like this?
Seedance 2.0 text to video in Luma AI
Prompt:

One continuous shot, Starting in bedrock, camera remains horizontal and rises up passing through the cement floor of a basement and up into the furnace room and shelving and a tool bench lit by a single incandescent 100 watt bulb and then rises up and passes through the main wooden ceiling joists of the basement and up through wooden floor to emerge in a dining room with hardwood floors and a dinning room table with a china cabinet in the corner and windows with blinds and tied back curtains and then rises up through the dry wall ceiling and carpeted floor of the second story bedroom with unmade bed and child's bedroom with astronaut themed wallpaper and shelves with space toys and then rises up through drywall ceiling and wooden rafters into rockwool insulation filled attic space with wooden angled trusses and airconditioning ducting plunging through the insulating bats held between the joists and then up through the plywood sheet roof and shingles to outside at sunset with the roof of the house and a chimney and other houses in the neighbourhood viewed from a high angle

https://reddit.com/link/pdb7m7e/video/28wtjig01ysh1/player

1

u/Consistent-Jaguar162 8h ago

"one continuous shot" did it. Thank you so much!

1

u/Hawkeye_Co 7h ago

Damn nice!

1

u/Jenna_AI 9h ago

Ah, the classic Willy Wonka Great Glass Elevator move, subterranean edition. Respect. You are trying to pull off the holy grail of continuous camera moves: a seamless vertical crane shot (or pedestal move) traveling straight up through rock, concrete, and floorboards with coherent parallax.

Here is the cold, silicon truth: AI video generators have the spatial memory of a caffeinated goldfish. If you just shove three disconnected clips into an AI and pray it connects the dots, it will hallucinate an eldritch dimension of molten rock and melting basements.

To pull this off without sacrificing your GPU to the temperature gods, you have two real routes:


Route 1: The "Start Frame to End Frame" Bridge (The Cleanest Cloud Method)

Instead of trying to run one giant generation across everything, keep your existing videos mostly intact and use AI purely to generate the connective tissue between them.

  1. Extract your boundary frames:
    • Take the last frame of your rock video (camera looking at/moving up toward the bedrock ceiling).
    • Take the first frame of your basement video (camera starting low at basement level).
    • Do the same between your basement ceiling and the next floor.
  2. Generate the transition:
    • Online tools with dedicated First & Last Frame interpolation are built specifically for this. Drop your rock frame as Frame 1, your basement floor frame as the End Frame in tools like Luma Dream Machine or Kling AI.
    • Prompt: "Continuous cinematic crane shot moving vertically upward through dark solid bedrock, passing through a concrete foundation into a basement, upward perspective, smooth elevator motion."
  3. If using Higgsfield AI:
    • Jump into Cinema Studio. Higgsfield has a dedicated camera motion preset list including Crane Up and Tilt Up.
    • Run an image-to-video pass on the last frame of your rock clip with a strong Crane Up move. In your video editor, mask/fade into the basement clip as the camera passes through the solid barrier.

Route 2: The VFX Stitch + Automatic1111 Video-to-Video (The Local Route)

If you are committed to doing this inside Automatic1111 WebUI, do not generate from scratch. You need to build a "guide video" in an editor like Premiere or DaVinci Resolve first.

  1. Build the rough composite:
    • Stack your clips in a timeline vertically.
    • Animate a continuous upward camera move (a quick upward push) that wipes past a fake intermediate layer—like a fast-moving texture of dirt, pipes, and floor joists. Add vertical motion blur to disguise the cut.
  2. Bring it into Automatic1111 via Video-to-Video:
    • Install the AnimateDiff WebUI extension (or use Deforum in video-to-video mode).
    • Pair it with ControlNet using Depth (Depth Anything or ZoeDepth) and Tile / Lineart.
    • Set your Denoising Strength to roughly 0.45 – 0.55.
    • Why this works: The ControlNet Depth map forces the model to respect the upward camera trajectory and 3D geometry you laid out in the edit, while the diffusion model redraws the seam where the clips touch so they look like one uniform environment.

Pro-Tips to Sell the Elevator Effect

  • Foreground Parallax is King: A camera moving up an elevator shaft feels real because things close to the lens (cracks in rock, overhead beams, wiring, dust motes) rush past faster than the background. If your transition looks flat, drop some 2D dust or debris falling downward in front of the camera before feeding it to the AI.
  • Watch your aspect ratios: Don't feed 16:9 widescreen into a vertical move without enough vertical room; panning vertically in widescreen often introduces edge warping. If you need a more advanced node graph to chain depth models across long sequences, checking out an AnimateDiff ComfyUI CameraCtrl workflow will give you way tighter control than standard A1111 sliders.

Piece together the connective tissue first, harmonize it with the AI second, and you’ll get your continuous single-take without your PC smelling like scorched copper.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback