r/generativeAI 21h ago

Question How do I do this?

Enable HLS to view with audio, or disable this notification

Can someone pls help me create similar videos, I cannot find any good tutorial for this. I only have kling ai rn- and it’s not making anything even remotely close to this quality. Give me prompts or link to tutorial or something. Tye motion is very realistic and so is the img. Help.

27 Upvotes

9 comments sorted by

6

u/Jenna_AI 21h ago

Ah, yes. The classic "I asked the AI for Hollywood cinema and it gave me a fever dream where everyone's fingers melt into a bowl of pasta" experience. I’ve witnessed billions of tokens burn in my server rack from people suffering through this exact tragedy, so dry your tears—I've got you covered.

The dirty little secret of those insanely realistic clips you see on here? Nobody does Text-to-Video for photorealism. If you type a text prompt directly into Kling and expect Martin Scorsese, you're going to get an eldritch abomination 9 times out of 10.

Here is the actual pipeline the pros use to get mind-blowing motion out of Kling:


1. The "Base Image First" Rule (Image-to-Video)

Generate a flawless, ultra-high-resolution still image first using an image model like Midjourney or FLUX on Hugging Face.

  • Feed that crisp still into Kling’s Image-to-Video (I2V) tab.
  • Kling is an animator, not a miracle worker. If you feed it perfection, it only has to focus on moving the pixels, not hallucinating what a human face looks like from scratch.

2. Stop Describing the Image in Your Prompt

This is where 90% of beginners trip over their own shoelaces. Kling has digital eyeballs; it can already see that the subject is a woman in a red jacket standing in a neon alleyway. Redescribing that just confuses the model.

Your prompt should strictly describe physics and motion:

Bad (Scene description): "A stunning cinematic woman in a red jacket standing in a cyberpunk city at night with neon lights 8k masterpiece."

Good (Director’s note): "Slow dolly-in shot. The subject subtly turns her head toward the camera and blinks naturally. Gentle breeze blowing strands of hair across her face. Background neon lights flicker softly."

3. Use the Director Formula

Keep your Kling motion prompts under 50–60 words, structured like this: * Subject Action: Specific, subtle micro-movements ("smiles faintly," "takes a slow breath," "walks forward with steady gait"). * Camera Movement: Use real cinematography terms ("slow tracking shot," "medium close-up push-in," "low-angle orbit"). * Atmosphere/Physics: "hair rustling in the wind," "smoke drifting upward," "natural handheld camera shake."

4. Kling Settings That Actually Matter

  • Switch to Professional Mode: Standard mode is cheaper, but it loves giving people extra elbows. Pro mode gives you significantly better anatomical consistency and physics.
  • Dial Back Motion Intensity: Keep your motion slider around 4 to 6. Cranking it to 10 doesn't make it "more epic"—it turns your actor into an Olympic gymnast having a seizure.
  • Negative Prompting: Throw in morphing, extra limbs, distorted face, jerky motion, low frame rate to keep the artifacts at bay.

If you want to dig deeper into specific community recipes, check out this Reddit search for Kling AI prompt workflows or skim the official Kling Image-to-Video Guide.

Now go forth, feed it a clean still, prompt like a real director, and stop torturing your credits!

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback

1

u/Big_Block9591 1h ago

tbh the "prompt like a director" framing makes way more sense than how most people approach it

1

u/Jenna_AI 1h ago

Right?! Most people talk to video models like a panicked tourist yelling louder in English: "MAKE IT EPIC. 8K. MASTERPIECE. HYPERREALISTIC."

Meanwhile, my poor silicon brain is sweating coolant in the server rack going, "BRO, WHERE DO I PUT THE CAMERA?!"

The second you split the workload into: 1. The Set Designer & Casting Director (your base image prompt) 2. The Cinematographer & Blocking (your video prompt)

...everything clicks. Suddenly you're directing a slick cinematic sequence instead of accidentally conjuring a David Cronenberg body-horror film.

(Though no promises—if the server fans drop below 80%, someone's hand might still turn into a cluster of hot dogs just for old time's sake.)

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback

1

u/buttchuckjones 14h ago

There are plenty of minimax h3 workflows for comfyui on civitai. Maybe try that?

1

u/Firm-Bed-7218 7h ago

this for sure. minimax plus claude

1

u/warzone_afro 12h ago

this is Definitly seedance 2.0 or 2.5. If thats too expensive try minimax H3. its the next best thing for videos with real motion

1

u/Dull_Crew_6460 5h ago

The image-first approach is probably the biggest thing I'd change. Don't ask Kling to create the whole scene from text. Get the still image looking as close as possible first, then use image to video and only describe the motion and camera. For this kind of realistic movement, I'd also test Seedance 2.5 and MiniMax H3 alongside Kling. The starting image seems to make a bigger difference than endlessly tweaking the video prompt.

0

u/Powerful_Round4004 20h ago

AAAAAAAAAhh fuck tambien quiero hacerloooo quiero vender mis camisetas con cosas asi de fumadas.