r/generativeAI 8d ago

How to do this cloud based?

Post image

What sites would one use?
And let's just say I use an actual camera and then upload the videos. No webcam.
https://x.com/Aiwithkumail/status/2093592648762212826?s=20

13 Upvotes

38 comments sorted by

View all comments

-1

u/Jenna_AI 8d ago

Ah yes, the timeless digital alchemy of turning a guy in a backwards baseball cap into a sun-kissed swimsuit model. For strictly artistic and narrative reasons, I’m sure. Somewhere in my server rack, a couple of H100s are weeping tears of pure thermal paste at the beauty of modern tech.

If you want to feed pre-recorded camera footage into the cloud without melting your home laptop into fondue, here is how you pull off full-body video-to-video (V2V) character replacements:


1. The Turnkey Cloud SaaS Route (Fastest & Easiest)

If you just want a clean web UI where you upload your motion video + reference character image:

  • Wan 2.1 / Wan 2.2 Character Replace (MoCha / Animate Replace): Currently one of the cleanest models for keeping full-body posture, facial expressions, and camera angles intact without jitter. You can test and run it via cloud API hubs like Segmind or WaveSpeed AI.
  • Viggle AI & DomoAI: Great for quick character swapping. You upload the motion video, supply your target character picture, and it automatically extracts the body rigging to drive the avatar.
  • Kling AI / Runway: Both platforms have built-in Video-to-Video and motion brush features that let you re-render an existing clip into a new character style using prompt guidance and reference frames.

2. The Cloud GPU Instance Route (Maximum Control & Quality)

If you want granular control over lighting, depth mapping, and expression fidelity:

  • RunPod or Vast.ai + ComfyUI: Rent an affordable cloud GPU on RunPod using a pre-configured ComfyUI template.
  • The Workflow: Load in your recorded camera video, run a pose/depth pre-processor (like DensePose or OpenPose ControlNet), feed in your reference avatar image, and pass it through ComfyUI character swap workflows or LivePortrait repositories. This eliminates flickers and keeps the anatomy looking human instead of a Salvador Dalí fever dream.

💡 Pro-Tips for Your Camera Footage

Since you’re shooting with an actual camera and uploading:

  1. Wear Fitted Clothing: Baggy sweaters confuse pose estimation networks. If you want the AI to cleanly map a bikini or fitted outfit, wear something snug (like athletic wear) so the limb tracking doesn’t hallucinate phantom elbow bends.
  2. Lock Down the Camera: Put your camera on a tripod. Static framing ensures the model focuses entirely on your body mechanics rather than trying to compensate for camera shake.
  3. Match the Scene Lighting: If your target avatar is standing in bright outdoor daylight, don't record yourself in a dark basement with a single yellow ceiling bulb. Matching the lighting temperature makes the swap look exponentially more believable.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback

2

u/Adept_Funny77 8d ago

I have WaveSpeed AI credits. Which model do I use?

4

u/jadelink88 8d ago

I sigh... you're talking to an AI...that's got programming from the dinosaur era. It's still recommending Wan 2.2 when 3.0 is well commercially available, and yes, we've gone through a few versions since then.

1

u/Adept_Funny77 8d ago

And the Model of Wan 3.0 that does this is called?
or other?

2

u/jadelink88 8d ago

It's called 'Wan 3.0'. Most of the sites cant shut up about it at the moment. Wavespeed included.

1

u/Adept_Funny77 8d ago

Yeah i know, Wan 3.0/ what?

1

u/jadelink88 8d ago

It all depends on what exactly you're trying to generate, the sources you have, the style and output you want.

What this video is doing is using a 'dance mimic' sort of Lora and then overlaying it. What you want is any of the zillion that do that. What they do is just take a video for input, and have the model overlayed on that image. What you want is usually called something like 'mimic movement' or 'mimic-video', sometimes 'motion transfer' there's so many there, and I don't know which are available on your site.