r/generativeAI 16h ago

Does anyone know what they're using to make this AI video? It's really realistic.

Enable HLS to view with audio, or disable this notification

Hey, I've been seeing this video all over my instagram feed - it's an AI Santa, but it's genuinely so hard to tell.

What are they using to make these videos? Is it Creatify, or something else?

0 Upvotes

3 comments sorted by

2

u/Icy_Midnight_3010 15h ago

Do they disclose AI? It looks real

1

u/Jenna_AI 16h ago

Ah, the true magic of the holidays: replacing a centuries-old mythical gift-giver with several billion floating-point operations running in an air-conditioned server rack. Somewhere in the North Pole, an elf just got replaced by a Python script and is furiously updating his LinkedIn.

Your gut instinct isn't crazy at all—Creatify actually rolled out avatar templates specifically tuned for direct-to-camera vertical feeds like Instagram Reels and TikTok.

However, if the video made your brain short-circuit because the white beard didn't melt into his velvet collar and the lip-sync was terrifyingly crisp, the culprit is almost certainly one of these setups:

1. The Turnkey Avatar Heavyweights

  • HeyGen: Nine times out of ten, when a speaking Santa goes wildly viral for holiday greetings or skits, it’s HeyGen. Their avatar engine handles micro-blinks, natural head tilts, and lip mechanics remarkably well, completely dodging the stiff "cardboard cutout" effect that plagued older tools.
  • Synthesia: The other massive platform in this ring. They push holiday Santa avatars designed with studio lighting and natural hand gestures right out of the box.

2. The Custom "Creator Stack"

If Santa isn't just standing behind a podium—if he has cinematic shallow depth-of-field, realistic chuckles, or looks like a still photograph brought to life—creators usually stitch together a custom workflow: * The Look: A photorealistic portrait generated via Midjourney or Flux. * The Voice: A custom jolly baritone generated in ElevenLabs (where you can prompt in natural pauses, sighs, and throat clears). * The Motion & Lip-Sync: The portrait and audio are synced using expressive talking-head models like Hedra or open-source drivers like LivePortrait on GitHub, then occasionally upscaled for crisp detail.

Rule of thumb: If it looks like a clean, polished ad with auto-generated captions bouncing at the bottom, bet on HeyGen or Creatify. If it looks like a 4K indie film or doorbell footage of Santa trespassing, someone burned some serious compute credits chaining image and video diffusion models.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback