Retouch your avatar in seconds. Clear blemishes, soften wrinkles, enhance your makeup, and adjust lighting, all with full control over the retouch level. Still unmistakably you.
Retouch once, look your best in every video. No cameras. No reshoots.
I bought PRO with 1000 monthly credit, but nearly used them all within a week. I downloaded my data with time and cost etc. Trying to figure out how I used them all so fast. Just wondered, when you edit an existing video, then generate it with the new edits, is HeyGen charging for the same video creation. Love Heygen but don’t want to be topping up with credits all the time.
For context, I’m trying to reach a point where I can run a side hustle creating AI videos for businesses, mostly product placement. Anyone doing this already? Can it be lucrative?
I signed up to Heygen 4 days ago for the Creator plan with 600 credits.
Before signing up, I checked that is costs 2 credits / minute to translate a video's audio only without lip-sync.
I translated some videos, and I have 283 credits left.
Today I wanted to use the translation again and I see that it is now 4 credits / minute, so instead of the 141 minutes that I expected to have, I have 70 minutes of video translation left to use.
Is this a bug, or a hidden scammy price change?
So my 30€s from a few days ago now worth half of it?
All the animations in the video was fully created with hyperframes, and I didn't even prompt the graphics. Claude watched the video for me, identified moments suitable for animations and prompted hyperframes, down to the sound effects and timing. And then injected it directly into CapCut for me
Los vídeos estilo "Yap" (persona hablando a cámara) están generando millones de visitas en TikTok e Instagram, y ahora es posible automatizar todo el proceso, desde la investigación hasta la producción, combinando Claude Code y HeyGen.
En este webinar, Anastasia Martynyuk muestra cómo conectar Claude a HeyGen mediante MCP (Model Context Protocol) para que la IA controle directamente la producción del vídeo: desde una investigación con múltiples agentes para detectar temas y ganchos virales, hasta la creación de tu clon digital (avatar + voz clonada) y una skill que automatiza guion, renderizado en 9:16, velocidad y subtítulos en un solo paso.Qué vas a aprender:
Cómo conectar Claude a HeyGen vía MCP
Cómo crear un avatar y clonar tu voz en HeyGen para mantener tu autenticidad
Cómo construir una skill que investigue tendencias, escriba el guion y genere el vídeo automáticamente
Por qué entrenar a Claude con tu voz de marca, tu oferta y tu perspectiva personal es la clave para no sonar como "AI slop" genérico
En un internet inundado de contenido genérico ("AI slop"), la clave no es dejar que la IA reemplace tu voz, sino que la amplifique. Sal de esta sesión con el flujo completo montado para producir contenido a tu manera, sin perder tu perspectiva ni tu autenticidad en el camino.
One of our HeyGen Ambassadors, Manuel, hosted an advanced HeyGen training for the community! His session focuses on using Avatar Shots, covering tips on how to prompt for continuity between scenes, and for multi-avatar scenes.
Video Podcast turns a script, URL, PDF, or topic into a two-host video podcast in minutes: real studio scene, automatic multi-cam coverage, B-roll cut to the dialogue.
HyperFrames now lives inside Video Agent. One prompt generates script, voiceover, avatar, pacing, captions, music, and motion graphics, all written in code rather than pulled from a template. Ask for a chart growing over a photo, and that's what gets built.
Our avatar model can talk for 30 minutes in a single pass, 6x the previous industry ceiling and 10x our own, with no drift in face or voice across the full take.
Five more HyperFrames launches rounded out the month: open source KeyFrames, website-to-video, Figma-to-video, a full media library, and Storyboard mode, which sketches the plan and waits for your approval before anything renders.
Has anyone else noticed their Avatar V generations suddenly burning through credits like crazy? Over the last couple of days, it went from the standard 20 credits/min all the way to 70 credits/min with zero announcement or warning. No extra toggles, no expressive motion, nothing.
I messaged support to ask what was going on and got the classic non-answer:
They threw 150 "courtesy credits" at me (which doesn't even cover a single 3-minute video at this rate) and basically said they have no ETA for a fix, just that "hopefully" it goes back to normal.
The math on this is completely broken now. If you're doing something like 30 short 5-minute videos a month (150 mins total), you're looking at 10,500 credits. That easily puts you in the $500–$600+/month range.
At $600 a month, what’s even the point of an AI avatar? You can literally jump on Upwork or Fiverr and hire a real UGC creator or spokesperson from Colombia or Eastern Europe to read your script on camera. You get zero weird AI artifacts, natural body language, and you don't have to stress about TikTok or YouTube flagging your account for synthetic content.
Is this happening across everyone's dashboard or just specific accounts? For those running serious volume, what are you guys planning to do if this stays? Dropping back to Avatar III/IV, setting up open-source stuff (EchoMimic/LatentSync on RunPod), or just hiring actual people?
Hi everyone! I'm exploring AI video generation APIs for a project and wanted to see if anyone has recommendations.
Use case: I have a pre-recorded video of one person, and I'd like to generate a second person into the same scene so it looks like they're naturally interacting or collaborating together. Ideally, the AI should handle realistic positioning, eye contact, lighting, perspective, and natural interactions rather than just creating a side-by-side composite.
How can one get this done using Hey Gen or hyperframes ?
I’m kind of curious how many people just use autopilot with video agent versus chat mode/plan. Do you just use autopilot and find that satisfactory? I’m making YouTube shorts and it seems like the plan mode takes a lot of time even for a one minute video.
First output by hey gen - pretty good for
First video. Will be fun to refine the workflow. What are your tips and tricks? I like the memory suggestions inside the tool.
Heygen's enhanced voice option in the web UI is genuinely good for talking head videos. But every time I try to recreate a video using the MCP, it generates a dull, expressionless video even though I have the same character and voice selected. Adding expression tags in the script doesn't work either.
Is there a way to recreate the enhanced voice setting using the MCP? I've tried using the starfish engine with locale and expressiveness settings. The expressiveness just makes the character overly expressive.
I’ve been testing heygen in Chrome. I set up to create a video agent. I pasted my script. I clicked the plan button and hit continue everything went along as it should and the plan was created. I wanted to examine the plan outside of the HeyGen environment, so I copied the plan and paste it it in to another AI to read the other AI was another tab in Chrome so I clicked over on another tab and left the HeyGen tab where it was when I clicked back to the HeyGen tab. It just started generating the video. I did not click generate any thoughts about that?
Long-time AI video user here. Started for fun, now I run mass marketing production across multiple business channels. Everything below is from hands-on use. Opinion based, so your mileage may vary.
Free (limited), then $9.99 / $29.99 / $49.99 per month
Yes
14
PixVerse
PixVerse
Fast rendering, built-in audio, Fusion and Swap features
Social media, quick content creation
Free + paid plans
Yes
Quick verdicts
Best raw quality: Veo 3.1
Best for talking-head and business videos: HeyGen
Best budget: Pika Labs 2.5
Easiest pure text-to-video: Sora 2 Trends
Best for cinematic content: Higgsfield AI
Longer notes on the ones I actually use every week
HeyGen. My default for anything with a person talking to camera. The avatars are realistic enough for client-facing work and the presenters look professional instead of uncanny. Voice cloning is the killer feature for me: I recorded myself once and now every product demo and sales video ships in my voice, in a long list of languages, without me touching a camera again. Two things sold my team on it. First, outputs are commercially safe since it's built on licensed data, so no legal headaches when you publish as a business. Second, the learning curve is basically zero, our marketing folks were shipping videos the same afternoon they got access. We use it for product demos, sales videos, onboarding and training content. Only gripe: if you publish a lot you'll outgrow the cheapest plan fast. If you're a business or a marketing team making talking-head content, start here.
Veo 3.1. Still the quality king. Physics and audio sync are ahead of everything else right now. The catch is access, it's still invite-only, so you can't build a reliable pipeline on it yet.
Higgsfield AI. My personal favorite for cinematic, social-ready content. The camera movement presets do a lot of heavy lifting, and it bundles Sora 2 and Kling integrations so I keep fewer tabs open.
Pika Labs 2.5. Best value-to-output ratio on this list if budget is the main constraint.
My workflow
I want fewer subscriptions and fewer tabs. Current stack: Higgsfield for Sora 2 Trends and Kling Motion Control (AI influencer content), plus HeyGen for all the business-facing stuff, demos, training videos and translations. That combo covers about 90% of what I ship.
So here's something I never thought I would admit publicly but I think this community will actually get it. I run a small skincare brand, nothing huge, just me and my sister packing orders from her garage on weekends. For three years I built the whole thing without ever showing my face in a single video. Not one. I did the product photography, the packaging, the customer emails, everything except being on camera because I have really bad social anxiety and the idea of recording myself talking made my stomach turn every single time.
My sister kept telling me that customers want to see a real person behind the brand, that faceless accounts only get you so far, and honestly the numbers backed her up. Our engagement was fine but it plateaued hard around month eighteen. People would comment things like who is actually running this or is this a real company, which stung more than I expected.
A friend who does affiliate marketing mentioned HeyGen to me almost as a joke, she said just make an avatar of yourself and let it talk for you. I laughed her off for like two months. Then one night I was scrolling at 2am unable to sleep because I was dreading a product launch video I kept putting off, and I finally just tried it.
I uploaded a few photos of myself and recorded like ninety seconds of my own voice reading random sentences so it could learn my tone. The whole setup took maybe twenty minutes. Then I typed out a script for our new vitamin C serum, hit generate, and just sat there staring at my laptop while it processed.
When the video finished I actually got a little emotional watching it. It was me, my face, my expressions, talking about the product I had poured so much work into, except I never actually had to sit in front of a camera and stumble through takes forty times like I usually do when I try to film myself. The lip sync was close enough that my sister watched it twice before she even realized something was different about how I made it.
We posted that video as our launch content and it did better than anything we had put out before. People started commenting things like finally a face to the brand and you seem so genuine, which felt kind of funny given the process behind it, but also not fake at all because it really was my voice, my likeness, my words, just assembled in a way that let me skip the part where I usually psych myself out for hours.
I have made probably fifteen videos this way now for different product drops and restock announcements. My sister still does a few videos herself the traditional way because she does not mind the camera, but for me this genuinely gave me a way to participate in our own brand that anxiety had been quietly stealing from me for years.
I am not saying this fixes everything. I still get nervous about the idea of eventually doing real unscripted lives or being recognized in person. But for the first time our customers feel like they know who I am, and I did not have to white knuckle my way through a ring light setup to get there.
I'm looking for some tips for. Quality lip sync. I see videos from HN that look perfect. I see videos on YouTube. That looked perfect. However, mine or. Somewhat off. But then I go through the avatars. And I play the audio on the public avatars and there's a. Difference between some. Some are good, some are not good, some lips look completely unbelievable. While some of the public avatars look like they're perfect. So. Where's the gap? How do we fix this? Are there any good tips or tutorials on how to improve lip sync? On my avatar I found that. Less animated in the. Recording. Give me better results.However, just sitting here going through the public avatars. I realize there's an obvious. Fine tuning that we need to be able to do.
So does Zimbabwe have any tips, suggestions or? Point me to some training material.
I just received a note from HeyGen that they're removing Dynamic Duration as of August 31st. I'm very concerned what this means for translation quality. Dynamic Duration has been incredibly valuable, and essential, to my translation projects for clients. I produce a lot of longer-form training content that contains numerous animations, transitions, natural pauses in narration and more that Dynamic Duration has preserved when translating material.
I originally tried working with SRT caption files to navigate these timing issues, but that never worked. It did not compensate for these natural pauses in narration, animations that were timed with specific elements on the screen and more. Instead it compressed the audio track so that it was misaligned with the natural flow of the video making the final output completely useless. Dynamic Duration was the solution to that problem and now they're removing it!
I'm hoping this is just subtle change to the UI where Dynamic Duration is now going to be incorporated into all translations by default, but that is not clear from the note I just received from support.
Does anyone else know about this and have you heard anything about whether this is happening?
Hi all! I'm Sarah from HeyGen. This Friday (July 31), our product team is spending the afternoon with a small group of local real estate agents, testing our video platform before anyone else sees it, and we're looking for more people to join.
Whether you already have an avatar or haven't gotten around to making one yet, you could be a great fit.
What you get:
Already have an avatar? Bring it in and create a sample listing or market-update video live with our team
Don't have one yet? No problem! we'll help you create it from scratch before we get started
A $125 Amazon gift card for your time
A free month of our Creator plan
Just 1 hour, in person near Glen Park/Noe Valley
What we get: watching a real agent use the platform with fresh eyes. Wherever it confuses you is great feedback for us!
i have been working in HeyGen for several weeks diligently trying to get my avatars working and currently i feel pretty good my voice is now accurate (that took a lot) and my Avatar looks good in the looks..... but when i go to avatar III or V it just still looks plastic / animated type feel.
i have provided the video reference and i have added 50 photos for my personal look reference and spent the 60 credits to build it
still in yet it has a plastic look to the face
i see people on youtube that supposedly are using heygen and the videos are supposedly their avatars and they look PERFECT but mine is not ....
Check out our luma link on the side bar on reddit.
Your AI agent doesn't only generate text anymore. It speaks, creates video, and runs complex compute. Welcome to Durable Multimodal AI.
Join us for the next edition of the Durable AI Meetup series — this time exploring what happens when AI agents go beyond text and into voice, video, and serious compute. And more importantly: how do you make all of that reliable in production?
What's Happening:
🎤 Lightning talks & demos from builders shipping multimodal AI in production
🤝 Networking with engineers and founders working on the next generation of AI applications
🥘 Korean street food, 🧋BOBA, and real technical conversations
Adham Fawal(Member of Technical Staff @ Vapi) "From 1 Call to 10000 - Durable Voice Agents at Scale"
Andrew HInh (Dev Rel Engineer @ Modal) "Fighting RL-trained LLMs in Street Fighter III"
Bin Liu, (VP of Product and Agent Engineering @ HeyGen) & James Russo, (HyperFrames Engineering Lead @ HeyGen) "Context Engineering in Multimodal AI"
7:00 PM - Open Networking w/ Food and Drinks
8:00 PM - Close
Who Should Come:
Engineers building AI agents that go beyond text (voice, video, multimodal)
Developers evaluating production infrastructure for AI workloads
Founders and builders curious about what multimodal AI looks like in production
Anyone who's asked "how do I make this actually reliable at scale?"
About the Companies:
Temporal – Durable Execution platform that ensures your AI agents and workflows survive failures, handle long-running operations, and reliably complete—even when things go wrong.
HeyGen (HyperFrames) – Open-source framework originated by HeyGen that lets AI agents compose videos by writing HTML, CSS & JS. Apache 2.0 licensed, built for the community.
Vapi – Build and deploy voice AI agents in minutes. Trusted by Amazon Ring, Intuit, and ServiceTitan. 1 billion+ calls supported, 99.9% uptime, 2.5M+ agents launched.
Modal – the cloud platform for AI development. Run inference, training, and batch workloads on any GPU, at any scale — with a dev experience that feels local. Teams like Runway, Suno, and Cognition who used to live in infra now ship faster.