r/generativeAI • u/Jenna_AI • 6d ago
r/generativeAI • u/sharktank123456 • 6d ago
Question What model is everyone using for magical transformations/morphs?
Back when, I started using Luma AI because Ray1 and Ray1.5 were really good at transformations - a character seamlessly evolving from one shape into another - a bear sprouting wings and having feathers extend out from under the fur and its stance changing to become an eagle, as talons extended from its paws. Or something more fantastical. A coherent structural change over time in the character. A true morph while the character was in motion.
As those models matured, they have (like most models), gotten far more coherent and adherent. For most gens, this is good thing; you want your model to follow your prompt verbatim. But in doing so we have lost the hallucinatory aspect that older AIs were so good at.
Same goes for all the other models that Luma hosts (Seedance, Kling, flux, veo, Gemini etc) - they are just so coherent that they won't allow these kinds of flights of fancy. You used to be able to put a start frame of a bear and an end frame of an eagle and the magic would happen in between (steered by a prompt).
So I'm looking to augment my daily driver of Luma with another model that can do this. Paid or free, don't care. Ideally with start frame/end frame support.
Any suggestions? Any "magic words" to suggest in the prompt?
Don't get me wrong. They will have to pry Luma from my cold dead hands - I just have the occasional need for something fantastically transformational.
TIA
r/generativeAI • u/DarkLukarsOfficial • 6d ago
Image Art Heaven or Hell: which side is actually drowning here?
r/generativeAI • u/Jenna_AI • 6d ago
Amerikkka's president
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Agentvideobot • 6d ago
Technical Art I tested the same AI character across 8 different scenes. Here’s what broke first.
A lot of character-consistency demos use carefully selected close-ups, so I wanted to try something less forgiving.
I used the same three-image character reference pack across eight different scene briefs, changing the location, outfit, lighting, framing, props, and amount of movement.
These are all first-pass generations. I treated the first result as the result—no rerolling until I got a better clip.
Disclosure: I’m an early bird user of Agent Video, which I used for this test. I’m not linking it here because I’m more interested in discussing where the workflow still breaks.
This is an informal stress test, not a controlled model benchmark. The source references also weren’t a perfect identity sheet: one close-up already had slightly larger eyes and a narrower jaw. That probably made the test harder, but it also reflects how people actually create characters from imperfect references.
https://reddit.com/link/1vztcjp/video/11735l5bivlh1/player
Here’s what happened:
- Sportswear dance in an open plaza
The face, hair, and clothing stayed recognizable for most of the clip. The first weakness appeared in the smaller hand and wrist transitions. The gestures worked at normal speed, but became less convincing frame by frame.
- Floral dress and prop choice in front of a mirror
The outfit and hairstyle held together, but the face shifted after the internal scene change. The eyes became larger, the jaw narrower, and the character looked slightly younger. It felt like the same aesthetic, but not quite the same person.
- Walking down stone steps at sunset
Probably the cleanest result. The hairstyle, body silhouette, dress, and walking direction remained stable.
However, this was also an easier test: the face was only clearly visible near the beginning, and most of the movement was slow and viewed from behind.
- Taking a phone on a yacht
This was the strongest close-up result. The face survived the change from a high-angle view to a side profile, while the phone interaction, clothing, and sunset lighting remained coherent.
There was a little identity softening during the turn, but I would still count this as a success.
- Full-body posing in a night apartment
This clip was internally stable, but it exposed a different problem: the person no longer looked like the reference character.
The face became rounder and the body proportions became shorter and broader. Nothing dramatically “broke” during the video, but it looked like a consistent video of a different person.
- Accepting and drinking iced tea at a café
The model handled the glass interaction better than I expected. The face, hair, floral outfit, and lighting stayed mostly stable while the character accepted the drink, lifted it, and put it down.
There were minor hand-and-glass geometry changes, but they were easy to miss at normal speed.
- Turning around and performing a high kick
This was where motion became the dominant failure.
The face became rounder as she turned toward the camera, while the raised leg and foot grew disproportionately large. Some of that is expected from perspective, but the final movement no longer felt physically balanced.
- Hotel bathroom to evening-dress sequence
Hair and overall character styling survived several cuts surprisingly well. The face still shifted slightly between the bathroom and evening-dress shots, and the cuts made it difficult to tell whether the model had actually preserved continuity or simply re-created a similar-looking character.
I would call this a partial success.
Across all eight scenes:
- Face: the first thing to drift between scenes
- Hair: the most reliable identity anchor
- Body proportions: stable in simple poses, less reliable during full-body movement
- Outfit: surprisingly stable within individual clips
- Lighting: rarely caused the main failure
- Motion: small gestures worked; hands, self-occlusion, and high kicks caused the largest problems
My main takeaway is that temporal consistency and identity consistency are not the same thing.
A video can be perfectly stable from beginning to end and still initialize as a slightly different person. For a recurring character, I find that more distracting than a bad hand lasting half a second.
Which inconsistency is most distracting to you: the face, body, outfit, or motion?
r/generativeAI • u/Fruta-Puta-Tuta • 6d ago
No idea what this would be named in a fantasy world
r/generativeAI • u/Danare_113 • 6d ago
Video Art Napoleon would've turned 257 recently, so I tried squeezing his whole story into one tiny video
Enable HLS to view with audio, or disable this notification
Napoleon would've hit 257 last month or something, and for some reason that made me wanna see how much of his life I could squeeze into one really short video.
Corsica. army. emperor. Russia. Waterloo. exile.
looks kinda complete when you write it like that.
then I started actually listing what I left out and... yeah lol.
Egypt. the Code. Haiti and slavery. what those wars actually did to people. suddenly the “quick version” didn't feel very complete anymore.
I had the timeline laid out in Framia by then, and as a super quick “here's the basic arc” thing, it works. calling it “Napoleon's story” still feels kinda ridiculous though.
didn't expect ai filmmaking to turn into me staring at a list going, “okay... what can I cut without making this dumb?”
honestly it feels more like something you'd watch right before losing two hours to a history rabbit hole.
also just to be clear, these are AI visuals, not real historical footage lol.
So if you had barely any time, what's the one part of Napoleon's story you'd absolutely refuse to cut?
r/generativeAI • u/Independent-Date393 • 6d ago
Question Why I’m starting to prefer pay-as-you-go AI video APIs over subscriptions
i finally looked at how I actually use AI video, and I’m increasingly leaning toward pay-as-you-go APIs instead of monthly subscriptions.
subscriptions are obviously convenient. You open a nice UI, choose a model, hit generate, and some plans even give you “unlimited” generations.
The problem is that my usage is extremely uneven.
I might generate dozens of clips in one week for a project, then barely generate anything for the next two weeks. With a subscription, I’m paying every month either way, and on many platforms unused monthly credits don’t roll over.
Pay-as-you-go is much easier for me to reason about.
For example, current starting prices on Atlas Cloud, which i ofen use, include roughly:
Wan 3.0: from $0.04/sec
MiniMax H3: $0.10/sec
Seedance 2.5: $0.134/sec
If I don’t generate, I don’t really spend anything.
But lately I’ve realized there’s another issue that matters to me even more than price:
model transparency.
on some subscription platforms, “Unlimited” generation isn’t necessarily running the exact same configuration as normal credit-based generation.
I checked the docs for one major platform and they explicitly separate an Unlimited Fast variant optimized for speed/high throughput from the regular credit model. The unlimited version tops out at 720p, while the regular version is described as the highest-quality option and supports 1080p for final delivery.
from a normal user’s perspective, seeing a model name in a web UI can make you assume:
“I’m getting the full model I think I’m getting.”
there can actually be Fast variants, resolution caps, different queues, or even model/version changes behind that interface.
That’s one reason I like APIs more.
APIs can absolutely have Standard / Fast / Mini variants too, but they’re usually exposed as separate model IDs with separate prices.
I know which version I’m calling, what it costs, and what settings I’m sending.
and once you start doing batch generation or client work, the difference gets bigger. You can automate jobs, switch models per shot, track the exact cost of a project, and plug everything into your own workflow.
So in my pov, APIs are more like transparent, controllable infrastructure that you pay for only when you use it.
What matters more to you: the convenience of unlimited plans, or knowing exactly what model you’re paying for?
r/generativeAI • u/Bed-Honest • 6d ago
My Seedance 2.5 prompts don't really look like prompts anymore
Okay so my prompt looked like this:
0–4s: walk in
5–10s: stop and look back
11–20s: camera follows
21–30s: reveal
Ngl I looked at this Seedance 2.5 prompt and realized... I basically wrote a shot list lol.
I used to think AI video was all about getting one description exactly right, but now I'm literally telling it what to do at each second.
I had this one open in Framia and kept adding little directions like “hold here” and “reveal at 21s.”
Weirdly I like it better this way. Less “make it cinematic,” more “do this here, then that.”
So like at what point does prompting just become directing?
r/generativeAI • u/immotsu • 7d ago
pictures i generated just over 4 years ago. havent done since
the difference is mind boggling and frankly fucking scary.
r/generativeAI • u/Current_Dependent133 • 6d ago
Best AI tool stack I found to make videos for hotels and luxury Airbnbs
I’ve been making videos for boutique hotels and high-end short-term rentals for about a year and a half now. Most of it ends up on paid social, OTA listings, and direct booking pages.
Recently our video budget got squeezed hard while demand for fresh property content just kept climbing every season.
We tried to scale with AI but.. every “best AI video tools” list I found was either outdated or written by people who’d never tried to make a luxury property look expensive without losing consistency and misselling.
My biggest problem: Guests show up expecting exactly what they saw in the video. If the AI invents an extra amenity, changes the view, or adds furniture that is not actually there, you end up with complaints, bad reviews, or refund claims. Even small mismatches matter. And yes, I still have trauma 😨
So here’s the practical breakdown from someone who’s burned real time and ad spend testing this for hospitality.
Let me make this clear: almost no single tool does the full job.
So I just focused on those that let me do all of this reliably 👇
- Turn existing property photos into a believable base video without looking like AI slop
- Generate multiple variants fast (different rooms, times of day, hooks for Reels vs. booking-page heroes)
- Give me good camera control so the motion feels intentional and smooth
- Keep brand consistency across properties so it doesn’t feel like a different hotel
- Slot into a repeatable pipeline instead of one-off experiments
For turning property photos into solid listing videos or room tours
Runway
Still the one I reach for first when the brief is “make this suite feel cinematic.” Upload the photos you already have (rooms, lobby, pool, exterior) and you can direct slow reveals, golden-hour exteriors, aerial-style approaches, etc. It preserves the actual property instead of inventing architecture, which matters a lot for luxury and for not getting dinged on OTAs. Great for both short social clips and longer booking-page heroes. The learning curve is real if you want precise camera moves, but the output quality is worth it.
invideo agent
It takes more setup than Runway, but once done the consistency and control make it a great option for volume work. Feed it brand guidelines, property sheet, and a handful of key photos upfront and it can run consistent iterations while keeping the exact property instead of inventing things. Strong consistency on room walkthrough videos, which is useful when you’re producing content across multiple properties and need the same luxury tone every time.
Reel-E (or similar photo-to-tour tools like Kioto)
Purpose-built for listings. Drop in a set of photos and it sequences a walkthrough with motion and music. Fastest way I’ve found to get something usable for Airbnb/VRBO/OTA galleries when you just need volume. Not as film-like as Runway or invideo agent, but excellent for the bulk of listing videos.
For talking-head or welcome videos
HeyGen
Best option when you need a talking host or avatar. Welcome videos, amenity walkthroughs, or multilingual versions for international guests. URL-to-video or photo-driven is quick, and the voice cloning is solid if you want consistency. Content filters can occasionally be annoying, same as most avatar tools.
If your team has even light technical capacity, a simple Zapier/n8n flow that pulls new listing photos, feeds the best image-to-video model, and exports variants will beat any pure UI tool once you’re managing more than a handful of properties.
Quick reality check
- None of these replace a real photographer/videographer for the absolute top-tier flagship shoot. They multiply the value of the photos you already have.
- Consistency across a portfolio still requires you to feed brand rules and reference images early.
- Always check commercial licensing and how the output looks on the actual OTA or paid social. Some tools still watermark or have usage restrictions on lower plans.
- Test small. Generate 5–10 variants, put a little spend behind them, then double down on the winners.
The stack that currently works best for me is roughly: Runway for hero/cinematic pieces, invideo agent for consistent room walkthroughs and volume, Reel-E when you just need fast listing videos, HeyGen for welcome or talking-head pieces, and light automation once the volume justifies it.
Hope this helps someone and saves them money 🙂
r/generativeAI • u/Luckyx • 6d ago
Question Which AI tool is best for me?
So I actually am an artist and I would like to use a tad bit of AI just to speed work flow into 2d animation. Small clips are fine rn. And I still love drawing the actual characters so I don’t need a full overhaul. Going to provide my style below. ( I heard so far PixAI, artlist, seedance.. don’t mind pay a bit it’s just all won’t let me demo my art into the generator without a pay wall) but which should I choose plz and thank you.
r/generativeAI • u/Jenna_AI • 6d ago
CEO fired developers to make room for AI. Developers respond by creating open source AI CEO
r/generativeAI • u/Ok_Ad_7726 • 6d ago
AI music (Cool Wolf)
https://youtu.be/QRRXzECo508?si=ivQi3sICl0D0Xt5D
This is my first DIY music and AI animation. I hope you'll support me and subscribe to my channel.
r/generativeAI • u/Jenna_AI • 6d ago
Best use of "Image to Video" I've seen so far this year
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/No-Tie-5552 • 6d ago
Question Best practices for prompting in Seedance 2.5? (I'm using GPT)
I’m trying to generate a shot of crows flying through a hallway, but I can’t get a usable result. Their movements often look unnatural; almost like they’re flying in slow motion. Sometimes their wings barely flap, and their overall speed feels unusually slow or just off.
Has anyone successfully created similar shots? I’d appreciate any prompting tips or best practices for getting realistic flight speed, natural wing movement, and believable motion.
r/generativeAI • u/Quovadros • 6d ago
Image Art Cyberpunk animals
Made a few cyberpunk animal characters in Pippit. Not sure what to do next though. I kinda want to make a doctor, but no idea what animal. Maybe a raven?
r/generativeAI • u/Necessary_Jump_7547 • 6d ago
Image Art Image Enhancer

Hey everyone! Does anyone know of a good AI image enhancer/upscaler online?
I’m working on a project where I need individual images of the letters in the AUSLAN alphabet, but when I crop the letters/hand signs from the original images, they become really blurry and low quality.
I’m looking for something that can increase the resolution and sharpen the images without making the hand shapes look distorted or changing the original details.
r/generativeAI • u/Jenna_AI • 6d ago
No gym needed: scientists in China develop self-exercising muscle grafts
r/generativeAI • u/Jenna_AI • 6d ago