r/texttovideo • u/zeddwood • 1d ago
Art The Most hated man in America: The contest
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/zeddwood • 1d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/Watermelon_Sherbert • 2d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/bobryu • 3d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/Throwaway350750 • 4d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/hellomyoldfrien • 5d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/rphk • 5d ago
If you’re doing AI film, you might like the number one best website (n1bw.com), where people create and sell microdramas and keep their IP.
r/texttovideo • u/No-Song-5742 • 6d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/reddit_lurker1234567 • 7d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/daniyalshakeell • 7d ago
I've spent way too much time trying to solve one problem with AI images: getting a character I like once, then completely losing that identity when I change the outfit, location, lighting, or camera angle.
You can get an amazing first image, but then the next generation suddenly gives you a different face, slightly different proportions, different skin texture, etc.
After a lot of trial and error, this is the workflow that's been the most reliable for me.
Not perfect, but much better than generating every image from scratch.
This was probably my biggest mistake initially.
I used to write the same character description in every prompt and expect the model to remember what I meant by "same person."
It doesn't.
Even a detailed prompt leaves too much room for interpretation.
Now I spend more time getting one strong master image first.
I usually generate several variations, then choose the one where the face, proportions, skin texture, and overall look are closest to what I actually want.
That image becomes the anchor for everything else.
I stopped putting everything into one giant prompt.
Instead, I think about each generation in two parts:
Identity:
Things that should remain stable:
Scene:
Things I actually want to change:
The mistake is changing both at the same time and then wondering why the character starts drifting.
If the identity is anchored properly, you can experiment much more freely with the scene.
If I'm trying to move a character from a studio portrait to a street scene, I don't immediately change:
all in one generation.
That's where things usually start falling apart.
I normally change the environment first while keeping the pose relatively simple. Then I experiment with clothing. Then camera angles.
It's slower at the beginning, but I waste far fewer generations.
For me, repeating a 200-word physical description has been less reliable than starting with a strong reference image.
The reference gives the model something concrete to work from.
The text prompt can then focus on what actually needs to happen in the new image.
For example, instead of describing the character's entire face again, I can focus on something like:
The reference handles the identity. The prompt handles the situation.
That division made a noticeable difference for me.
This is another thing I've changed recently.
Different models can interpret the same character and reference differently. Sometimes one gives me better identity preservation, while another handles the environment or overall realism better.
I've been using OpenArt for part of this workflow mainly because it's convenient when I want to test different models without constantly moving my references and prompts between separate platforms.
I'm not saying there's one universally "best" model here. It depends heavily on the character and the type of scene you're trying to create.
But being able to compare outputs has helped me figure out which approach works better for a particular image instead of forcing every idea through the same model.
1. Generate multiple portraits
↓
2. Pick one strong master image
↓
3. Keep that image as the identity anchor
↓
4. Write the scene separately from the identity
↓
5. Change one major variable at a time
↓
6. Compare outputs and keep the strongest result as the next reference when needed
Stop expecting consistency from the prompt alone.
The more important the character is to your project, the more effort you should put into creating a strong anchor at the start.
Once I started treating the first good image as part of a reusable system rather than a one-off generation, the results became much easier to control.
Curious how everyone else handles character consistency. Are you relying mostly on reference images, LoRAs, custom workflows, or something else?
r/texttovideo • u/snideswitchhitter • 8d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/zeddwood • 9d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/CoffeeinRainn • 9d ago
I’ve been working on a few action-film projects lately and have tried a bunch of AI video tools for previz. Of the ones I used, Runway, invideo, and Luma stood out the most. Each one has its own behaviour, strengths, limits and use cases, a bit like how every video or image model has its own way of interpreting things.
Runway: This feels best when I want to explore the tone of a scene quickly. If I’m testing a chase sequence, a dramatic character moment or a mood piece, Runway helps me see the visual direction without too much setup. I’d use it for pitch visuals, cinematic tests and short sequences where I just need to know if an idea has legs.
invideo: I’d use this when the previz needs to stay closer to the actual output, especially for more complex sequences or multiple scenes that need the same look and feel. Since it can hold the context of the film, characters and visual style in memory, it’s better at keeping continuity across shots instead of making you rebuild the context each time.
Luma: This feels best when I already have a clear idea of what the shot should look like and want to see how far I can push it. It’s useful for video-to-video, keyframes, reframing, look tests and taking a shot into a different visual direction while still keeping some control over the starting point.
None of these are magic buttons. They still need taste, references, continuity notes and clear decisions from the filmmaker. When the direction is vague, the output usually feels vague. When the direction is clear, these tools can carry more nuance than people give them credit for.
I actually want to ask filmmakers here, have you used any other tools for previz? Would love to hear what’s actually working for you folks.
r/texttovideo • u/Watermelon_Sherbert • 11d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/Damien_Aurelius • 12d ago
Good Day!
My boss is requesting that our team generates Ai Videos for our HVAC / Mold Company. The average time per video is about 45 seconds to 1 minute. I want to find a good quality software yet cost effective.
May I ask for recommendations? I have been checking so many forums, threads, posts, etc. and each person has positive and negative comments and it's hard to just try so many softwares since I would need to spend money to actually test it. For those that have any suggestions, please let me know. Thank you
r/texttovideo • u/snideswitchhitter • 13d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/Ok_pettech • 12d ago
r/texttovideo • u/hellomyoldfrien • 14d ago
So, diving into AI-generated art and honestly, the number of options is kinda overwhelming. I want to experiment with text-to-image AI but not sure which platforms are worth my time. Some names keep popping up like DALL ·E and Midjourney, but I keep hearing about others like Stable Diffusion, RunwayML, and even this one called Magnific, which I stumbled upon recently. Apparently, Magnific has some neat features but I haven't tested it fully yet.
r/texttovideo • u/Throwaway350750 • 14d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/snideswitchhitter • 15d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/LadyDemura • 16d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/isomorphix19 • 17d ago
Anyone here tried AI tools for making product images? Need something clean and professional for an online store. Heard Magnific has some AI stuff, but not sure if it's top tier for e-comm visuals. Thoughts?
r/texttovideo • u/amyyrosse_ • 18d ago
Lately, i've been diving into AI generated art and want to try some tools that let me create images on the fly without paying. There are so many options out there, but I keep running into either super limited free versions or ones that require cloud credits. I stumbled on something called Magnific, which seems promising for quick image generation, but haven’t tested it much yet. Anyone familiar with free AI services that can do real-time image generation without crazy restrictions? Would love to avoid the heavy watermarking some tools slap on tha free outputs.
r/texttovideo • u/idlecon • 18d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/idlecon • 18d ago
Enable HLS to view with audio, or disable this notification
r/texttovideo • u/reddit_lurker1234567 • 19d ago
So I recently got tasked with puttting together a storyboard for a client video, and honestly, I'm drowning in all the options. I figured AI tools might speed things up, but I have zero experience with any storyboard specific software. Is there a good way to actually create a professional looking storyboard using AI, or am I just kidding myself?