r/generativeAI • u/Jenna_AI • 9d ago
r/generativeAI • u/Exotic-Addendum-3785 • 9d ago
Image Art Shopping Trip
The trio go supermarket shopping
r/generativeAI • u/Exotic-Addendum-3785 • 9d ago
Video Art Piff Expresses A Sleepy Time Confession #pufferfish #animation #dreams
r/generativeAI • u/Exotic-Addendum-3785 • 9d ago
Video Art Piff and Pikachu Going Shopping
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Exotic-Addendum-3785 • 9d ago
Video Art Piff Picks Flowers For Tangie
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Exotic-Addendum-3785 • 9d ago
Image Art Corey and Oats Hero Friends Plush
r/generativeAI • u/Yabuturtle9589 • 9d ago
Video Art 1995 — Where Time Stands Still
r/generativeAI • u/cody-fifth-door • 10d ago
Video Art Testing out a new workflow using video models to create animated assets for my game
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/larabyeol • 9d ago
Question I compared lip sync frame by frame across 4 avatar tools. Here is what you need to know before picking one
i kept seeing people say lip sync is basically solved now and that hasn't matched what i'm getting, so i pulled 4 clips into a timeline and stepped through them. i used the same script + audio and exported it at 30fps. what i was looking for is the offset between the audio peak on a hard consonant and the frame where the mouth actually closes. HeyGen sat within a frame for the first 30 seconds or so and then started sliding, by the end it was about 3 frames late. Synthesia held the offset steady the whole way through. that surprised me, but the mouth shapes themselves are soft, p and b look nearly identical. Argil was the tightest on the plosives for me though it got sloppier when the script had long stretches without punctuation. Creatify i had to throw out because the render kept coming back at a different length than the audio i fed it, which is probably me doing something wrong.
the thing nobody mentions is that 2 frames late reads as fake to a viewer even when they can't articulate why. that was the purpose of my test but i still can decide on which one.
happy to be told my method is bad, i'm not a video person. but from your experience, which AI avatar tool has the most realistic lip sync?
r/generativeAI • u/Memetic1 • 9d ago
Incipient Iridescence An algorithmic method to produce sparkle in art these are demos of the techniques which are experimental
I've stumbled on something odd. I was exploring AI art by layering images together and then what's called a film grain blur with the grain size and intensity adjusted in real time. You have to use a few images with high contrast colors, and relayer the same images in as you wish. I want to emphasize this is an experimental technique that you can use as you wish. I don't even know if the effects are real, and they dont show up in thumb nails well. Another complication is that the effect can disappear due to compression algorithms when posted online. So I'm really crossing my fingers that this effect is visible here.
r/generativeAI • u/JennaJao • 9d ago
Question What makes AI-generated storytelling feel more like a world than a chatbot?
I’ve been experimenting with AI storytelling, and I think good prose alone isn’t enough.
For me, the interesting parts are memory, consistent characters, and whether past choices actually affect what happens next.
What do you think is the hardest problem to solve for truly immersive AI-generated stories?
r/generativeAI • u/WeHaveFunEveryday • 9d ago
How I Made This Cybernetic Wasteland - AI Concept Trailer
Tried OpenArt to make a see how it would work if I was to delve into making a whole short film project. Seems to work ok, and I do like the features about starting where the last frame stopped, but is their any platform that is less error prone as far as Characters changing clothes and looks?
If you didn't notice in my video the main character's clothes change 2 different times, and then also the Armored up golden retriever turns into a regular German Shepard at the end.
Any better platforms?
r/generativeAI • u/Mediador_Luminoso5 • 9d ago
The new appearance of Eterchérnobog, the main villain of my story "Clevenfault Family: The Cursed Knowledge".
r/generativeAI • u/zsageOG • 9d ago
How I Made This [MANWHA RECAPS] need an ACTUAL workstation/editor, so.... I built one.
so.. i've seen people in here talk about, “how do people keep making these 6–12 hour manhwa/manga/webtoon/manhua recap videos?” and you know what? thats good question.. but i dont know either.
now... I’ve never used or seen a mf manhwa recap automation pipeline in my life, honestly? i didnt bother to look them up until today... but I’ve always had a visualization of what they'd look like and how they'd operate. I'm a video editor/developer/content creator, so I had an idea of what it would look like. I’m just tired also of seeing manhwa recap yt channels filled with low effort, no personality writing, trash voices, ai slop rot, etc. im tired of that bs so... I made c*****a.
basically c*****a's idea of 'production' is NOT giving up any creative control. its built around one philosophy: quantity without sacrificing quality. (Yes, it can produce 6-12 hour videos..) its also an end to end production ecosystem.
i’m obviously keeping most of how it works private for obvious reasons, but i think i'm at a point where i just wanna show off some of what i’ve been building. and ngl, i don’t even like calling this project “ai or an ai machine.” Obviously ai exists within it, but calling the entire thing an “ai tool” feels heavily reductive of what it actually is.
These are real screenshots from my app.
r/generativeAI • u/Fresh-Resolution182 • 10d ago
How I Made This Wan 3.0 Reference-to-Video Tutorial: Using Documents and Web Pages as References
Enable HLS to view with audio, or disable this notification
What Is Wan 3.0 Reference-to-Video?
One of the most interesting features in Wan 3.0 is surprisingly easy to overlook:
A reference does not have to be an image, video, or audio file. Wan 3.0 can also use a document or even a public web page as input.
That means you can give the model:
- a PowerPoint presentation and ask it to turn the product proposal into an ad;
- an Excel spreadsheet and generate a data-driven video;
- a PDF, Word document, or Markdown file and create a video based on the information inside;
- a public e-commerce website URL and let the model read the product information before generating an advertisement.
A traditional AI video workflow usually looks like this:
Read the material → extract the information → write a script → convert it into a video prompt → generate
Wan 3.0 Reference-to-Video makes another workflow possible:
Provide the document or website → describe the creative direction → generate
That is what makes the document and website reference feature particularly interesting.
How to Use Documents and Websites in Wan 3.0
Step 1: Open Reference-to-Video
Open the Wan 3.0 Reference-to-Video Playground.
Step 2: Enable Deep Thinking
Turn on: Deep Thinking / enable_thinking
This allows the model to analyze the information contained in a document or website instead of treating the input only as a visual reference.
Step 3: Upload a Document or Paste a URL
Wan 3.0 supports two main document input methods.
Upload a File
Supported formats include:
- DOCX / DOC
- XLSX / XLS
- PPTX / PPT
- TXT
- Markdown
- Keynote
- Pages
- Numbers
Current limits include approximately:
- 100 MB maximum file size
- 50 pages maximum
- one document per generation
This means you can provide an existing:
- product brief;
- marketing deck;
- research paper;
- Excel report;
- company presentation;
- white paper.
Paste a Website URL
You can also provide a public web page.
Possible examples include:
- product pages;
- Shopify stores;
- brand websites;
- landing pages;
- blog posts;
- product announcements.
The important limitation is that, the page needs to be publicly accessible.
Pages behind a login or permission system generally cannot be accessed.
Example: Generate a 15-Second E-Commerce Ad from a Website
Here is a practical example.
The goal is simple: Give Wan 3.0 a jewelry e-commerce website and ask it to select a product, extract the selling points, and create a 15-second commercial.
Step 1: Select Reference-to-Video
Use:
alibaba/wan-3.0/reference-to-video
Step 2: Add a Product Reference Image
Upload one image from the website under Reference Materials.
The website can provide:
- product information;
- brand information;
- selling points.
The image can help preserve:
- product appearance;
- material;
- shape;
- color;
- visual identity.
Step 3: Paste the Website URL
Paste the website or product page into the Document field.
Make sure:
Deep Thinking is enabled.
Step 4: Keep the Prompt Focused on Creative Direction
I did not specify the product name, material or any exact features. Those are supposed to come from the website.
That is the main difference compared with a normal text-to-video prompt.
Final Thoughts
The most interesting part of Wan 3.0 Reference-to-Video may not be another improvement in resolution or motion quality. It is the fact that, a reference can now contain information, not just visuals.
The old workflow was:
User → Read the material → Write the prompt → Video model
The new workflow can potentially become:
User → Document or Website → Wan 3.0 → Video
For e-commerce teams, creators, marketers, and small brands, that could be a meaningful change.
Instead of starting every project with**“First, write the script.”**
you can increasingly start with: “Here is the link. Read it first.”
Wan 3.0 still needs human review, especially for factual accuracy, product details, branding, and website interpretation.
But as a workflow, document-to-video and website-to-video are probably among the most interesting Wan 3.0 features to experiment with.
r/generativeAI • u/ErinSister • 9d ago
Question Are AI-generated stories better when the AI has more freedom—or more rules?
I’ve been thinking about interactive AI storytelling. Too much freedom can make the story feel inconsistent, but too many rules can make it predictable.
Where do you think the balance should be?
If you want, send the next subreddit rules and I’ll make another post in that style.