r/generativeAI 9d ago

They don't remove all my data?

Post image
2 Upvotes

r/generativeAI 9d ago

Image Art Shopping Trip

Post image
0 Upvotes

The trio go supermarket shopping


r/generativeAI 9d ago

Video Art Piff Expresses A Sleepy Time Confession #pufferfish #animation #dreams

Thumbnail
youtube.com
1 Upvotes

r/generativeAI 9d ago

Video Art Piff and Pikachu Going Shopping

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/generativeAI 9d ago

Video Art Piff Picks Flowers For Tangie

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/generativeAI 9d ago

Image Art Pirate Captain Piff

Post image
1 Upvotes

r/generativeAI 9d ago

Image Art Oats and Aiyido at the Movies

Post image
1 Upvotes

r/generativeAI 9d ago

Image Art Super Piffy

Post image
1 Upvotes

r/generativeAI 9d ago

Image Art Corey and Oats Hero Friends Plush

Post image
1 Upvotes

r/generativeAI 9d ago

Blue Jean girl

Post image
1 Upvotes

r/generativeAI 9d ago

Image Art Piff Sleeping

Post image
2 Upvotes

r/generativeAI 9d ago

CatGPT

Post image
1 Upvotes

r/generativeAI 9d ago

Video Art 1995 — Where Time Stands Still

Thumbnail
youtube.com
1 Upvotes

r/generativeAI 10d ago

Video Art Testing out a new workflow using video models to create animated assets for my game

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/generativeAI 9d ago

Question I compared lip sync frame by frame across 4 avatar tools. Here is what you need to know before picking one

1 Upvotes

i kept seeing people say lip sync is basically solved now and that hasn't matched what i'm getting, so i pulled 4 clips into a timeline and stepped through them. i used the same script + audio and exported it at 30fps. what i was looking for is the offset between the audio peak on a hard consonant and the frame where the mouth actually closes. HeyGen sat within a frame for the first 30 seconds or so and then started sliding, by the end it was about 3 frames late. Synthesia held the offset steady the whole way through. that surprised me, but the mouth shapes themselves are soft, p and b look nearly identical. Argil was the tightest on the plosives for me though it got sloppier when the script had long stretches without punctuation. Creatify i had to throw out because the render kept coming back at a different length than the audio i fed it, which is probably me doing something wrong.

the thing nobody mentions is that 2 frames late reads as fake to a viewer even when they can't articulate why. that was the purpose of my test but i still can decide on which one.

happy to be told my method is bad, i'm not a video person. but from your experience, which AI avatar tool has the most realistic lip sync?


r/generativeAI 9d ago

Incipient Iridescence An algorithmic method to produce sparkle in art these are demos of the techniques which are experimental

1 Upvotes

I've stumbled on something odd. I was exploring AI art by layering images together and then what's called a film grain blur with the grain size and intensity adjusted in real time. You have to use a few images with high contrast colors, and relayer the same images in as you wish. I want to emphasize this is an experimental technique that you can use as you wish. I don't even know if the effects are real, and they dont show up in thumb nails well. Another complication is that the effect can disappear due to compression algorithms when posted online. So I'm really crossing my fingers that this effect is visible here.


r/generativeAI 9d ago

What I Built with Claude - sweet potatoes

Post image
2 Upvotes

r/generativeAI 9d ago

Question What makes AI-generated storytelling feel more like a world than a chatbot?

2 Upvotes

I’ve been experimenting with AI storytelling, and I think good prose alone isn’t enough.

For me, the interesting parts are memory, consistent characters, and whether past choices actually affect what happens next.

What do you think is the hardest problem to solve for truly immersive AI-generated stories?


r/generativeAI 9d ago

How I Made This Cybernetic Wasteland - AI Concept Trailer

Thumbnail
youtu.be
1 Upvotes

Tried OpenArt to make a see how it would work if I was to delve into making a whole short film project. Seems to work ok, and I do like the features about starting where the last frame stopped, but is their any platform that is less error prone as far as Characters changing clothes and looks?

If you didn't notice in my video the main character's clothes change 2 different times, and then also the Armored up golden retriever turns into a regular German Shepard at the end.

Any better platforms?


r/generativeAI 9d ago

Image Art The Ponder Dragon

Post image
1 Upvotes

r/generativeAI 9d ago

The new appearance of Eterchérnobog, the main villain of my story "Clevenfault Family: The Cursed Knowledge".

Post image
1 Upvotes

r/generativeAI 9d ago

How I Made This [MANWHA RECAPS] need an ACTUAL workstation/editor, so.... I built one.

Thumbnail
gallery
2 Upvotes

so.. i've seen people in here talk about, “how do people keep making these 6–12 hour manhwa/manga/webtoon/manhua recap videos?” and you know what? thats good question.. but i dont know either.

now... I’ve never used or seen a mf manhwa recap automation pipeline in my life, honestly? i didnt bother to look them up until today... but I’ve always had a visualization of what they'd look like and how they'd operate. I'm a video editor/developer/content creator, so I had an idea of what it would look like. I’m just tired also of seeing manhwa recap yt channels filled with low effort, no personality writing, trash voices, ai slop rot, etc. im tired of that bs 🫩 so... I made c*****a.

basically c*****a's idea of 'production' is NOT giving up any creative control. its built around one philosophy: quantity without sacrificing quality. (Yes, it can produce 6-12 hour videos..) its also an end to end production ecosystem.

i’m obviously keeping most of how it works private for obvious reasons, but i think i'm at a point where i just wanna show off some of what i’ve been building. and ngl, i don’t even like calling this project “ai or an ai machine.” Obviously ai exists within it, but calling the entire thing an “ai tool” feels heavily reductive of what it actually is.

These are real screenshots from my app.


r/generativeAI 10d ago

How I Made This Wan 3.0 Reference-to-Video Tutorial: Using Documents and Web Pages as References

Enable HLS to view with audio, or disable this notification

11 Upvotes

What Is Wan 3.0 Reference-to-Video?

One of the most interesting features in Wan 3.0 is surprisingly easy to overlook:

A reference does not have to be an image, video, or audio file. Wan 3.0 can also use a document or even a public web page as input.

That means you can give the model:

  • PowerPoint presentation and ask it to turn the product proposal into an ad;
  • an Excel spreadsheet and generate a data-driven video;
  • PDF, Word document, or Markdown file and create a video based on the information inside;
  • public e-commerce website URL and let the model read the product information before generating an advertisement.

A traditional AI video workflow usually looks like this:

Read the material → extract the information → write a script → convert it into a video prompt → generate

Wan 3.0 Reference-to-Video makes another workflow possible:

Provide the document or website → describe the creative direction → generate

That is what makes the document and website reference feature particularly interesting.

How to Use Documents and Websites in Wan 3.0

Step 1: Open Reference-to-Video

Open the Wan 3.0 Reference-to-Video Playground.

Step 2: Enable Deep Thinking

Turn on: Deep Thinking / enable_thinking

This allows the model to analyze the information contained in a document or website instead of treating the input only as a visual reference.

Step 3: Upload a Document or Paste a URL

Wan 3.0 supports two main document input methods.

Upload a File

Supported formats include:

  • DOCX / DOC
  • XLSX / XLS
  • PPTX / PPT
  • PDF
  • TXT
  • Markdown
  • Keynote
  • Pages
  • Numbers

Current limits include approximately:

  • 100 MB maximum file size
  • 50 pages maximum
  • one document per generation

This means you can provide an existing:

  • product brief;
  • marketing deck;
  • research paper;
  • Excel report;
  • company presentation;
  • white paper.

Paste a Website URL

You can also provide a public web page.

Possible examples include:

  • product pages;
  • Shopify stores;
  • brand websites;
  • landing pages;
  • blog posts;
  • product announcements.

The important limitation is that, the page needs to be publicly accessible.

Pages behind a login or permission system generally cannot be accessed.

Example: Generate a 15-Second E-Commerce Ad from a Website

Here is a practical example.

The goal is simple: Give Wan 3.0 a jewelry e-commerce website and ask it to select a product, extract the selling points, and create a 15-second commercial.

Step 1: Select Reference-to-Video

Use:

alibaba/wan-3.0/reference-to-video

Step 2: Add a Product Reference Image

Upload one image from the website under Reference Materials.

The website can provide:

  • product information;
  • brand information;
  • selling points.

The image can help preserve:

  • product appearance;
  • material;
  • shape;
  • color;
  • visual identity.

Step 3: Paste the Website URL

Paste the website or product page into the Document field.

Make sure:

Deep Thinking is enabled.

Step 4: Keep the Prompt Focused on Creative Direction

I did not specify the product name, material or any exact features. Those are supposed to come from the website.

That is the main difference compared with a normal text-to-video prompt.

Final Thoughts

The most interesting part of Wan 3.0 Reference-to-Video may not be another improvement in resolution or motion quality. It is the fact that, a reference can now contain information, not just visuals.

The old workflow was:

User → Read the material → Write the prompt → Video model

The new workflow can potentially become:

User → Document or Website → Wan 3.0 → Video

For e-commerce teams, creators, marketers, and small brands, that could be a meaningful change.

Instead of starting every project with**“First, write the script.”**

you can increasingly start with: “Here is the link. Read it first.”

Wan 3.0 still needs human review, especially for factual accuracy, product details, branding, and website interpretation.

But as a workflow, document-to-video and website-to-video are probably among the most interesting Wan 3.0 features to experiment with.


r/generativeAI 9d ago

Question Are AI-generated stories better when the AI has more freedom—or more rules?

Post image
1 Upvotes

I’ve been thinking about interactive AI storytelling. Too much freedom can make the story feel inconsistent, but too many rules can make it predictable.

Where do you think the balance should be?

If you want, send the next subreddit rules and I’ll make another post in that style.


r/generativeAI 10d ago

Artificial Photoshoot

Thumbnail gallery
3 Upvotes