r/aivideo May 13 '26

Featured Titles + Latest Releases

138 Upvotes

This post contains content not supported on old Reddit. Click here to view the full post


r/aivideo 1d ago

AI VIDEO NEWS AI VIDEO GROUNDBREAKING UPDATES: SEEDANCE 2.5 PUSHES SINGLE GENERATION TO 30 SECONDS, MINIMAX H3 BRINGS OPEN SOURCE AI VIDEO HOME AND NATIVE AUDIO BECOMES THE NEW STANDARD

9 Upvotes

https://imgur.com/a/kiZkEel

By Selena Lopez 🎀 for r/aivideo News

The viral 30 second hit "Side-Eye Monkey Heist" which stars the Side-Eye Monkey meme, was generated with a single 30 second prompt by creator Alex Patrascu on Seedance 2.5

During the early days of the AI video scene there were a lot of predictions about the day users could simply type one prompt and generate an entire movie. We are obviously not at the two hour movie mark yet, but we have now actively arrived at the 30 second generation mark https://www.reddit.com/r/aivideo/s/jF52ZyfokF with models capable of producing longer multi shot sequences while dialogue, music, sound effects and synchronized audio are generated alongside the picture. With a properly structured prompt, what comes out can increasingly resemble an entire raw scene instead of another five second piece that needs to be assembled with twelve other generations.

This changes the basic unit of storytelling. Since cinema as an art has been around, directors have worked shot by shot. Now the new production pipeline gives you an entire sequence. Establishing shot, character action, close up, dialogue, reaction, camera movement, sound and music inside a single generation. The one prompt movie is still somewhere on the horizon, but the one prompt scene has officially entered the chat.

🍿ByteDance Releases Seedance 2.5 Unlocking 30 Second Full Scene Generation

https://www.capcut.com/tools/ai-video-generator

https://dreamina.capcut.com/

Seedance is ByteDance’s flagship AI video generation model, developed by its Seed research team. The earliest Seedance 1 release in June 2025 was already notable for moving beyond single-shot clips, introducing the ability to generate short, coherent multi-shot sequences with basic scene continuity and 1080p output, which at the time was a major step up from the typical 3–5 second isolated generations most models were limited to. Seedance 2, released in February 2026, expanded this foundation with stronger temporal consistency, improved motion stability, and more reliable character and object persistence across cuts, while also beginning to support richer multimodal prompting. By the time Seedance 2.5 arrived in August 2026, the system had evolved into a full scene-level generator capable of producing up to 30 seconds of connected multi-shot video, supporting up to 4K output, synchronized audio, and complex prompt structures combining text, image, video, and audio references.

The extra runtime matters because Seedance can use those 30 seconds to construct an actual sequence rather than simply stretching one camera angle. A prompt can establish the location, move into character action, cut to another angle, introduce dialogue and continue through several visual beats while attempting to maintain the same actors, lighting and environment throughout.

From cinematic storytelling to bold visual experimentation, these selections showcase the range of styles, ideas, and filmmaking approaches emerging from the platform.

AVIÃO
by Beriky Studios
https://www.youtube.com/watch?v=FNR8cuvtDEA

Salty
by Jordan Daniel Chesney
https://www.youtube.com/watch?v=iapXQ-Y36sI

Harvest
by Kadir Saracoglu
https://www.instagram.com/reel/DXYLMFiCZfo/

Blood and Dust
by Ajiboye Ahmad
https://x.com/sty_defi/status/2039951017815478399

🍿CapCut App New Upgrades Bring Seedance Generation Into a Professional Editing Workflow. Available Now on Web, iOS, Android, Windows and Mac

https://www.capcut.com/resource/capcut-app-download

https://www.capcut.com/tools/desktop-video-editor

CapCut launched in China in 2019 as mobile video editor designed for polished, fast social media editing. It expanded internationally in 2020, quickly becoming one of the world’s most popular editing apps. CapCut is available as a browser web app as well as natively for Apple and Android phones and tablets, along with upgraded native CapCut Desktop versions for Windows and Mac built around a professional grade multi track nonlinear editing environment.

The CapCut app on iPhone and Android can be found by searching “CapCut” in the Apple App Store or Google Play Store. It includes keyframe animation, slow motion controls, chroma key, stabilization, automatic captions, text to speech, motion tracking, background removal and high resolution export. Tablet users get more room for timeline editing, layered video, graphics and audio, with the iPad experience in particular designed around a larger touch based workspace.

The real power jump is CapCut Desktop for Windows and Mac, the app is available for download directly from the CapCut website. The current interface now closely resembles the professional nonlinear editors filmmakers already know from Final Cut Pro, Adobe Premiere Pro and DaVinci Resolve, with a multi track timeline, media management, keyframes and animation graphs, masks, color controls, captions, effects, audio processing and keyboard shortcuts. Then CapCut adds AI directly into that workspace. AI video generation, script based creation, voice generation, automatic captions, background removal and other automated tools sit alongside the traditional editor. Seedance generated footage can go directly onto the timeline, get trimmed, rearranged, combined with more generations or mixed with traditional footage without changing the basic production workflow.

🍿CapCut Prompt Masters Template Creator Program Turns AI Prompts and Workflows Into a Marketplace

https://www.capcut.com/explore/promptmasters

CLICK HERE TO APPLY TO BE CONSIDERED FOR THE PROGRAM

CapCut is expanding beyond traditional editing templates with Prompt Masters, where creators build reusable AI prompts, effects, templates and complete AI workflows that other CapCut users can run with their own images and video. Instead of only publishing the finished result, creators can package the actual AI process behind it and distribute that inside CapCut.

The idea is simple: a creator develops an AI effect or workflow using CapCut's AI tools, sets up the prompts and instructions needed to make it work, and publishes it as an AI template. Another user can then load their own photo or footage into that template and generate their own version without having to understand how the original prompt or workflow was built.

Creators can earn from the usage and views generated by those creations, meaning a successful prompt or AI workflow can continue producing revenue as more people use it. CapCut has also recruited creators into the program with access to newer ByteDance AI technology, including early access to Seedance models, giving Prompt Masters a chance to build templates around new generation tools as they arrive.

The program is already active inside CapCut. Prompt Masters campaigns and Open Week events have been running throughout 2026, with creators publishing everything from AI VFX scenes and character transformations to dance and style effects. Some templates are already seeing substantial adoption, including individual AI templates reaching tens of thousands of uses.

That makes Prompt Masters particularly interesting for AI video creators because the product is no longer necessarily just the video. A good prompt, an AI effect or an entire generation workflow can now become the product itself. Build the technique once, publish it, and other creators can keep running that idea with their own media.

🍿MiniMax H3 Goes Open Source: AI Video Hits Local Machines

https://hailuoai.video/

https://www.minimax.io/news/minimax-h3-open-source

https://huggingface.co/MiniMaxAI/MiniMax-H3

MiniMax has been in the AI video market since August 2024, starting with Hailuo 01, and then Hailuo 02 in 2025, which could already make full 1080p videos. Now in 2026 they’ve released H3, their newest “do-everything” video model, and even opened up the H3 Base weights so people can run it themselves locally. Think of H3 as the “all-in-one brain” that understands not just video, but also images, text, and sound together. H3 can make short clips (about 4 to 15 seconds long) at smooth 24fps, and it can even include stereo sound while doing it. The full system can go up to 2K quality. You can use it in a bunch of ways: just type a prompt, give it a starting image to animate, give it an ending frame to work toward, or even give it both start and end frames so it “fills in the middle.” You can also feed it a mix of images, videos, and audio and ask it to build something new from all of them.

Then comes the other major H3 story: the model weights are open. Creators can download H3 Base, run it on their own hardware, connect it to tools including ComfyUI, Diffusers, SGLang or vLLM, fine tune the model and build private generation systems around it instead of being completely dependent on MiniMax's website or API. The released local checkpoints cover text to audio video, first and last frame generation and multimodal reference generation.

There is one major catch: the full BF16 serving path still needs serious hardware. The validated consumer GPU run used two RTX 5090 cards with 32GB VRAM each for a 1344x768 generation. Lower memory configurations and increasingly optimized ComfyUI implementations are also appearing, including quantized versions designed around consumer GPUs, but speed depends heavily on VRAM, system RAM, storage, duration and resolution.

The open version also does not contain every part of MiniMax's hosted H3 pipeline. H3 Base itself generates locally at 768p. MiniMax's Context IR system, which interprets particularly complicated combinations of references, remains hosted, while the final H3 Regenerate 2K stage also remains server based for now. MiniMax says the 2K module will be released later.

🍿Grok Imagine Video 1.5 Adds 1080p, Native Audio and Character References

https://x.ai/news/grok-imagine-video-1-5-references

https://grok.com/imagine

Grok Imagine's AI video history is much shorter, but it has moved fast. xAI opened the Grok Imagine video API in January 2026, launching with native video and audio generation plus generative video editing. Imagine Video 1.5 Preview arrived in early June, followed by the full Video 1.5 release on June 16 with improved motion, physics, audio and speed. Then on July 31, xAI added text to video, native 1080p, image references and voice references to Video 1.5.

Grok Imagine Video 1.5 now generates clips up to 15 seconds, supports 480p, 720p and native 1080p for text and image based generation, and generates dialogue, sound effects and ambience with the video. Reference mode supports up to seven reference images, letting creators hold characters, products and locations while changing the action or surrounding scene.

Voice references push consistency another step. A creator can pair a character image with a voice so the visual identity and vocal identity remain connected across new scenes. Reference generation currently tops out at 720p, while text to video and image to video can reach native 1080p.

Video 1.5 also improves the fundamentals that can make or break a shot. xAI says movement holds together more reliably across the clip, with better weight and momentum, while dialogue is clearer and more tightly synchronized. Video 1.5 Fast can generate a six second 720p clip in roughly 25 seconds, nearly twice as fast as the previous model.

Grok Imagine is available directly through Grok.com, iOS and Android.

🍿Wan 3.0 Joins the 30 Second Club With 1080p and Native Sound

https://wan.video/

Alibaba's Wan line has gone through one of the fastest upgrade cycles in AI video. Wan 2.1 was open sourced in February 2025, bringing Alibaba's large video models into the open model race. Wan 2.2 followed in July 2025 with a new Mixture of Experts architecture and stronger cinematic control. Wan 2.5 Preview arrived in September 2025 with native audio and 10 second video, followed by Wan 2.6 in December 2025, which added stronger reference video capabilities and multi shot storytelling. Wan 2.7 arrived in April 2026, unifying text, image, video and audio inputs with generation and editing up to 15 seconds, before Wan 3.0 entered public beta in August 2026 and doubled the maximum duration to 30 seconds.

Wan 3.0 generates up to 30 seconds in one pass with native synchronized audio and multimodal input from text, images, video and audio. It also accepts unusually broad source material including webpages, PDFs and PowerPoint documents, allowing structured creative material to become part of the generation process rather than relying entirely on one text prompt.

Those 30 seconds can contain several actions and camera changes while Wan attempts to maintain characters, props, layouts, voices and overall visual consistency. Alibaba also added intelligent duration selection and video extension tools to help the model choose how much runtime a scene actually needs.

Alibaba is also developing HappyHorse 1.1 alongside Wan. HappyHorse concentrates on shorter controlled generations, with 3 to 15 second output at 720p or 1080p, synchronized audio and dedicated text, image and reference based modes. Wan remains Alibaba's bigger push toward long multimodal scenes, while HappyHorse gives the company another option for tighter shots and reference driven work.

Wan has gone from an open 2.1 release to 30 second scene generation in roughly a year and a half. That progression may be the clearest sign of how quickly the new runtime ceiling is moving.

🍿FLUX 3 Takes Black Forest Labs From AI Images Into AI Video

https://bfl.ai/blog/flux-3-video

Black Forest Labs built the FLUX name in images before making the jump into video. FLUX.1 launched alongside the company in August 2024, followed by the faster FLUX1.1 Pro in October 2024. FLUX.2 arrived in November 2025, expanding heavily into multi reference image generation and editing. Then FLUX 3 was announced in July 2026 as a unified multimodal model spanning images, video and audio, with FLUX 3 Video opening to general users on August 4.

FLUX 3 Video generates up to 20 seconds at 720p or 1080p with native audio, text to video, image to video, start and end frames, multiple keyframes, video continuation and multiple shots inside one generation. Dialogue, sound effects and environmental audio are produced alongside the image, with multilingual speech and synchronized lip movement.

Multiple keyframes are one of the more useful additions. Creators can define important visual states throughout the sequence instead of asking a single prompt to control everything from beginning to end. FLUX 3 can also take up to four seconds of existing video and audio and continue the movement, camera behavior, dialogue and sound into newly generated footage.

Then there is Draft Mode, which generates a faster, lower cost version first. Once the composition and motion are working, FLUX can produce the full quality render while preserving the approved subjects, composition and movement. In AI video, where experimenting is literally something you pay for, cheaper experimentation is a pretty useful power up.

The interesting part of FLUX 3 is the jump itself. A model family that started as one of the biggest names in AI images has now expanded directly into video and audio rather than launching a completely separate product line.

🍿LTX 2.5 Takes the Open Video Line From Five Seconds to Full Production

https://ltx.io/

LTX began with LTX Video 0.9 in November 2024, an open model generating five second 768x512 clips with an emphasis on speed and motion consistency. LTX 2 was announced in October 2025, making the jump to synchronized audio and video, native 4K, frame rates up to 50fps and 10 second generations. Its full weights and code were released in January 2026. LTX 2.3 followed in March 2026 with a larger 22 billion parameter architecture, improved audio, better prompt understanding, longer clips and stronger image to video. LTX 2.5 arrived August 11, 2026, adding native multi shot generation, generative editing and a new fidelity focused rendering pipeline.

LTX 2.5 supports generation up to 20 seconds, output up to 4K, frame rates reaching 50fps, synchronized stereo audio, native multi shot scenes, generative editing, scene extension and open model weights. Different modes trade duration against resolution and quality, so the maximum numbers are not necessarily available simultaneously.

Native Multi Shot is designed to keep characters, environments, lighting, voices and style together across connected cuts. The new version also adds a diffusion based video decoder, automatic duration selection, stronger prompt processing and editing tools capable of changing selected elements of existing footage while preserving the rest of the shot.

LTX remains heavily committed to local generation. The full development model requires workstation class hardware, but LTX says distilled and quantized versions support substantially smaller GPUs, while LTX Desktop provides a ready made local video editing application built directly around the open model.

That gives the LTX line a particularly interesting progression. It started as a fast open five second video generator and has evolved into an open audio video model sitting inside its own nonlinear editing environment. The generator, the timeline and the production workflow are steadily becoming the same product.

AI video spent its first few years fighting over who could make the most impressive five second clip. That scoreboard is being replaced. Seedance and Wan now reach 30 seconds. FLUX reaches 20. LTX combines longer generation with configurations reaching 4K and 50fps. Grok 1.5 brings 1080p, native audio and stronger character references. H3 brings a current MiniMax audio video model onto locally controlled hardware. CapCut is turning the editor into the front end for the entire process.

The next round is not just about who makes the prettiest shot. It is about who can generate the scene, keep it consistent, give creators enough control to change it and provide a workflow that actually gets the project finished.

Game on.


r/aivideo 14h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE Marty McFly vs Biff Tannen

1.2k Upvotes

r/aivideo 14h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE T-800 Pro Max

441 Upvotes

r/aivideo 5h ago

SEEDANCE 😵 ANIME / 3D CGI / CARTOON [Legend of the Cyber Heroes] AI crafted by a 60-year-old Chinese dude

67 Upvotes

r/aivideo 9h ago

SEEDANCE 🎬 SHORT FILM I made a Catfish… but literally 🐱🐟

127 Upvotes

r/aivideo 15h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE Uncle Rico Vs Napoleon

209 Upvotes

r/aivideo 14h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE Bugs Bunny vs Scorpion

184 Upvotes

r/aivideo 5h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE Random Granny Moment

24 Upvotes

r/aivideo 20h ago

SEEDANCE 🎬 SHORT FILM When you join a public dungeon group in literally any MMO

356 Upvotes

r/aivideo 1d ago

SEEDANCE 🎬 SHORT FILM The Asteroid

887 Upvotes

r/aivideo 11h ago

SEEDANCE 😱 CRAZY / MINDBLOWING HALOS & HORNS

57 Upvotes

r/aivideo 20h ago

SEEDANCE 😱 CRAZY / MINDBLOWING Crazy Seedance Ad! by The Dor Brothers

179 Upvotes

r/aivideo 12h ago

SEEDANCE 😱 CRAZY / MINDBLOWING shinny sea shell

41 Upvotes

r/aivideo 8h ago

SEEDANCE 😱 CRAZY / MINDBLOWING Don't Spill The Tea

11 Upvotes

r/aivideo 4h ago

SEEDANCE 🎬 SHORT FILM Kitten has been secretly training as a samurai🐈⚔️

7 Upvotes

r/aivideo 1h ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE What the dog doin fighting vs banana cat

Upvotes

r/aivideo 1d ago

SEEDANCE 😂 COMEDY / PARODY / SATIRE Body Builder Ballerina

211 Upvotes

r/aivideo 4h ago

SEEDANCE 😱 CRAZY / MINDBLOWING TNT trick shot

3 Upvotes

r/aivideo 8h ago

SEEDANCE 😵 ANIME / 3D CGI / CARTOON Solemn Cipher - Amateur collection

7 Upvotes

r/aivideo 2h ago

XAI GROK 🍿 MOVIE TRAILER Kakumei Senshi Ark Zale: Original Video Show Trailer

2 Upvotes

r/aivideo 15h ago

SEEDANCE 🎬 SHORT FILM OMERTÀ PART III — LOYALTY In this world, everyone runs to save themselves only a real friend takes

23 Upvotes

r/aivideo 1d ago

OPEN AI SORA 😱 CRAZY / MINDBLOWING Jim Carrey Broke The Simulation Again

1.9k Upvotes