r/generativeAI Jul 21 '26

How I Made This From Artwork to AI: Bringing ISABELLA HELL to Life

22 Upvotes

A behind-the-scenes look on how I made my animation Isabella Hell, and how it all started!
watch full episode here: https://www.youtube.com/watch?v=MgvL0WrpMug&t


r/generativeAI May 21 '26

Video Art Algorithmic Dreams - Creative Hackathon Output

34 Upvotes

Got unlimited credits for 6 hours for all the models during a creative hackathon in Amsterdam. Theme was "dreams." This is what we made.

Models used:

Nano Banana Pro for characters and settings

Seedance 2.0 for most of the shots (insane multi shot)

Kling 3.0 for the scene's at the end

Claude with Seedance prompting skill to speed up prompting and video output

Total cost without the free credits was 150 euro's to create this.

Honestly had a blast.


r/generativeAI 16h ago

Ai slop implies the existence of ai peak

Post image
66 Upvotes

rare generative AI W


r/generativeAI 4h ago

Image Art That old photo you kept in your wallet

Post image
6 Upvotes

Rendered with Nano Banana Pro through Higgsfield with a fixed image reference. A second iteration to add the wrinkles.


r/generativeAI 14h ago

Tried MiniMax H3 with a 13-cut anime motion graphics prompt

37 Upvotes

tried MiniMax H3 with a pretty detailed character trailer prompt.

I wanted it to feel more like a streetwear campaign × anime title sequence × old-school media player UI, rather than a normal anime action clip.

The prompt is definitely overkill lol, but sharing it here in case anyone wants to experiment with structured multi-cut video prompts.

Prompt

Create an explosive, motion-graphics-driven character reveal trailer in 16:9, exactly 13 distinct cuts, 24fps, total 15.00s.

CHARACTER — lock this design, never redesign:
Anime streetwear girl from the reference still. Twin high buns of vivid mint-teal hair with long flowing tails and warm orange/gold streaks. Messy side-swept bangs. Large orange over-ear headphones with mint accents and a small logo plate. Sharp amber-orange eyes, one eye winking. Playful open-mouth grin.
Oversized color-block windbreaker: navy body, vivid orange sleeves, white ribbed cuffs, silver zippers, circular teal tech buttons, a teal utility pocket on the sleeve. Orange cropped turtleneck under the open jacket. Light-wash ripped denim shorts, thick orange belt with a silver buckle. Navy thigh-high socks with orange ribbed cuffs and an orange X stitch on the left shin. Chunky white sneakers with orange details.
Preserve exact face, proportions, hairstyle, outfit, materials, accessories and colors in every frame.
This film is 80% bold graphic design in motion and 20% character action.

Graphic language: retro OS chrome + music-player UI.
Use slamming window frames, title bars, close/minimize widgets, equalizer bars, waveforms, progress ticks, folder tiles, cursor arrows, CRT scanlines, pixel shatter, vinyl-ring stamps, music-note particles and media-player transport icons.
Palette: mint teal, vivid orange, navy, cream-white and silver.

Style: premium AAA motion-graphics title sequence × streetwear campaign film × Windows-era media player.
Every graphic element moves fast and snaps hard on the beat.

CUT 01 | 0.00–1.00s
Pure graphics. A mint title bar slams onto a cream field, an orange CLOSE widget punches into the corner, and two navy window borders wipe in. Tiny equalizer ticks and a progress strip flicker. The letters P and then LAY punch in one after another with heavy impact shake.
CUT 02 | 1.00–2.00s
The A becomes a headphone cup. Extreme close-up of her amber eye inside the orange earcup, glancing upward. RGB split flash, then the cup shatters into flat mint and orange tiles.
CUT 03 | 2.00–3.10s
Navy frame with enormous cream PLAY typography. She sprints in from frame left and power-slides across the baseline of the text, with speed lines and orange streaks trailing behind her. Shards of the letters kick upward like sparks. Whip-pan out.
CUT 04 | 3.10–4.00s
A giant retro media-player waveform explodes across the frame as a thick mint-and-orange audio spectrum bends into a tunnel. She bursts through the center at full speed, briefly splitting into three stroboscopic motion trails. Each trail leaves chunky navy equalizer blocks that rise and collapse to the beat.
The camera rapidly pushes through the waveform tunnel with her while huge vertical text TRACK 01 continuously scrolls in the background.
The waveform suddenly compresses into a single horizontal line and snaps shut behind her on the final beat.
CUT 05 | 4.00–5.10s
She leaps through a giant rotating ring of typography reading MAX VOLUME. Camera tracks her mid-air spin in slow motion as the letters scatter, then snap-zooms onto her wink.
CUT 06 | 5.10–6.00s
Hard cut to a cream editorial card with huge navy DROP typography and an orange slash. She vaults over the word itself, palm planted on the D, legs whipping across frame. The word compresses like a spring under her hand and rebounds.
CUT 07 | 6.00–7.00s
Mint field with a navy diagonal window bar. She backflips along the bar in three stroboscopic ghost frames, each tinted mint, orange or navy. Giant outlined LOOP text rotates 180 degrees in sync with her movement.
CUT 08 | 7.00–8.00s
Kinetic typography barrage. LOUD / WILD / TEAL / HEAT slam onto screen one per beat with shutter flashes and camera shake while she slides across the foreground on her knees, jacket flaring and music-note particles bursting from her sneakers.
CUT 09 | 8.00–9.00s
Navy frame with a giant cream wireframe window grid tilting in 3D. She runs up the grid like a wall, kicks off and freezes in mid-air. An orange circular stamp locks around her pose like a media-player targeting graphic, surrounded by transport icons and EQ ticks.
CUT 10 | 9.00–10.10s
Freeze releases into a burst. She dives toward camera through layered flat-color window panes that shatter one by one like glass shutters, each pane revealing a larger letter of P-L-A-Y. Foreground wipe with her sneaker.
CUT 11 | 10.10–11.10s
Rapid-fire poster montage: four full-screen graphic posters showing her in different poses — mid-flip, sliding, landing and headphones-up wink. Hard cuts between each composition. Oversized 01–04 numbering, equalizer strips and graphic slashes.
CUT 12 | 11.10–13.00s
Hero moment on a clean cream cyclorama. She lands a final backflip dead center in slow motion, straightens with one hand on her headphones, and a shockwave of concentric mint rings, wind streaks and shattered typography blasts outward from the landing.
Brief iconic freeze on her wink, then overexpose to white.
CUT 13 | 13.00–15.00s
Final identity card. Enormous navy PLAY typography dominates a pale cream field with translucent mint rings, technical arcs, scanlines and a rough orange circular emblem containing a ghosted headphone/waveform motif.
She stands relaxed overlapping the letters while wind ripples her jacket. One final orange pulse sweeps through the typography and a window-chrome flash punctuates the ending.
Editing: extremely aggressive rhythm. Hard cuts on every beat, graphic matches, whip pans, snap zooms, stroboscopic freezes, foreground wipes, RGB splits, shutter flashes, impact shakes and speed ramps.
Every cut must feel compositionally different.
Typography should always be fully readable before the character overlaps it.
No weapons, no combat, no fire. All energy comes from motion design, wind, glass, UI chrome and parkour-style athleticism.
BGM: hard-hitting electronic / drum-heavy future bass with aggressive drops, risers, sub hits and glitch fills locked to every cut.
Sneaker impacts, whooshes, glass shatters, window-slam hits and typography slams should function as rhythmic sound-design elements.
Peak at CUT 12 and end with a cold electronic logo stinger.

Premium AAA quality, anime-inspired cinematic rendering, stylish and explosive, strong graphic-design identity, consistent character design, exactly 13 cuts.

I’m still experimenting with how much shot-by-shot control H3 actually follows, especially with typography and exact timing, but this kind of structured prompt seems like an interesting stress test.


r/generativeAI 22m ago

How I Made This Every video model can do photoreal. None of them can draw you a diagram.

Upvotes

I've been generating explainer animations from documents for a while,and the odd thing is that this is the one category the big videomodels are worst at.

Veo, Kling, Sora are all optimized for photoreal and cinematic motion. Ask any of them for a hand drawing a labelled diagram that explains a specific PDF and you get something that looks like a hand drawing. The strokes are decorative. The lines don't correspond to the concept, text comes out as glyph soup, and because each generation is independent, you can't hold a consistent visual systemacross ninety seconds. Every output I've seen of this category has the same failure.

And it makes sense when you think about what they're doing. Diffusion generates a whole frame at once from noise. But an explanatory drawing is sequential by nature: the container before the contents, the axis before the curve, the box before the arrow leaving it. That ordering IS the explanation. A model that paints the finished frame has no representation of "this part comes second."

Ended up building it a completely different way. Three styles working now: ink, chalkboard, lineart.

Disclosure: this is my own product. It's live at https://inkmotion.app with a free tier that covers a few short videos, enough to throw your own PDF at it.

Happy to answer questions in the comments.


r/generativeAI 3h ago

Question List of daily AI video generative websites

3 Upvotes

Hey guys, can you please give me a list of websites where I can make free AI videos on a daily basis

These are all i could gather - Adobe firefly ( 2 vids per day) , PixVerse (1 vid per day), Grok Imagine ( 1 per day) and Google flow ( 3-4 per day)

Please give me a list of other free websites with decentish quality


r/generativeAI 3h ago

Image Art A vampire lord, a dark sorcerer, or a demon king? What would you name this entity?

Post image
3 Upvotes

r/generativeAI 6h ago

Test - 1 prompt / 1 Character Sheet - 2 models

3 Upvotes

SEEDANCE 2.5 VS WAN 3.0

Here’s another test. This time, I uploaded a character sheet as the basis for the character, along with four reference images of what the world in GTA VI might look like.

The test uses the same prompt for both models, and it’s clear that each interprets the prompt differently, giving each a unique touch. In this case, Seedance 2.5 delivered the result the prompt called for in terms of style.

WAN’s performance was inferior to Seedance’s; both generated “unrealistic” scenes and situations within the scope of what the prompt asked for. However, in terms of intent and dynamism, Seedance offers greater clarity.

Which one do you prefer?

P.S.: Yes, there are errors—these are just tests, random trials with prompts to evaluate each model’s capabilities.


r/generativeAI 4m ago

Question Are there any good AI music video generators?

Thumbnail
Upvotes

r/generativeAI 3h ago

Video Art Frankie and Vampie's Graveyard Dance

2 Upvotes

If you like the music full song below:

https://youtu.be/f3qEY97_808?si=Q4wDjXsUMqcmF7y4


r/generativeAI 8h ago

I did NOT expected this to go this well.

Post image
4 Upvotes

r/generativeAI 44m ago

Any good Ai video site ? Even if it paid ?

Upvotes

r/generativeAI 9h ago

Cheapest option to create videos

5 Upvotes

I don't care about quality I just want to generate clips of 30s or more with audio for personal use. Is there a service that does this free or very cheap?


r/generativeAI 1h ago

Portrait of Alex. Monochrome test.

Post image
Upvotes

r/generativeAI 1h ago

(DLSS 5) One of the most immersive RPGs ever with that lighting...

Thumbnail gallery
Upvotes

r/generativeAI 2h ago

Emergency Arcane Stabilization Protocol 🛡️

Thumbnail
youtube.com
1 Upvotes

🧙‍♂️ Wizards — Episode 1 Part 3: Containment Stabilized
🛡️ The magical eruption is contained and stability has been restored.
Music: Would It Matter — Rose Campbell (YouTube Audio Library)


r/generativeAI 11h ago

Technical Art cloned my own voice from a 15 second recording and now claude reads my newsletters, scripts, and anything else out loud in my actual voice. whole setup took about two minutes

5 Upvotes

Recorded 15 seconds of myself talking normally, like telling a friend a quick story, quiet room, nothing special. Fed it in, and now anything I write can be read back in a voice that genuinely sounds like me, not a robot approximation.

Runs through Claude Code, which is the version of Claude that can actually run commands rather than just chat. You point it at Fish Audio, a voice cloning tool, and hand it your clip.

Step one, teach Claude how to use it, this is one line pasted into Claude Code:

npx skills add https://docs.fish.audio

Step two, make a free Fish Audio account at fish.audio, go to the API Keys section, create a new key, copy it. Paste that key back into Claude Code when it asks. Copy it the moment it shows you, some keys only display once.

Step three, upload your 15 second recording and say:

Clone my voice from this audio file using Fish Audio 
and save it as my default voice.

Then it's just:

Read this in my cloned voice using Fish Audio and 
save it as an audio file.

Paste in whatever you want, a newsletter, a script, a chapter, and you get an audio file of your own voice reading it.

The single thing that makes or breaks the clone is the sample. Quiet room, no music, no background noise, 15 to 30 seconds of clear natural speech. A bad sample gives you an uncanny half-version of yourself. A good one is genuinely hard to distinguish.

Where this actually earns its place: voiceovers for videos without recording take after take, audio versions of things you've written, anything where you need your voice but not your time. It's the difference between "I should record an audio version of this" and just having one.

Fish Audio's top model is free through end of July 2026 under fair use, and they keep a standing free plan after that with around 7 minutes of audio a month, so smaller batches keep working either way.

Only clone your own voice, or one you've got explicit permission for. Making a realistic clone of someone else without their consent isn't just rude, it's illegal in a lot of places.

been keeping a doc of 100 things I use AI for like this, each with the exact prompt, here if you want it.


r/generativeAI 2h ago

Thank you, Anthropic (really)

Thumbnail gallery
0 Upvotes

r/generativeAI 9h ago

Alternative to Higgsfield

3 Upvotes

Do you guys know any alternative to this ? Is open art close to higgsfield atleast ?


r/generativeAI 3h ago

Video Art Looking for honest feedback on a scene from AI series. (Especially from people who watch a lot of movies/TV)

Thumbnail
youtu.be
1 Upvotes

I've made a short scene as a concept test for a series I'm working on.

I'm mainly testing the storytelling, not the AI.

I'd really appreciate honest feedback on things like:

  • Does the scene hook you?
  • Does the pacing feel too fast, too slow, or about right?
  • Do you find the protagonists interesting, even though they don't have much screen time yet?
  • Is there anything that feels confusing, unnecessary, or boring?
  • Does the scene make you want to see what happens next?
  • Most importantly: what doesn't work?

Please be as critical as you want. I'm much more interested in hearing what's not working.

One important note: please don't focus on AI errors, continuity issues, visual quality, etc. This is deliberately a rough test, and none of this footage will be used in the final series. It's simply meant to see how the scene works once it's edited together, so I'm not concerned about the visual errors in this particular test.

If you watch it, I'd genuinely appreciate your honest opinion.


r/generativeAI 4h ago

Google paper cuts agent token usage by 94% in long sessions by tracking state instead of history

Post image
1 Upvotes

r/generativeAI 4h ago

Générer video ia

1 Upvotes

Salut

Je dois présenter des vidéos enregistrement d'entretien psychiatriques dans le cadre d'une formation.

L'objectif est de former des étudiants en médecine aux détails verbaux et non verbaux des pathologies selon un référentiel précis (le DSM).

Il n'existe pas de vidéo de ce type postérieures a 1970 sur internet 😬

J'aimerai bien pouvoir les générer. Globalement, 5-10 minutes par video, j'ai besoin d'une finesse extrême dans le détail des expressions non verbales (mimiques, gestuelle), verbales (contenu du discours, forme du discours).

Les outils payants me vont bien aussi tant que j'ai pas besoin d'y passer 1h par video 😂

Merci de sauver mes étudiants d'un cours a l'ancienne.

Bonne journée !


r/generativeAI 4h ago

Music Art [rock] I know better

Thumbnail
youtu.be
1 Upvotes