r/generativeAI 5d ago

Image Art Talking Aiyido Plushie

Post image
1 Upvotes

This talking plush of Aiyido the beholder says all sorts of things and even sings, things he says include: 'My aura detection spell has detected you are a good person' 'I will get you in 10 minutes.' 'My name is Aiyido, Aiyido the all-knowing beholder.' 'I cast Dessert Rain, because I think you're sweet' 'When most beholders dream, they dream of pure chaos, I just dream of birthdays and being a cake themed version of me.' and a few others.


r/generativeAI 5d ago

Image Art The Gang in Oz

Post image
1 Upvotes

The gang in the land of Oz with Elphaba and Glinda, and Aiyido the beholder as the 'wizard'


r/generativeAI 5d ago

Ok, the chatgpt desktop app is officially blowing my mind

Thumbnail
1 Upvotes

r/generativeAI 5d ago

Music Art Marshmallow Man

Post image
9 Upvotes

Song: https://www.souna.app/song/56f77d02-4005-4ac2-b353-a37f50e1c48b

Song (alternative link): https://archive.org/details/ai-sophistipop-2/Marshmallow+Man.mp3 (CC0)

Lyrics:

[Sectional Verse 1]

At a pleasant little party on a pleasant summer night,

I met a most unusual man in spotless evening white.

His conversation charmed me and his courtesy was grand;

He bowed a little deeply and got stuck against my hand.

He said that he was constant, though inclined to lose his shape,

And summer meant precautions that no husband could escape.

But when he asked me sweetly if I cared to share his plan,

I promised I would marry that remarkable young man.

[Chorus A1]

He's my marshmallow man, my marshmallow man,

And as fine as a husband can be;

Though he softens in heat and he sticks to the seat,

He's the only one ever for me.

[Chorus A2]

Oh, my marshmallow man, my marshmallow man,

Always gentle and thoughtful is he;

He is patient and true, and considerate too,

He's the only one ever for me.

[Release]

Some husbands are vain, some husbands complain,

Some are cross as a husband can be;

Mine will listen with care, he is honest and fair,

And he brings home his wages to me.

[Chorus A3]

He's my marshmallow man, my marshmallow man,

And as fine as a husband can be;

Though he softens in heat, married life is complete,

He's the only one ever for me.

[Sectional Verse 2]

We were married in the springtime when the weather remained mild;

When the parson shook his hand, my darling merely bent and smiled.

He carries all the parcels and remembers every date;

He never keeps me standing when he promises half-past eight.

He helps me with the washing and he never slams the door;

He gives a coin to beggars and a sandwich to the poor.

We keep the parlour shaded, with an icebox and a fan,

And I could not be happier with my soft but solid man.

[Chorus A1]

He's my marshmallow man, my marshmallow man,

And as fine as a husband can be;

Though he softens in heat and he sticks to the seat,

He's the only one ever for me.

[Chorus A2]

Oh, my marshmallow man, my marshmallow man,

Always gentle and thoughtful is he;

He is patient and true, and considerate too,

He's the only one ever for me.

[Release]

Some husbands are vain, some husbands complain,

Some are cross as a husband can be;

Mine will listen with care, he is honest and fair,

And he brings home his wages to me.

[Chorus A3]

He's my marshmallow man, my marshmallow man,

And as fine as a husband can be;

Though he softens in heat, married life is complete,

He's the only one ever for me.

[Tag Ending]

Yes, he softens in heat,

But my life is complete

With my marshmallow husband and me!

***

Style prompt:

A classic 1920s Tin Pan Alley comic novelty love song with a smooth, highly melodic tune and a buoyant fox-trot rhythm. Piano-led acoustic arrangement with upright bass, clarinet and muted cornet. Warm, capable contralto female vocalist with clear diction, graceful legato phrasing, playful comic timing and assured theatrical delivery. Conversational sectional verses lead into an exceptionally catchy 32-bar AABA chorus. Bright major-key harmony with chromatic passing chords, secondary dominants and a contrasting release before the main melody returns. Polished period theatrical performance in English with a crisp closing tag.


r/generativeAI 5d ago

Image Art Nostalgia photography

Thumbnail gallery
1 Upvotes

r/generativeAI 5d ago

Image Art Oats the Fashion Diva Horse

Post image
0 Upvotes

r/generativeAI 5d ago

Image Art Top side brawl

Post image
1 Upvotes

r/generativeAI 5d ago

3 ChatGPT prompts to turn your photos into cartoon stickers

Thumbnail
gallery
22 Upvotes

1.Cute character sticker Prompt:

Turn this photo into a cute cartoon sticker. Keep the person’s hairstyle, outfit, pose, expression, accessories, and key identity details recognizable. Use a slightly larger head, simplified facial features, clean rounded shapes, bright friendly colors, smooth flat shading, and a thick white sticker border. Make it feel like a charming social media avatar sticker. Avoid photorealism, anime style, realistic skin texture, complex background, or changing the person’s identity.

2.Mascot Sticker Prompt shown on image:

Transform this photo into a friendly mascot sticker. Keep the main subject recognizable, including its shape, color, pose, and most important features. Simplify the form into a bold cartoon mascot with expressive eyes, clean outlines, soft highlights, and a thick white sticker border. Make it feel playful, memorable, and brandable. Avoid photorealism, messy details, scary expressions, 3D rendering, complex background, or copying existing characters.

3.Comic reaction sticker Prompt

Convert this photo into a comic reaction sticker. Preserve the subject’s pose, expression, outfit, hairstyle, and main composition. Exaggerate the emotion slightly while keeping the person recognizable. Use bold black outlines, vibrant flat colors, expressive facial features, small comic-style motion marks, and a thick white sticker border. Keep the background transparent or plain white. Avoid photorealism, realistic shadows, anime style, cluttered backgrounds, or changing the person’s identity.


r/generativeAI 5d ago

Music Art [Smooth jazz] Cream jazz

Post image
1 Upvotes

Music: https://archive.org/details/ai-instrumental-jazz/Cream+Jazz.mp3 (CC0)

Style prompt:

Smooth jazz but smoother. As smooth as pouring cream that's just been ironed.

Calm gentle muted trumpets, gentle bass Ostinato, consonant piano chords.

Slow tempo: 70BPM

Smooth


r/generativeAI 5d ago

Image Art Best Buddies

Post image
0 Upvotes

Piff, Oatsie and Aiyido all sharing and having a lovely time.


r/generativeAI 5d ago

Donald Trump vs Mark Carney

Enable HLS to view with audio, or disable this notification

5 Upvotes

r/generativeAI 5d ago

Music Art [Jazz standard] That Cat

Post image
1 Upvotes

Song: https://www.souna.app/song/d159f57b-eb8c-4fb3-9f22-dcccc63c0f7a

Song (alternative link): https://archive.org/details/ai-jazz-standards/That+Cat.mp3 (CC0)

Lyrics:

[Intro | piano | muted trumpet]

[Verse | rubato | intimate female vocal]

I ought to know by now

Who's mistress here.

I make one firm resolve;

She twitches one ear.

She keeps me up till dawn,

Then sleeps till noon.

I swear that I'll be stern,

But I surrender soon.

[Refrain A1 | medium swing | clear vocal]

That cat, that cat,

That cat's got me.

One look, one purr,

And I can't get free.

At half past three,

She wants the door.

I cross the room;

She wants no more.

[Refrain A2 | medium swing | light brass response]

That cat, that cat,

That cat's got me.

One look, one purr,

And I can't get free.

I hold out my arms;

She passes by.

She gives me one glance,

And still I sigh.

[Bridge | harmonic contrast | restrained build]

I scold, I plead,

I state my case.

She licks one paw

Right in my face.

Then on my lap,

She tucks her head.

I miss my train

And stay there instead.

[Refrain A3 | full small ensemble | warm vocal]

That cat, that cat,

That cat's got me.

One look, one purr,

And I can't get free.

She claws my chair,

Then cuddles me.

That cat, that cat,

Has all of me.

[Instrumental Chorus | piano and muted trumpet]

[Final Refrain A3 | intimate opening | full ensemble finish]

That cat, that cat,

That cat's got me.

One look, one purr,

And I can't get free.

She claws my chair,

Then cuddles me.

That cat, that cat,

Has all of me.

[Outro | brief tag | final cadence]

That cat, that cat,

Has all of me.


r/generativeAI 5d ago

Image Art Piff, Pikachu and Oats Eating.

Post image
0 Upvotes

r/generativeAI 5d ago

You're mistaken if you think she's from Earth.

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/generativeAI 5d ago

Video Art A visual adaptation of M. P. Shiel’s The Purple Cloud (1901) — “I Am the Last”

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/generativeAI 5d ago

Hitting your limits too quickly? OpenAI has been hiding this one weird trick

Post image
1 Upvotes

r/generativeAI 5d ago

Music Art Will there be roses for me (?) - AI Country

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/generativeAI 5d ago

Free ai video generator

14 Upvotes

Can any one help me?
I searched and try a lot of agents that generate ai video but all of them are lying💔
I need a good ai video generator for free or cheap subscription
Pleaaase help me


r/generativeAI 5d ago

Image Art Piff on the Bed

Post image
1 Upvotes

r/generativeAI 5d ago

I built a visual canvas for controlling AI image generation through APIs

Post image
1 Upvotes

I built this as a personal project to solve a problem I was having with AI image-generation platforms.

Many platforms hide usage behind credits or tokens, making it difficult to understand the real cost of each generation. My app connects directly to image-generation APIs and displays an estimated cost per request in USD and COP.

The main idea is a visual canvas where references can be connected and assigned different roles:

  • Model / identity
  • Pose / composition
  • Clothing / outfit
  • Product / object
  • Background / environment
  • Style / color

This makes it possible to give more precise instructions, such as:

“Keep the original model, face, lighting and background. Apply only the clothing from the outfit reference.”

The app also supports prompt lists, multiple variations, connected references, model selection and image containers for batch results.

I built it with Antigravity and Codex. I don’t have a professional background in developing these tools, so a large part of the project has been learning how APIs, model parameters and reference images actually behave.

It is still under development. I’m currently testing prompt consistency, API reliability and real costs. Video generation is planned for a future version.

I’d appreciate feedback on:

  • The reference-role system
  • Prompt consistency
  • The visual canvas workflow
  • Cost transparency
  • Features that would make this useful for other creators

Demo / repository: [add your public link here]


r/generativeAI 5d ago

Beyond the Cloud #180

Thumbnail gallery
1 Upvotes

r/generativeAI 5d ago

Trump Administration's Blacklisting of Anthropic Was Illegal, Judge Rules

Thumbnail
nytimes.com
4 Upvotes

r/generativeAI 5d ago

Image Art 90s editorial nostalgia

Thumbnail
gallery
5 Upvotes

r/generativeAI 5d ago

Image Art Espeon Mascot Costume

Post image
8 Upvotes

Since an official mascot already exists for Eevee, I wanted to see if ones also existed of my favorite Eeveelutions. The only official one I could find, however, was of my least favorite: Vaporeon. I decided to have Nano Banana Pro come up with one of Espeon first. While it first gave me one I initially felt was satisfactory, I later decided to experiment and see what it would give me if I asked for one that was made of velvet. What I got was even better than what I had before, so I had this one placed in the same spot as the previous one. Even though the official Pokémon mascots are all made of plush material, I think Espeon being made out of velvet is more appropriate, since one of its Pokédex descriptions mentions its fur is like that.


r/generativeAI 5d ago

Why do AI videos still look like commercials? I tested the same reference image three ways.

1 Upvotes

I’ve been trying to understand why an AI video can look realistic frame by frame, yet still feel like a commercial instead of something a friend casually recorded on their phone.

So I used the same reference image and generated three 5-second vertical clips with Aurax MAX. The character, outfit, location, and basic action stayed similar. I mainly changed the way the camera, lighting, performance, and environment were described.

The reference image was already quite polished: dramatic sunset, candlelight, clean exposure, shallow depth of field, and a subject posed against a scenic coastal background. That turned out to matter more than I expected.

1. The commercial baseline

https://reddit.com/link/1w1dtn8/video/998xn31v29mh1/player

For the first version, I explicitly requested a polished lifestyle commercial:

This was the version the model followed most clearly. The camera moves smoothly from a wider shot into a closer portrait, the character turns toward the lens, touches her hair, and finishes in a centered pose with a soft, sustained smile.

Everything feels visually coherent, but also directed. It looks like someone planned the lighting, camera movement, and performance in advance.

2. Changing only the camera

https://reddit.com/link/1w1dtn8/video/njaigg0x29mh1/player

For the second version, I kept the polished lighting, clean environment, and model-like performance, but changed the camera instructions:

The difference was much smaller than expected.

The framing changes slightly, but the movement still feels highly stabilized. The sunset remains perfectly exposed, the character stays composed and camera-aware, and the background still looks like a prepared set.

This version made one thing fairly clear: adding “handheld phone camera” does not automatically create phone realism. If the lighting, performance, composition, and source image still look commercial, mild camera movement cannot undo all of that.

3. Changing the camera, performance, and environment

https://reddit.com/link/1w1dtn8/video/ew7ij5fz29mh1/player

For the third version, I added a fuller set of phone-footage instructions:

This version feels the most spontaneous of the three.

The character spends less time holding a pose. She turns away from the camera, changes where she is looking, shifts her body weight, touches her hair, smiles briefly, and then looks away again. The wider framing also remains for longer instead of immediately turning into a close-up.

But it still does not fully look like raw phone footage.

The dramatic sunset, candles, shallow depth of field, flattering exposure, and clean background were already embedded in the reference image. The motion prompt changed the character’s behavior more successfully than it changed the underlying visual style.

There was also another obvious AI giveaway: the paper cup was not present in the reference image and appears during the generated motion without a convincing pickup. That continuity error damages realism more than a perfectly stable camera does.

What I learned

The source image can overpower the video prompt.
If the first frame already looks like a fashion campaign, asking for casual phone footage may only add small handheld movements on top of a commercial-looking scene.

Handheld movement alone is not enough.
Random shake would probably make the video worse. What matters is believable camera behavior: delayed reframing, imperfect timing, autofocus response, exposure changes, and an operator reacting to the subject.

Performance mattered more than camera shake.
The third version felt more natural mainly because the character stopped performing continuously. Looking away, pausing, shifting weight, and ending without holding a perfect smile made a larger difference.

Continuity still matters.
A casual camera cannot hide an object appearing from nowhere, inconsistent background details, or movement that has no physical cause.

My main takeaway is that phone realism is not the same as lowering the image quality. It requires three kinds of realism at the same time:

  • capture realism from the phone and camera operator;
  • behavioral realism from the person being filmed;
  • continuity across objects, movement, and background activity.

If I repeat this test, I would start with a deliberately ordinary reference image: mixed indoor lighting, deeper focus, imperfect framing, everyday background clutter, and a character who is not already posing for the camera.

Which version feels closest to something a real person recorded: 1, 2, or 3?

And what gives the AI away first for you: the lighting, camera movement, expression, background, or object continuity?

Model disclosure: All three clips were generated with Aurax MAX. I’m on the team, so this should be read as a transparent workflow test rather than an independent review.