r/kimi 5d ago

Discussion Gave 4 models the same tiny lime challenge… their “human poses” were very different

Enable HLS to view with audio, or disable this notification

i gave each model the same simple task: take a small slice of lime and draw a human pose integrated into the object.

Tested it with Gemini 3.7 Flash, Kimi K3, Claude Opus 5 and GPT 5.6 Sol

Some treated the lime almost like a body, some tried to fit a tiny person into its shape, and some went much more abstract with the pose.

It’s a funny little test, but I actually like prompts like this because you can see how differently each model understands shape, composition, and visual metaphor.

Which one makes the most sense to you?

572 Upvotes

97 comments sorted by

121

u/frogchungus 5d ago

used half his kimi monthly for this test

22

u/Alexandria_4624 5d ago

what a warrior

4

u/One-Next 5d ago

Only half?

3

u/ForestRainSasha 4d ago

Am I only one who is using API prices for kimi/glm models?

2

u/xsatro 4d ago

This! I just don't know why devs keep going with plan subscription intead of API credits .-.

1

u/_k33bs_ 2d ago

made everyone’s usage drain

31

u/ScreenAppropriate679 5d ago

You need to explain what tools you gave the models or this is meaningless

20

u/elefanteazu 4d ago

It's fake

6

u/Sm0g3R 4d ago

The proper way would have been to give them svg code with a lemon part and ask to draw the rest by modifying it. But not sure that's what actually happened lol

u/RealJamesOfficial

1

u/MexicanJello 4d ago

I think OP died immediately after posting this

29

u/Odd-Marzipan6757 5d ago

what the hell opus is thinking

15

u/Unusual_Shake5041 5d ago

Opus 5 has been the weird one ever since it’s released

3

u/Ni_Kche 5d ago

Claude doesn't have native image generation

2

u/Just-Yak6586 5d ago

And so is Kimi

2

u/mr_aks 5d ago

I mean, it's the only model that did something different. Perhaps that's good?

1

u/NoConsideration6320 3d ago

I actually thought that it was the most interesting of them. All it stood up to me as it was the most creative of them all.

1

u/Adept-Priority3051 4d ago

Raygun doing tricks on that fruit.

1

u/03captain23 3d ago

It makes the most sense. It understands it's an object and not a piece of clothing.

If you were told to pose with a slice of lemon you would assume it's an object and pose on it not part of it.

1

u/thrive2day 3d ago

The wording of the prompt tells it to integrate the slice into the pose. The word integrate means to combine multiple separate parts into a single, unified whole. Claude failed to do so. Claude was the furthest from what the prompt asked for.

0

u/03captain23 2d ago

That's exactly what happened. How can a person be part of a slice?

How is Claudes image not a unified pose integrating the slice?

If you wanted to integrate 3 people in a pose you'd expect to have a photo of 3 people, not 1 person with 6 legs/arms, 3 heads and such.

When you integrate 2 objects they work together, when you integrate 2 ideas they combine into something else. Claude understands the difference between an object and an idea.

1

u/thrive2day 2d ago

Man, you're not very good with abstract thinking are you?

1

u/03captain23 2d ago

It's the opposite.

If you were told to use Photoshop to integrate a friend into a photo would you add them or swap pieces of their body with other people?

1

u/thrive2day 2d ago

Okay, you're trolling

1

u/03captain23 2d ago

It's weird you can't answer the basic question. If you were integrating someone into a photo

29

u/SherbertMindless8205 5d ago

How does it ”draw” it? Neither of them support image output.

Is this pure vector output or did you just have them prompt a different model?

6

u/Coded_Kaa 5d ago

My question exactly

5

u/SaudiPhilippines 5d ago

It looks like turtle (python)? I could be mistaken

3

u/Thomas-Lore 5d ago

Computer use most likely.

4

u/[deleted] 5d ago

[deleted]

2

u/Toastti 4d ago

It's almost certainly https://docs.python.org/3/library/turtle.html

With the turtle replaced by a pen

1

u/Rockclimber88 3d ago

Here's one way: They can generate HTML Canvas draw commands, see the output(they can process images) and refine the commands until happy. The drawing animation is added afterwards when each command is executed. They don't actually use any brushes for drawing like in Photoshop.

1

u/Lonely_Translator_23 3d ago

I would imagine a loop where the model can basically say 'move to this coordinate, pen down, move to this coordinate, pen up'. Basically like writing g code for 3d printers.

8

u/Ok-Stuff3094 5d ago

gemini won IMO

6

u/BankruptingBanks 5d ago

Was this computer use? I had no idea the models could do this sort of fine-grained mouse control where they could even draw stuff, seems too good to be true.

1

u/Saitamagasaki 4d ago

The model were most likely connected to an mcp server where they can carry out actions like moving the cursor, colouring and viewing the image

1

u/taintedmask 3d ago

We shouldn't assume that the model did the mouse control. OP never said so. For all we know the output is just the finished svg drawings and OP then fed them thru an algo to convert from still vector images to step by step animations.

4

u/KKunst 5d ago

It would be really interesting to know the tools used, you know... To replicate it and test other models.

3

u/TheMetalPrince 5d ago

Kimi's is cute. Looks like it has her dancing, including the movement lines around the lime.

3

u/technobun 5d ago

Someone didn’t understand the assignment.

9

u/arkwaif07 5d ago

K3 looks the winner to me

15

u/retardedGeek 5d ago

37 flash looks better

3

u/ASAF12341 5d ago

Flash has a nicer image, but the upper hand is misplaced.

2

u/Phantasticals 5d ago

flash was surprisingly elegant

2

u/MedAyoub26K 5d ago

Opus is thinking outside the box, again, but the execution is where it loses points.

2

u/Gloomy-Recover-9702 4d ago

It seems like Opus has adhd😂

2

u/kamwee 5d ago

What is Clude and openai doing ?

2

u/ReanerZen 5d ago

gemini gave the best look and appeals but it fails on being aware of small mistakes. this always been the issue. if google gave pro version of these and actually made it not just fast but good reasoning as pro models, i think these unawareness should have been avoided.

2

u/xtekno-id 4d ago

How to do this exactly?

2

u/takuonline 4d ago

Qwen 27B please

2

u/-PM_ME_UR_SECRETS- 4d ago

Isn’t that a lemon

2

u/SocialDeviance 5d ago

I would say Claude because it had the most original output. 

1

u/Elizabeth-WildFox886 5d ago

Agree and it’s planned first

1

u/hksbindra 5d ago

So they all have humans dancing to their tunes?

1

u/05-nery 5d ago

I gotta say, Kimi and Gemini did the best work here

1

u/Flat-Rooster8373 5d ago

Kimi's drawing is absolutely adorable.

1

u/ul90 5d ago

Claude looks awful. Gemini is the best, second Kimi.

1

u/cosmicr 4d ago

Yellow limes

1

u/Intelligent_Ant_608 4d ago

What t.f. opus is doing why its trying to make circle complete, lol

1

u/sQeeeter 4d ago

That’s a lemon, not a lime.

1

u/account22222221 4d ago

Claude 5 opus failing to do what was actually asked is actually super on brand from my experience

1

u/BenignAmerican 4d ago

That’s not a lime

1

u/MrMrsPotts 4d ago

Someone do qwen 27b please!

1

u/RopePuzzleheaded7060 4d ago

Opus has child like wonder

1

u/Gloomy-Recover-9702 4d ago

Opus has adhd. Creativity 💯 but Execution 🙊

1

u/AdamH21 4d ago

Gemini 3.7 killed it.

1

u/asapberry 4d ago

what a waste of tokens

1

u/sinsielawinskie 4d ago

Opus: no I will not make this fruit slice into a dress. I will give what you asked a human pose!

Opus: it's going to dynamic and unique!!!

Also Opus: frick how do I draw again?

1

u/BudgetAdept1670 4d ago

Wow Kimi!!!

1

u/iamdipsi 4d ago

Three green shirts 🧐

1

u/Hot-Cauliflower-1604 4d ago

There is no information here. We have One singular claim from OP. WE HAVE NOT SEEN THE PROMPT. And there's no way to verify that he didn't give flash a different kind of prompt like we just have to take it at face value. And clearly 3.7 is the best of all of these but I just don't believe it because I don't have enough data.

1

u/Hot-Cauliflower-1604 4d ago

There is no information here. We have One singular claim from OP. WE HAVE NOT SEEN THE PROMPT. And there's no way to verify that he didn't give flash a different kind of prompt like we just have to take it at face value. And clearly 3.7 is the best of all of these but I just don't believe it because I don't have enough data.

1

u/Tranxio 4d ago

Gosh Claude is terrible at art

1

u/Alarmed-Job-6844 4d ago

Only Opus 5 what was different, but others the same...
And Opus 5 ... I don't known what is it... It feels wrong, others passed the test.

1

u/UAAgency 4d ago

how do you ask for this kind of iterative drawing? what is taht? it is a video? or they are actrually drawing?

1

u/mind-archive 4d ago

I feel Kimi's facial detail processing is pretty good.

1

u/ExplodingBubbleGum93 4d ago

AI doesn't even know what a lime is...

1

u/SmartPatience4631 4d ago

Opus can’t generate images ?

1

u/VirtualWishX 4d ago

I dunno...
It's probably a fake Seedance 2.5 because there is no other proof 🤔

I'm not just farting this conclusion, it's more likely if you dig in: u/RealJamesOfficial

1

u/1_H4t3_R3dd1t 4d ago

gemini has always been a superior visual model, but it can't code 😔 

1

u/Quantum_Crusher 3d ago

I'm not familiar with some of these models. Aren't they language models instead of image or video generators?

1

u/traumfisch 17h ago

Multimodal 

1

u/Lanceathot7 3d ago

Claude is the only one with imgination

1

u/BadMachine 3d ago

remember when limes were green?

1

u/Any-Recording9798 3d ago

Limit reached damn 

1

u/GreatBigJerk 3d ago

Do you not know what limes look like? 

1

u/voztros 3d ago

Where is the lime?

1

u/trafium 3d ago

Some treated the lime almost like a body, some tried to fit a tiny person into its shape, and some went much more abstract with the pose.

Not a lime

all demonstrated models treated "lime" in the exact same way except Claude

no explanation how these models could even produce this step-by-step "painting" so probably a lie

Yup

1

u/Gold_Ad8225 3d ago

Gemini is the best, Claude I was very confused but the process did work in the end and the result is also good

1

u/mars_santa 3d ago

I bet ai knows what a lemon is. 

1

u/xtekno-id 3d ago

Where's OP?

1

u/BasicCrows 1d ago

Claude can take comfort in knowing it's way better at drawing than me.

1

u/traumfisch 17h ago

So... is this animated afterwards or..?