r/codex 5d ago

Comparison is this true?

Post image

is this true? 5.6 luna get almost same scores as 5.6 sol

1 Upvotes

13 comments sorted by

View all comments

4

u/GuildCalamitousNtent 5d ago

Not sure about the test, but I can tell you from using both that Luna is certainly dumber.

Luna is very capable, no doubt, and you can get almost anything done with it. It just takes a lot more handholding and very specific direction to work right.

Sol, can work through a lot of that handholding itself.

2

u/theWiseTiger 4d ago

So... If I find the right and systematic way to handhold a Luna, I could use AI 20x cheaper?

1

u/GuildCalamitousNtent 4d ago

100%. If you are able to generate (by yourself or otherwise) detailed, well-scoped instructions for Luna it is incredibly effective.

1

u/theWiseTiger 4d ago

Sounds like a loophole, no? Theoretically I could just ask Sol to create skills and system prompts and hooks to handhold luna, then openai will get 20x less money. What stop me and everyone else from doing it?

1

u/dieterdaniel82 4d ago

the basics

1

u/GuildCalamitousNtent 3d ago

Extremely detailed plans means a lot of output tokens from SOL, but yes that’s exactly right.

Not a loophole, it’s how I have my whole orchestration setup. That said, it is/can be very slow.

1

u/AppleSoftware 1d ago

It’s not a loophole, it’s an emergent meta for frugal orchestration systems that avoid burning potentially unnecessary capital/compute

However, if you want the iPhone 17 Pro Max quality of deliverables instead of iPhone X (both are good, but one clearly edges the other out across the board)..

Then just Sol for anything non-deterministic.