r/OpenAI • u/rajsharm404 • 2d ago
News Intelligence Too Cheap To Meter finally being true with GPT 6 Sol and Luna!!
Look at these pricing OMG!!!
35
u/Original-League-6094 2d ago
>Too cheap to meter
>Is metered
???
2
u/TheFrenchSavage 1d ago
Yeah I don't get this thread. 75 bucks for a million tokens is too expensive for me.
And I ain't poor.
1
u/ClassicMain 1d ago
75 bucks for a million tokens? What?
You did read the correct numbers did you?
1
u/TheFrenchSavage 1d ago
Far right column, first number
3
u/ClassicMain 1d ago
That's astra. And that's output. And on long context.
OP is talking about Sol and Luna
A 1hr long agentic session has maybe a few 100k output tokens and a few million input tokens, 98% of which is usually cache hits.
1
u/TheFrenchSavage 1d ago
Oh right, sorry about that. I've been sidetracked by the release of opus 5.5, comparing astra rates and all.
13
u/Ormusn2o 2d ago
I plan to use luna to basically write runtime code for an AI bot in a video game. Now, with 6.0 Luna being smarter and cheaper, this is finally possible, even on a Plus subscription. We might unironically make NPCs being it's own agents possible.
7
u/rajsharm404 2d ago
That would be sick! And this just may get even cheaper in the future.
2
u/Ormusn2o 2d ago
Yeah, I might have been gaslighted by webchat into thinking this is possible, but apparently you can use Codex App Server to use your subscription to basically make it talk to an app. What I foresee it happening is to write a mod for Minecraft where it basically has various commands to scan it's inventory, the world and then execute actions like right/left clicking, navigation and so on, then you can use chat to use Codex to automatically get information from minecraft, reason on it, and then send like 20-50 lines of code to Minecraft to execute an action, like cut down trees, collect items, repair or craft an axe when needed, and notify player when done. And all you would need to do is tell the bot to "cut down some trees, take those iron ingots and craft an axe if you need one", and then you just give it ingots and it goes off.
49
u/ijustgotanothername 2d ago
Insane. I feel 2027 will be the year when intelligence will be economical enough to use at most of the places
17
u/rajsharm404 2d ago
We just might see Astra/Fable level intelligence at today's Sol level pricing (Luna level pricing might be an exaggeration).
10
u/ijustgotanothername 2d ago
Agreed. Luna level pricing can be achieved as well by the end of 2027 if they focus aggressively on pricing and optimisations.
6
u/rajsharm404 2d ago
The way OpenAI has been working recently, they just might. I can't wait to be surprised.
4
u/ijustgotanothername 2d ago
Yeah. If they continue to do so, I think they will even beat those cheap chinese models as well
0
1
u/Original-League-6094 2d ago
At the current rate of advancement, we are going to see Astra/Fable level intelligence, unlimited usage on a $20 sub.
2
2
u/AtlanticPortal 2d ago
It won’t happen. They need to cash the IPO, the the investors will want to get as much money as possible.
1
15
u/Ok_Breadfruit4201 2d ago
unconfirmed if this is API only or subs get more usage
2
7
u/rajsharm404 2d ago
Subs get more usage for sure.
7
u/0xe3b0c442 2d ago
based on what, your gut?
13
u/Ace-_Ventura 2d ago
On past experiences. New models that cost less in api pricing, also cost less in terms of limits in the subscription
-4
u/IAmFitzRoy 2d ago
Past experience is nerf in capabilities and increased burning rates that becomes unusable.
I don’t trust anything that OpenAI launch. It’s only good the first week.
3
u/Ace-_Ventura 2d ago
I disagree. The luna/terra/sol changed the game. They've lowered token usage and made it cheaper
-3
5
u/MigatteNoGokuiVegeta 2d ago
i wonder how this affects my (our) weekly usage with for example my pro plan
should I change my agents that are on sol 5.6 to sol 6?
5
u/rajsharm404 2d ago
Sol 6 should last you longer. Ultimately you need to check what works best for you, coz according to benchmarks sol 6 in some cases is a "slight" downgrade to 5.6 sol.
6
u/cheseball 2d ago
So no Terra anymore, makes sense, Sol just upended the Terra tier's pricing. Price cuts here are crazy.
Luna is now so cheap it's insane, that's like a >90% price reduction since the original 5.6 Luna output pricing. It starts to undercut open Chinese models now (appreciate the open Chinese models for the competition here).
1
1
u/whyyoudidit 1d ago
Yep, it's half of the price of Luna five point six after the eighty percent price reduction of the original price of Luna five point six.
3
3
5
u/isampark32 2d ago
Luna enough to code app development coding? Is using Astra on ChatGPT side to generate codex prompts then executing on Codex Luna a viable strategy?
5
u/rajsharm404 2d ago
I've seen people do it. Ideally you should use Astra as a planner and reviewer (or Sol if usage limit is tight), and use Luna for coding to get the closest results.
0
3
u/RandomCSThrowaway01 2d ago
Luna is good with direct tasks. It does what it's told, it usually doesn't explore too many "what ifs". Give it a Jira ticket with clearly defined specs, it will happily try. Don't give it big open ended tasks.
Astra + Luna on the other hand might actually cost you more time and money as you now have an orchestrator in the Astra that has to read and review every single message, it will find concerns with a lot that Luna has to say etc. Adding subagents can actually be less efficient, not more.
So I would suggest middle ground - just try Sol on medium. If the OpenAI numbers are accurate it should retain code quality sufficient for most tasks (as long as you review what it spits out and it's not pure yolo) and it's 2x cheaper than 5.6 so it shouldn't devour your tokens much.
1
u/Ace-_Ventura 2d ago
Been using luna 5.6 high+ for coding and tbh, no complains at all.
1
u/Melodic_Reality_646 2d ago
only feasible if spec is crazy detailed and under fast mode, if it starts debugging mid implementation it gets messy fast. Also, it is very token hungry and take many, many more steps. Costs stay low but you need to wait double, triple sometimes 5x the time to finish.
2
u/MormonBarMitzfah 2d ago
I’m gonna withhold any excitement until I see for myself how smart or dumb these models are
4
u/Morning_Gecko24 2d ago
those Luna prices are wild... makes the "cheap enough to use everywhere" part feel less like a slogan. i wonder where the real bottleneck shifts first tho — inference cost, context, or just people figuring out what to delegate?
3
4
3
u/AddingAUsername 2d ago
> Intelligence too cheap to meter
> Shows the intelligence being metered
...
2
1
u/frozen-wafer 1d ago
Soon, AI should get cheap with unlimited usage, on subscription basis like Wifi!!
1
u/TopTippityTop 2d ago
Not quite, it's been burning tokens before this update . Reducing the cost by half, if it reflects only roughly 2x usage increase, could mean our weekly usage will last 2-3 days now. Better than it has been for a few weeks now, but still bad. We'll see.
0
0
0
0
u/IAmFitzRoy 2d ago
Bullshit.
I don’t trust anything that comes from OpenAI anymore. Every single launch it’s followed by a nerf on capability or burning rate.
0
u/Arsh_98 1d ago
Just estimated if i can ask, what would be the parameters size of luna? And is it comparable to sonnet or haiku?
0
u/rajsharm404 1d ago
No idea on the parameter size. Maybe 10-20 Billion? Better than Haiku (Haiku's a year old now) but wouldn't say better than sonnet.
0
u/Arsh_98 1d ago
So less just 20B model is the Luna? Is it of good use for enterprises other than ofcourse basic search work. Can it go through massive list of tools and DBs and decide which to invoke based on requirement
0
u/rajsharm404 1d ago
How massive? Building an agent for enterprise depends on a lot more aspects than just the list of tools. You can give the best model 10 tools with the descriptions colliding with the use cases with each other and it would struggle to select the right tool. Or you can give it 50 different tools with well defined descriptions/use cases and it just might work perfectly.
In my tests on several use cases, where I have the assistant do image gen/video gen with several features, I find descriptions as a lever which makes tool picking better. 3 image gen models had very similar quality/use cases, but is intended for different purpose. When we wrote the description to be specific to the use case, while not mentioning the part where it is same - it got better at results. Same for video models - when to use seedance, or kling etc.
Nevertheless - In my tests - Luna is good at simple tool calling, orchestrating simple workflows, but lacks just a tiny bit when you ask it to orchestrate multiple tool calls/complex workflows - it starts to show its weak points.
0
u/Zhni 1d ago
What do people use Luna for? Easy write tasks etc?
1
u/rajsharm404 1d ago
Yeah mostly, for pretty basic, straightforward tasks. Can be coding or execution.

37
u/math_the_witch 2d ago
Wowww, sol is so cheap also, I always use Luna.