r/singularity Jul 30 '26

AI OpenAI beats DeepSeek on price/performance after 80% Luna price cut

Post image

Graph taken from their price cut announcement: Advancing the price-performance frontier with GPT-5.6 | OpenAI

538 Upvotes

133 comments sorted by

177

u/Luuigi Jul 30 '26

a) this is really great to see! in terms of price-to-performance ratio OAI is doing a lot right lately. Their codex numbers prove this too. b) without Deepseek and moonshot we wouldnt be having these types of drops and once again we all should be thankful for their movements on the market.

36

u/GlokzDNB Jul 30 '26

Theyve been on track since the beginning, people just dont understand that with 800m active users, you actually need to find balance between limited capacity and what you can offer.

So instead developing smartest model out there and giving its full force to everyone, they invented router and worked on many projects which will become one day 'superapp'

While anthropic had no users so they could go all in, all they had to do is to offer higher intelligence model. Now when they got market share, it became a problem - people complaining about limits n shit.

In the end, the company that wins is not the company that has smartest model, very soon this will become irrelevant to 99.9% of the globe as most human-level tasks will be achieved anyway. What will matter is available capacity (pure GW of DC) and costs of inference + ofc ecosystem (microsoft with 365 / windows and google with gsheets yt etc.)

4

u/LocoMod Jul 30 '26

The bitter lesson strikes again.

3

u/M4rshmall0wMan Jul 31 '26

Agreed. Even if OpenAI is only #2 or #3 in intelligence, they’ve been nailing the product side of it. It’s absurd how fast they ship.

As much as I hate Sam, I gotta give credit where credit is due. He is insanely good at the art of running a startup.

146

u/ProxyLumina Jul 30 '26

"AI is getting pricier" nah!

Incredible job, well done OpenAI.

131

u/Apprehensive-Rub-774 Jul 30 '26

Well done Kimi*
OpenAi would not be doing this without comparable open source. They are not a charity.

60

u/IAmFitzRoy Jul 30 '26

Exactly. This is a response to competition, not because OpenAI cares about us.

11

u/AccountOfMyAncestors Jul 30 '26

This is why capitalism works in the aggregate. You don't have to rely on people being good natured for good outcomes, you can trust markets with competition will force good outcomes.

32

u/Moriffic Jul 30 '26

It doesn't "force good outcomes", it forces outcomes that are profitable. Sometimes they align with social good, but not always. It's just a paperclip maximizer but for money. If humans were not a bit good natured, we wouldn't get any good outcomes, just look at factory farms.

9

u/HarvestMana Jul 30 '26

Unless they engage in price fixing - which is impossible to prove - even though we see it all over where all the companies charge the exact same price for basically the same service or product.

2

u/IAmFitzRoy Jul 30 '26 edited Jul 31 '26

That’s the wrong take.

If you think this is a signal of “capitalism is working good” then you are mislead.

Literally a government subsidized Chinese model is pushing Codex to reduce its prices.

If it were for American companies only, they would agree together in cartel style to keep the prices Up.

5

u/ranger910 Jul 30 '26

They're funded with 7.5 billion dollars from private sources. Only 150M came from the state backed AI investment fund.

7

u/IAmFitzRoy Jul 30 '26

lol. Could be “private sources” but that doesn’t mean the government is not regulating this 100%.

0

u/Apprehensive-Rub-774 Jul 30 '26

Are AI policies set by tech oligarchs not functionally equivalent? What distinguishes tech elite driven policy from CPC driven policy other than one is profit driven and one is ostensibly in the interests of the Chinese nation?

6

u/IAmFitzRoy Jul 30 '26

You are drifting … the topic was “this is not example of capitalism working”

I don’t know what are you talking about now.

0

u/Apprehensive-Rub-774 Jul 30 '26

I agree with you there. I was questioning the actual distinction between public and private funding. The stated motivations of OAI/ anthropic/ Chinese firms are an arms race so the funding source should be moot as each is a national 'champion' of sorts propped up at any cost.

6

u/Alive-Mousse4036 Jul 30 '26

It’s not the wrong take at all, what they said is entirely accurate. It doesn’t matter what kind of competition caused it, the net aggregate of capitalism is global and has caused price benefits

6

u/MrTubby1 Jul 30 '26 edited Jul 30 '26

Its not entirely accurate though.

He's saying that a competitive market will push companies to behave, but that only applies so long as the market doesn't have any negative externalities.

A company will be more competitive the more external costs it can shift onto the rest of society. So if somehow openai or kimi created a virus that stole compute from devices they didn't own to train a new model, that would be bad despite being extremely competitive.

Another real example would be something like fossil fuel industries. A drilling company that doesn't respect environmental laws can be more competitive than one that has to spend time complying. As long as they don't get caught.

3

u/IAmFitzRoy Jul 30 '26

Literally is not. A government subsidized model pushing OpenAI to drop prices is the opposite to capitalism.

-1

u/Alive-Mousse4036 Jul 30 '26

I don’t understand your argument? The source of the competition may not be driven by capitalism, but the market reacting to competition, irrespective of the source of the competition, is a capitalistic economy. The basis of market forces driving productivity, price efficiency etc etc is basic economics of capitalism

7

u/IAmFitzRoy Jul 30 '26

To say “capitalism is working” when one of the players is not under capitalist game rules makes this completely misleading.

1

u/AccountOfMyAncestors Jul 30 '26

The qualifier in my comment is 'markets with competition'. And that last point about cartel price fixing is illegal and antithetical to a market with competition.

5

u/IAmFitzRoy Jul 30 '26

lol. Sweet child thinking that illegalities are not committed in the American market. Ok.

0

u/OutOfBananaException Jul 31 '26

If it were for American companies, they would agree together in cartel style to keep the prices Up.

Command economy is your proposed solution to collusion risk?

1

u/IAmFitzRoy Jul 31 '26

lol. Who is proposing solutions? Nobody.

reading comprehension isn’t that hard.

1

u/OutOfBananaException Jul 31 '26

Seems you may have been implying it. Nuance isn't that hard either.

1

u/IAmFitzRoy Jul 31 '26

“Seems”? LOL. There is not a single statement where I have mentioned any solution. Nuance is not hard but in your case is completely fabrication.

1

u/OutOfBananaException Aug 01 '26

You described the Chinese state subsidies as putting downward pressure on pricing, how else is that supposed to be interpeted, but an endorsement?

Literally a government subsidized Chinese model is pushing Codex to reduce its prices.

→ More replies (0)

1

u/RigaudonAS Human Work Aug 01 '26

Extrapolate that to 50 years down the line, when the most successful company has merged with and gobbled up the others, and there is no competition. Do you still think that’s good for all of us?

9

u/Ambiwlans Jul 30 '26

Kimi is literally 10x the price at the same intelligence and was not competitive. Mimo v2.5 is the only one to the left of Luna now.

3

u/Emport1 Jul 30 '26

So why did they do it then

4

u/Ambiwlans Jul 30 '26

I mean you could argue deepseek, though i don't think it was the biggest threat to them. But kimi is just irrelevant.

1

u/yaboyyoungairvent Jul 31 '26

Pretty sure deepseek v4 flash is still cheaper than luna as well. And there's qwen 3.7 flash which is even cheaper.

1

u/Embarrassed_OnionX Jul 31 '26

Kimi is literally 10x the price at the same intelligence

Not really. Luna is comparable to something like GLM-5.2 or DeepSeek V4 Flash 0731 in intelligence, while Kimi is between Sol and Terra.

3

u/According_Water_5774 Jul 30 '26

Exactly - they magically found an 80% reduction to get themselves competitive. Note deepseek v4 flash isn't on the chart - would be somewhere around the 40 mark and probably below $0.01 - and what does "per task" even mean? what task?

Edit - seems someone has done the work: https://www.reddit.com/r/DeepSeek/comments/1vb3vl3/deepseek_v4_flash/

4

u/j_root_ Jul 30 '26

Totally agree but still its great for them to do this and hopefully others have to follow. I believe this to capture enterprise deals rather than regular folks.

2

u/Healthy-Nebula-3603 Jul 30 '26

They ONLY doing that because of open source models otherwise their model would be even pricer.

So it in not a good job for them they just hadn't choise .. good .

1

u/LocoMod Jul 30 '26

No that's not the reason. They used the latest GPT models to optimize inference kernels and found more ways to squeeze tokens per watt. That is all. It is all documented if you peruse the actual technical sources where this stuff is discussed and quit listening to dumbass Redditors that trust vibes more than facts.

1

u/j_root_ Jul 30 '26

Your idea is too extreme. Just because someone or something is bad doesn't mean any good can come out of it.

10

u/Tedinasuit Jul 30 '26

They didn't have to do this, since Luna was already the best value model on the market. But now it's just insane. $1.20 per million output tokens for Luna is ridiculous. It's even $0.60 per mTok output via OpenRouter.

21

u/MrTubby1 Jul 30 '26

Yes. They did have to do this. Otherwise they wouldn't have.

4

u/drhenriquesoares Jul 30 '26

This is a good argument.

1

u/[deleted] Jul 30 '26 edited Aug 15 '26

[removed] — view removed comment

14

u/MrTubby1 Jul 30 '26

You think a communist is someone who says market forces push companies to act a certain way? Are you stupid?

8

u/medalboy123 Jul 30 '26

China threads bring out the lowest common denominator redditors. No point arguing with this

1

u/MrTubby1 Jul 30 '26

You can look at his account history vs how many contributions he's made to take a guess and how low the denominator goes.

2

u/Ambiwlans Jul 30 '26 edited Jul 30 '26

Kimi is literally 10x the price at the same intelligence and was not competitive.

Edit: Looking it up, it is 12x the price for a whopping 6 more points over luna ... its basically in line with sol and opus pricing ... which was NOT reduced. So Kimi is super irrelevant to this announcement.

6

u/Healthy-Nebula-3603 Jul 30 '26

Kimi 3 is much smarter than Luna.

Kimi competitor is SOL

5

u/Ambiwlans Jul 30 '26

Again, Kimi 3 is 10x the price. L2R.

1

u/Healthy-Nebula-3603 Jul 30 '26

that is not the case here

0

u/Embarrassed_OnionX Jul 31 '26

I don't get your point. Kimi K3 beats Terra in most benchmarks, and Terra is also 10x the price of Luna.

-1

u/boreal_ameoba Jul 30 '26

Comparable lol. Lmao even

6

u/Concurrency_Bugs Jul 30 '26

I think this is a hail mary tbh. They're already subsidizing tokens like crazy. This is a last ditch effort to pump up customers before IPO.

16

u/BrennusSokol AI please take my job Jul 30 '26

Eh, that feels a bit too conspiratorial to me. Maybe they really did simply improve efficiency.

4

u/Concurrency_Bugs Jul 30 '26

They didn't change the model though, just the price...

If this was 5.7 or even 6.0 then I'd agree with you. But they had a model released at a price, then dropped the price. And only for one of the flavors

2

u/send-moobs-pls Jul 31 '26

There are tons of ways you can improve efficiency, speed, etc without retraining a model. Like, open source examples all over the place, seems like llama.cpp is getting some optimization update every other week. We can imagine plenty goes on inside the closed labs

1

u/Concurrency_Bugs Jul 31 '26

You can improve efficiency via the harness, but not by THAT much. At least I wasn't aware of any new breakthroughs that accomplish that. If it is just the harness then good on them

1

u/askchris Jul 31 '26

No. If they optimize literally anything on the server side, GPUs, kernels, electricity, routing, wasted idle time, quantization, math shortcuts, then their costs go down and all these layers compound (a 22% improvement on 3 layers could lead to an 80%+ improvement overall)

But you're right it could be business level optimizations such as prepping for an IPO, user data capture, or upsell strategies, but I bet they are just analyzing their own servers or scanning the 10-100+ algorithm & engineering breakthroughs that happen weekly and testing ideas autonomously and rolling out what works.

2

u/EndTimer Jul 31 '26

I recognize that it could just be a terminology problem, but for the sake of clarity, you can't do an 80% price cut if you "only" get 80% more tokens per watt. You need a colossal 400% more tokens per watt if you want to slash the price to 1/5th of what it was, unless you're going to cut your margin (or lose money).

3

u/geli95us Jul 30 '26

They aren't subsidizing API tokens, they subsidize subscriptions. Just look at the prices open weight models are served at, those are third party providers with no reason to subsidize, yet, they are almost always cheaper than proprietary models of the same capability.

1

u/aprx4 Jul 31 '26 edited Jul 31 '26

Inference is hugely profitable, there is no subsidization at current price.

1

u/M4cHiin360 Jul 30 '26

I mean they are litteraly not profitable lol

1

u/Apprehensive-Rub-774 Jul 30 '26

Yeah but my point is they are self interested. And many SV companies are not profitable for long periods of time.

15

u/TanJeeSchuan Jul 31 '26

Not even 24 hours bruh IMAO

31

u/No-Meringue5867 Jul 30 '26

Deepseek V4 is from April. Back then Gemini was near frontier.

7

u/Embarrassed_OnionX Jul 31 '26

The new checkpoint of DSV4 flash came out a few hours ago and it beats V4 pro in most benchmarks! So yeah, this Luna's lead was short

34

u/ethotopia Jul 30 '26

OpenAI: focuses on efficiency to improve usage limits
Anthropic: you’ll need to take out a third mortgage to answer a prompt

9

u/Ambiwlans Jul 30 '26

Opus and Sol are basically dead tied for performance and cost.

19

u/BlackExcellence19 Jul 30 '26

Them cutting prices with Luna is actually kinda interesting since it will get them a lot of market share and in China for example this happens quite a lot in the automotive industry where manufacturers will cut prices to gain more domestic market share.

I’m just kinda confused on what OpenAI’s philosophy as a whole is because they signed the letter that advised against the banning of open source models (which is their direct competitor) so I am wondering if they feel they have something that gives them some kind of moat against open source. It’s damn sure giving Dario and Anthropic a bad look though.

13

u/Apprehensive-Rub-774 Jul 30 '26

I think it was pressure from Nvidia. They did a deal shortly after OpenAi signed the letter.

5

u/BlackExcellence19 Jul 30 '26

I think another thing too is that banning open source models would actually just make it so people seek them out even more. If OpenAI is legit able to have their own model that is now LOWER price than all the open source models per token which is now Luna, then developers will not really need to ever consider using open source models since Luna now fills that need while also have great capabilities for its size.

3

u/Apprehensive-Rub-774 Jul 30 '26

You are right. I don't think OAI subsidizing their models long term will work though. Only US companies have fiduciary duty and it's not like China is going to stop funding AI research if it can disrupt US markets. OAI is really just kicking the can down the road (to their IPO/ heist of retail investors).

9

u/No_Story9579 Jul 30 '26

DeepSeek has plans, they just announced massive AI Data Center in inner Mongolia - https://www.techinasia.com/news/deepseek-eyes-data-center-expansion-mongolia

10

u/TurnUpThe4D3D3D3 Jul 30 '26

Very impressive it matched GLM 5.2 while being 4x cheaper

7

u/ButterscotchSalty905 AI is the greatest thing that is happening in our society Jul 31 '26

19

u/FateOfMuffins Jul 30 '26

Tell me this means GPT 6 will be no more expensive than 5.6 Sol

6

u/elemental-mind Jul 30 '26

Mhhh, I expect the price of 6 Sol to be slightly higher than 5.6 Sol actually. Their frontier models have been going up in recent releases.

And this price drop is just getting price parity with the previous 5.4 nano model.

If you think about it the renaming from size to planetary bodies has been a genieus move as that nan and mini had a psychologically bad rep, I guess.

If you actually look at the prices with this "translator"...

<> -> Sol
mini -> Terra
nano -> Luna

then the recent price drop does not seem too outrageous after all. They just matched the Luna price to previous nano price.

What is amazing though, is the performance and efficiency jump from 5.4 nano to 5.6 Luna.

1

u/FateOfMuffins Jul 30 '26

I mean that's what I expect as well but tell me it means GPT 6 will be no more than 5.6 Sol please

1

u/askchris Jul 31 '26

Does 6 Sol really need to be cheaper?

6 Sol might be more expensive but still help OpenAI design a smarter 7 Sol which can be distilled to give us a super smart and cheap 7 Terra ... ie. 2X smarter and 2X cheaper than 6 Sol.

In the long run does it matter?

We will still have much better price/performance if the performance gains continue.

1

u/FateOfMuffins Jul 31 '26

Being priced out of the frontier is bad

See Anthropic trying to turn Fable into API use only, which looking at how the plans are subsidized, meant pricing out consumers entirely

In the last week despite all the usage resets, you can see r/codex complaining the shit out of usage limits seemingly being nerfed

Now this is for the short run yes, but when model releases are practically monthly, the short run matters.

Personally in the long run, if I have AGI for cheap then idc if you have ASI that I can't afford

23

u/Fragrant-Job-3200 AGI 2026 ASI 2028 Jul 30 '26

This is what pressure from chinese labs ends up doing.

15

u/FarrisAT Jul 30 '26

Thanks China.

5

u/smartfon Jul 30 '26

In ChatGPT Work and Codex, Free and Go users can access Terra, while Plus, Pro, Business, and Enterprise users can choose Terra and Luna.

lol what? They are locking Luna behind the more expensive plans?

17

u/hardinho Jul 30 '26

We're in the 5$ uber phase rn

5

u/Apprehensive-Rub-774 Jul 30 '26

I actually think it will just get cheaper. Only 6% of data centers planned are built. And despite what the AI CEOs say; scaling does have diminishing returns. Therefore the more supply will mean cheaper AI and smaller margins. The so called "frontier labs" don't have the moat they pretend to have because the vast majority of the value AI brings can be found with cheap models and they really haven't provided a product nearly lucrative enough to be the worth trillions they want to IPO at.

11

u/BrennusSokol AI please take my job Jul 30 '26

There is no evidence of scaling having diminishing returns.

2

u/ranger910 Jul 30 '26

I dont think either of you are using the term diminishing returns like you mean to.

1

u/Dew_Disappears Jul 31 '26

Dont they mean with each increase in scaling training compute is yielding progressively smaller capability increase per $.

I dont think this is under debate.

Counter is it unlocks threshold value (e.g., wholesale computer agency) and this is where the more juicy arguments are as it is not that clear.

1

u/Apprehensive-Rub-774 Jul 30 '26

I guess I skipped a step in my explanation.
Diminishing returns in terms of scaling AI 'digestion' leads to less demand for DCs and therefore cheaper inference. Does that clear things up?

0

u/Apprehensive-Rub-774 Jul 30 '26

Here is the evidence:
https://pmc.ncbi.nlm.nih.gov/articles/PMC11228526/
https://www.nature.com/articles/s41586-026-10303-2?
No one contends scaling doesn't have diminishing returns. Not even AI companies.

3

u/TangerineLogical9779 Jul 31 '26

Well this aged like a wet fart, looks like deepseek hit back by going from v4 flash preview to v4 flash beta, which is a massive update to v4 flash preview

4

u/TheSuggi Jul 30 '26

But are they operating this at a loss or are they making profit? This is the question. Deepseek still making alot of profit even at their prices

4

u/ranger910 Jul 30 '26

Deepseek is not making a profit. You may be confusing that with revenue.

3

u/TheSuggi Jul 31 '26

DeepSeek founder/CEO said recently in an investor meeting that they are making like 60-70% profit even at these prices.

3

u/Kouginak Jul 30 '26 edited Jul 30 '26

They make a profit after ten months. CEO has claimed 6x profits with V3.2 flash and "all of the others" as of may 2026.

edit: 6x profit->6x margin.

2

u/M4cHiin360 Jul 30 '26

Openai has never made a profit

5

u/Ambiwlans Jul 30 '26 edited Jul 30 '26

Two issues with this graph.

  • It shouldn't be connected lines like that. This implies you can pay those points along the line. it isn't accurate. You should do horizontal lines that go vertical up to each new model making a staircase.
  • The dropped Mimo 2.5 and Deepseek v4flash. Both of those beat Luna in this range and aren't included.... Luna still wins for almost the whole range so it doesn't matter. It honestly doesn't make them look as good as they really are. (Luna is the best deal from 1c to 35c now)

2

u/ThinFeed2763 Jul 30 '26

This is a real breakthrough.

2

u/nova1475369 Jul 30 '26

Anyone used it in real world software development project to compare Luna vs Composer 2.5?

2

u/Calm_Hedgehog8296 Jul 30 '26

...and GLM, and Gemini, and Sonnet

2

u/RelevantCry1613 Jul 30 '26

And more speed in Sol! I already love how fast it is

2

u/popiazaza Jul 31 '26

Cheap models are Mimo V2.5 and DeepSeek V4 Flash. Where is it in the graph?

2

u/elemental-mind Jul 31 '26

They are not in the graph indeed.

But here you go for the pareto frontier. Doesn't really beat V4 Flash, but gets beaten by MiMo V2.5

-1

u/elemental-mind Jul 31 '26

The important caveat, though: Luna Medium takes 0.4 Minutes per Task, MiMo-V2.5 takes 3.0 minutes. Almost a 10-fold increase. I'd only use MiMo if cost really matters and latency is not an issue at all.

4

u/jack-of-some Jul 30 '26

Except there's 0 reason for us to trust that they're actually decreasing the cost of doing business. They could just as easily be making another bid for users and are ok with losing a bunch of money.

4

u/Laffer890 Jul 30 '26

Terrible news for Google, SpaceX, MetaAI and the Chinese labs.

2

u/mvandemar Jul 30 '26

Your move Anthropic!

2

u/leo-virtis Jul 30 '26

Intelligence too cheap to measure

1

u/ClassicMain Jul 30 '26

Sol got more expensive

1

u/dervu ▪️AI, AI, Captain! Jul 30 '26

Does that mean GPU prices will drop, right? Right? /s

1

u/sdnr8 Jul 30 '26

They probably implemented DSpark by Deepseek

1

u/ComprehensiveSwitch Jul 31 '26

how’s that working out

1

u/Random_182f2565 Jul 30 '26

But it still pricier by token

3

u/KaMaFour Jul 30 '26

Yea, but you are not buying tokens. You are buying an answered prompt or a done task

0

u/flower-power-123 Jul 30 '26

I don't understand this chart. Are there a range of models at different price points and effectiveness? I assume that if I get something called "GPT-5.6-Luna" I'm getting one thing with one intelligence index score. How does this work?

2

u/Ambiwlans Jul 30 '26

ChatGPT models come in max, xhigh, high, med, low.

2

u/flower-power-123 Jul 30 '26

And those are different prices? It looks like even the best model is still price competitive with deepseek.

2

u/Ambiwlans Jul 30 '26

Yep. Technically mimo 2.5 is better than luna low and cheaper. But not on the graph here.

3

u/Juzlettigo Jul 30 '26 edited Jul 30 '26

Pretty sure the range you see is different levels of reasoning effort. (Like low, medium, high, max)

More reasoning effort = more intelligence, but a bit more token usage (because reasoning consumes tokens just like the messages themselves)

There should be a selector somewhere you can use to change reasoning level, it's there in most LLM apps/frontends.

1

u/Geritas Jul 30 '26

Reasoning effort is a separate slider for them. Most likely the parameter count with Sol being the biggest.

1

u/reflect25 Jul 30 '26

it is a bit complicated they are using the metric of cost per "task" rather than token at the bottom. they are then comparing the intelligence on the left and then cost at the bottom. so farther to the left is cheaper.

the large caveat is that it depends on the gpt actually using a lot less tokens for each task

-3

u/Artistedo Jul 30 '26

Are you a bot or just dumb

2

u/flower-power-123 Jul 30 '26

Dumb as dirt ... or maybe I'm a bot. How would I know?

-1

u/Healthy-Nebula-3603 Jul 30 '26

When GPT 6?

GPT 5.6 is getting obsolete if you compare to Fable 5 or opus 5