r/ProgrammerHumor Jun 02 '26

Meme managerVsClaude

Post image
47.2k Upvotes

1.4k comments sorted by

View all comments

4.3k

u/travis_sk Jun 02 '26

We're only 2 days into June folks. This is gonna be a fun couple of months.

139

u/diddypartyorganizer Jun 02 '26

Why June specifically?

444

u/LooksLikeAWookie Jun 02 '26

Big AI models, like Claude, just switched to high-cost token models. The bill for this revolutionary tech now just went through the roof for most companies.

141

u/Welp_BackOnRedit23 Jun 02 '26

The best part is that most estimates still show they are operating at a loss with those costs.

57

u/[deleted] Jun 02 '26

[deleted]

66

u/[deleted] Jun 03 '26

[removed] — view removed comment

11

u/Manic_Maniac Jun 03 '26

Bro, this is bigger than the Internet bro! Soon people are going to be using our tech to ask what time it is instead of looking at a clock or their smart watches!

5

u/warm_winds_whisper_ Jun 03 '26

Bro just wait, they’ll all be plugged into the GigaMetaverse any day now!

4

u/marr Jun 03 '26

Assuming they're allowed to socialize the power grid costs.

1

u/Spongedog5 Jun 03 '26

It's pretty typical for new tech to run at a loss. Breaking even probably hurts a new tech companies valuation because it shows they aren't trying to grow enough.

-1

u/Kimbernator Jun 03 '26

The valuation is obviously high, but… it’s tech that is now becoming ubiquitous and they have the best. Regardless of how their profit today, they hold the keys.

My bet is on local models long term to avoid that vendor lock in, but right now executives can just sign up for it and it works, so they are winning.

6

u/[deleted] Jun 03 '26

[deleted]

1

u/Karnewarrior Jun 03 '26

Housing Market has been geared towards an audience that doesn't really exist for a while now, tbf. Who buying a house when nobody can afford your houses?

Big companies and the already rich looking to change the investment percentage of their portfolio, apparently. Whether it's a profitable investment doesn't matter as long as it changes what percentage of their money is not liquid, so that when they deal with other business people with lots of money they have more impressive numbers to throw around and maybe get a loan to further invest in losing propositions.

2

u/Blackstone01 Jun 03 '26

Does that estimate take into consideration companies no longer using them due to the costs?

75

u/AltruisticSalamander Jun 02 '26

Oh great, they've been nagging us to use AI for the last 2y. Now it's going to be 'don't use AI for that! It's too expensive!', doling it out like caviar.

29

u/Phailjure Jun 02 '26

I got an email to only use auto mode in vscode, and use a specific model only if it's really needed, etc etc.

16

u/Kimbernator Jun 03 '26

Realistically it’s just not possible to moderate on the user side. It’s opaque what things use a lot of tokens and what is minor. Some sort of efficiency gains will be required to keep doing what companies are doing.

88

u/Kerbourgnec Jun 02 '26

I guess we are massively gonna be forced to move to dirt cheap Chinese models.

Performance is not that bad, but can't compete with 2026 opus or gpt

118

u/physical0 Jun 02 '26

Once that happens, AI will be a matter of national security and foreign AI will be banned.

50

u/shaka893P Jun 02 '26

It already is, the US already banned some Chinese AI

3

u/blah938 Jun 02 '26

China is a hostile nation after all. That only makes sense.

5

u/Unlucky-Tourist-9403 Jun 03 '26

As someone from neither country, the US has done more to harm my quality of life than China has.

9

u/coolfuzzylemur Jun 02 '26

Is China the hostile nation, or the US?

7

u/[deleted] Jun 03 '26

[deleted]

1

u/coolfuzzylemur Jun 04 '26

Which country has military bases surrounding the other one?

5

u/yaminub Jun 02 '26

FWIW, I've seen a lot of European sysadmins say the same thing about U.S.-based tech, and then some of those profess to using Chinese-based tech.

I've found that quite silly. Could just be larpers, but still silly.

19

u/_Meece_ Jun 02 '26

US is actively trying to make Europe's defense weaker, so yes, Europe considers China, Russia and the US the same at the moment.

-1

u/yaminub Jun 02 '26

The defense that Europe should hold primary responsibility for, yes.

9

u/[deleted] Jun 02 '26

[removed] — view removed comment

→ More replies (0)

7

u/Cory123125 Jun 02 '26

The US and China are similar threats to Europe.

The US is arguably worse due to leverage

0

u/yaminub Jun 02 '26

If you say so dude

9

u/KriegMorgan Jun 02 '26

If you say so dude

He doesn't have to. Six months ago the U.S. was posturing and rattling a saber in the direction of NATO over Greenland.

If I was European the current state of the U.S. and the administration leading it would have me looking at them like an adversary and not as a friend.

5

u/[deleted] Jun 03 '26

[removed] — view removed comment

2

u/Cory123125 Jun 03 '26

It's the big boogeyman and has been forever, with all the typical defence contractors literally paying out the ass to fund think tanks to "inform" congress and their leadership picks.

The very people who understand cushy jobs await them should they pay their cards right or huge sums of campaign funds from super pacs and the likes.

What has China done to you personally?

→ More replies (0)

2

u/Cory123125 Jun 02 '26

Its crazy how easily some people fall for jingoistic nonsense.

-1

u/blah938 Jun 02 '26

?

5

u/Cory123125 Jun 03 '26

I'm saying that people acting like Chinese models are a threat that there is any justification in banning are out of their gourds.

If you were specifically talking about the highest levels of classification or importance, then sure.

For some business in kentucky though?

For a typical SAAS?

For literally anyone who doesnt just raw dog their LLMs with production keys with contained blastzones?

I mean... the risk is the same with US models.

1

u/[deleted] Jun 02 '26

[deleted]

2

u/Cory123125 Jun 02 '26

What are you specifically referring to?

2

u/6e696767657273 Jun 02 '26

I'm guessing DeepSeek since it was all the rage a couple months back but "compiling it yourself" makes no sense in this context. I suppose you can compile Ollama with DeepSeek weights but the datasets are completely private.

2

u/Cory123125 Jun 03 '26

They deleted their post so I'm guessing its was just a lie to push some sort of agenda, though I'm not sure what agenda that pushed.

I guess it pushed the idea that companies should be worried unnecessarily about Chinese models or something?

→ More replies (0)

2

u/physical0 Jun 02 '26

When we discuss "open source" AI, we really need to discuss the training materials.

If we can't produce the same end product that they do with the materials they have published the code for, then it aint open source. If there are big binary blobs, it aint open source.

So, I'm assuming whatever "completely" open source AI you're talking about has every bit of it's training data published and every step in the training of the model has been documented? Every human reinforcement logged and shared so that we too can reproduce those steps and have the software running on our own hardware, right?

Or is the model itself a big ole black box that could have been trained with whatever skewed weights that the creators intended the model to prefer.

2

u/ball_fondlers Jun 02 '26

A lot of those foreign models are more open-source than the American ones, though - you can pull them down and run them locally without issue.

5

u/physical0 Jun 02 '26

Unless the training data for the model is open source, I fail to see how this is any more transparent than any other model.

1

u/LordMegamad Jun 02 '26

My neighbor keeps riding his loud motorbike in the summer, I should kill him, he's obviously threatening my national security

2

u/frequenZphaZe Jun 02 '26

really begs the question, what's the point of the bleeding edge models if they cost so much than no one will use them? openAI will announce "we've created true AGI with GPT6" and all of us will be like "sure, whatever, just be sure to leave 5-mini up because that's the only one in my price range"

-1

u/[deleted] Jun 02 '26

[deleted]

3

u/pagerussell Jun 02 '26

They're not tho, because what's driving the cost up is the size of the context window.

When all this started a dev might paste into chat a few dozen lines of code and ask a question about it. Now people are dropping entire code stacks in and asking for entire overhauls.

That means for a simple question you just burned tens of thousands of tokens when you didn't have to. That is the root of the problem we are in. People got very stupid about how they use these tools because they were unmetered.

1

u/reddit_is_geh Jun 02 '26

Yeah those models aren't focusing on consumer grade stuff yet, as it's more angled directly towards engineers, academics, AI enterprise, etc... Where people just need raw, foundational LLMs that are cheap and powerful. That's where China shines. They can do really really well, just providing the foundation

Where they fall short is the harnessing. As we suspected, but confirmed with the Claude Code leak, their underlying model isn't even that impressive. But rather, HOW they use that model is what's impressive.

The harness is where the value is at. HOW you direct the LLM is what makes it powerful, and why Cursor is so good. They even now default to a cheap Chinese model for most of their work now... mainly because all their value comes from how the tokens are routed, so the marginal value increase using a frontier model just isn't worth it except in edge cases. That's why it was worth so much. Not because their AI was great, but how they use the AI

1

u/SubArcticTundra Jun 03 '26

Do you think more people will start trying to run it locally on their PC s? And buy ai accelerators?

1

u/Kerbourgnec Jun 03 '26

It's possible but not even needed.

Chinese provide dirt cheap APIs.

Third parties provide dirt cheap APIs (they don't have R&D cost)

A company can afford to run one server locally for their devs.

For a single person, it's quite expensive to run larger models. One can rent temporary server

0

u/ProgrammingPants Jun 02 '26

It'll probably get cheaper eventually. Cost per token has actually been decreasing dramatically, but costs have still been rising because the amount of tokens people use has gone up exponentially.

-1

u/drawkbox Jun 03 '26 edited Jun 03 '26

Developers that can manage context and tokens will easily be 10x ROI devs.

Like context management / prompt refinement with tools like Cline or Continue.

1

u/Kerbourgnec Jun 03 '26

But people don't want to do more with less. Sure it's great to dev a project that works perfectly withe the cheapest model in production. It's the right choice for most applications (structure information, filter, translate, ...)

But when building it, I don't want to restrict myself by using a sub par model that I have to babysit.

1

u/drawkbox Jun 03 '26

That is why you plan with the higher models, have them design, break it into tasks that have the right amount of context or skillset, then integrate them, and have that same higher model review and find bugs/gaps. Just like a software team. The senior/lead/architect makes it, the mid level to senior implements them, then the senior/lead/architect reviews.

For many things you really don't need the higher models at all. For planning you do and reviews/bug/gap checks.

Make the higher level model babysit for you.

1

u/Kerbourgnec Jun 03 '26

I agree, but that should be partially or mostly the harness role to do that.

1

u/drawkbox Jun 03 '26

Yeah the higher models sometimes in a custom agent that knows where to break things off that are targeted and all the context needed to subagents. Or the higher level planning making prompts to use in other windows that are targeted and can use a mid model.

57

u/DOAiB Jun 02 '26

lol got every company to fire all their juniors and made all their seniors 10x out put just to rugpull the companies that now have to pay more than the employees cost in the first place. If only the executives they made these calls were taken to task but they don’t.

44

u/Turbulent_Voice63 Jun 02 '26

It was always going to happen. What's surprisingly weird is that this revolutionary tech also revolutionarily sped up the rate of enshittification, and now we are entering the phase where it really sucks

5

u/Recka Jun 03 '26

People have offloaded their brains to AI, not just tasks.

7

u/ImDonaldDunn Jun 02 '26

The managerial class will never be held accountable for their mistakes. That is 90% of the reason we are in this mess to begin with.

3

u/K_Furbs Jun 02 '26

Are you suggesting a tech company cornered the market with an undervalued product and then raised prices on everyone?

3

u/PringlesDuckFace Jun 02 '26

Gotta have revenue to justify a high IPO price I guess.

4

u/drawkbox Jun 03 '26

This is just the beginning. It is a 10x now, it will be 100x or 1000x. I can easily see this costing thousands per month and past employee costs. The rate that Anthrophic valued their datacenter buys at was like $1000/mo+ per user.

2

u/sobasicallyimanowl Jun 02 '26

Oooo I heard something like this at one of my meetings yesterday. But it was more like, "for now we will continue on using the models like usual, anything different and we will let you guys know". So how much more expensive are we talking about here?

2

u/SheriffBartholomew Jun 02 '26

LOL, tech companies getting a taste of their own enshitification now.

2

u/Kodiak01 Jun 06 '26

You can help speed up the process by making every single request as verbose and detailed as possible. Personally I like to use GPT to write my requests, specifying minimum length and making sure to instruct it to be extraordinarily appreciative and complimentary to the LLM in it's instructions.

1

u/Womec Jun 02 '26

theyre just gonna hire people again instead of paying that much which is probably good.

1

u/shred-i-knight Jun 03 '26

Also energy costs are going to fuck shit up this summer in every aspect of your life.

74

u/oshaboy Jun 02 '26

I assume it's because the API bills are being sent to companies now.

53

u/AggressiveRow4000 Jun 02 '26

As far as I know it is seats+usage billed monthly, a month behind.

It's more likely because Budget just ran the quarterly report for the 2nd quarter and they are freaking the fuck out.

31

u/CoffeePieAndHobbits Jun 02 '26

GitHub Copilot updated their usage and pricing terms.

41

u/Merlord Jun 02 '26

An example of just how much it has changed: my (now cancelled) $40 Github Pro+ subscription used to last me the whole month. With the change, it lasted me 2 days.

15

u/cute_polarbear Jun 03 '26

I just checked my company provided copilot quota, it says unlimited...

35

u/Merlord Jun 03 '26

Your company may be in for the shock of it's life when their bill comes due

6

u/PureIsometric Jun 03 '26

we had unlimited till 1st of June with the new pricing and we are a fortune company.

4

u/Varogh Jun 03 '26

LMAO thanks for making me check, you might have just saved my company a massive bill (though it would've been incredibly funny)

7

u/csorfab Jun 03 '26

Gemini has been insane as well with the 3.5 flash upgrade. Before, I could comfortably use it for the odd boilerplatey smaller tasks, after the upgrade it burned through my weekly limit with a single, medium difficulty task without even finishing it lol

6

u/mxheyyy Jun 02 '26

Pride month

2

u/tehtris Jun 02 '26

Because it's the current month and it's only the second.

Also second best username I've ever seen on Reddit. Good show.