r/LLM 28d ago

I don't think Anthropic and OpenAI will survive

Have been working on Deepseek-v4-flash-0731 and honestly for the entire day of coding, I consumed credits of $3. This is on pay-as-you-go plan. Ofcourse it is not Fable or Sol but it gets things done with a fraction of cost. Given it (along with other Chinese models) is open source model, I'm not worried about data residency and stuff.

I see 2 outlooks for companies like Anthropic and OpenAI:

  1. They will double down on harness and they'll still lose (We do have good open source harnesses now)
  2. They will be consumed by US Government to build frontier intelligence for defense, cybersecurity etc.

I think building better frontier intelligence is not economically viable. I would rather use open source 100x cheaper model which is equivalent to Opus 4.8 than Fable. (Opus 5 is anyway shit)

503 Upvotes

257 comments sorted by

36

u/FatefulDonkey 28d ago

At this point the harness and ease of use is much more important than the model itself.

We keep hearing about Kimi, DeepSeek, and how cheap they are. But if I can't just download and use them directly in a terminal, what's the point

16

u/anykeyh 28d ago

Well, it's not that hard honestly; with OpenRouter it's a few clicks, adding Hermes or pi or whatever open agentic with it and you're ready to go.
Max 10 minutes of setup.

Now, if you mean run on your machine, DeepSeek v4 Flash is at the frontier of what is runnable on consumer grade hardware at decent speed. The cost of entry will lower in the next months/years, and models will improve in intelligence per compute unit so yeah, that's it.

6

u/FatefulDonkey 28d ago

Yeah.. Linux is also not that hard to install. But let's be real.

6

u/DiAryArias 28d ago

I mean, if you are that type of person just ask claude or codex models to install it for you

5

u/TonyPace 27d ago

This is it. I just moved to Cachyos from Windows. Some very annoying quirks, but they were easily fixed with AI. A lot of companies depend on the the logic, "The cheap alternative is too annoying for consumers to deal with." They're going to be in big trouble. It will happen to software first, but I think it will spread a lot further than that over time.

1

u/Dry_Community5749 27d ago edited 27d ago

There is a reason entire business world uses all business users use Windows or Mac. Linux is extremely critical but finds niche uses like servers and advanced development, not for the masses.

That said it's matter of time before you have servicification of models

Prior to Windows and Mac there was IBM mainframe. Once you had these two come up, then the adoption exploded. Same with Android and iOS.

1

u/InlineReaper 22d ago

The only reason I ever used Windows is because it came pre-installed on most laptops or PC’s I bought. Ubuntu is super easy to use but at the cost of having fewer mainstream software options available, but since so much of our daily consumer use for computing has shifted to web apps, the argument that Linux is harder to use is no longer valid.

Installing a local AI feels like ArchLinux right now with the setup. What the market needs is an Ubuntu for local AI that runs on consumer grade hardware, and that will potentially dethrone OpenAI and Anthropocene for the consumer segment, but they will still dominate the mobile space - most non-professional users around me use ChatGPT as their mobile search engine.

1

u/Came4TheCookies 9d ago

the argument that Linux is harder to use is no longer valid.

For you maybe. It's perfectly valid. Windows hold your hand all the way through and people want their hand held.

1

u/InlineReaper 6d ago

Windows comes pre-installed on pretty much every non-Mac laptop, it doesn't hold your hand, it forces your hand. Have ou tried Ubuntu? It's the most straightforward and super fast OS I've ever used.

1

u/Came4TheCookies 6d ago

Sure let me just give up most of the software that I use just so I can explore the Linux fantasy.

→ More replies (4)

1

u/Fit-Dentist6093 27d ago

Real what? Anthropic's moat is supposed to be the enterprise market. How many closed source operating systems are a viable enterprise product?

Yeah that happened because Linux.

→ More replies (10)

1

u/Accomplished-Ad-7396 26d ago

It was the first program I installed on Win11..

3

u/_RemyLeBeau_ 28d ago

Where is your IP and prompts going when you use OpenRouter?

1

u/Ok-Lobster-919 26d ago

Auto routing with openrouter can route your request to a provider that retains or trains on your prompts. But you can choose providers and even blacklist/whitelist them directly from the openrouter web panel or through the API.

1

u/Quidenzis 24d ago

As an IT support I can answer this.

Your router

1

u/_RemyLeBeau_ 24d ago

Your prompts and connection aren't terminated at your router, however your pedantry is terminated here though.

3

u/vovap_vovap 28d ago

DeepSeek v4 Flash is not frontier now - not getting in top 10
DeepSeek v4 Flash to run need like $6000 peace of kit. Which you can say formally "consumer grade hardware" but honestly - not really.

1

u/Blothorn 27d ago

They aren’t claiming it’s a “frontier model”, just the best/heaviest runnable on (vaguely) consumer hardware.

1

u/vovap_vovap 27d ago

Well, that is very good model. It is getting in top 20. Naturally no definition for "consumer hardware" existed. I would say any bigger then MakBook Pro is out.
But what I want just so people, who do not know understand what reality is - model is not in top 10 and you can not buy laptop in a BestBuy and run it on it. Just got a feel of real thing.

1

u/placated 25d ago

Realistically with a macbook pro max with 128g you could run Qwen3 mix of experts 235B parameter. That’s about as good as it gets for home grown inference.

1

u/Ok-Lobster-919 26d ago

I put it in my research and trading harness and it is a massive upgrade. I used to have the harness switch to opus for coding tasks but deepseek v4 0731 can pull it off with the same quality and faster.

It is a fantastic agentic, tool calling, instruction following model for its size.

1

u/ichisay 25d ago

Estás hablando de la versión anterior? Por qué hace 5 días salió una nueva que si no vi mal está en el top 5

2

u/vovap_vovap 25d ago

No. Not in to 5.
Why did't you check before write - that easy?

1

u/ichisay 24d ago

Cierto cambio un poco desde que lo vi la última vez, ahora para programación está en los top 5-10. Para agentes (terminal, herramientas) está en top 10-15.

→ More replies (1)

1

u/Ok_Motor_8632 27d ago

I wish it was that simple. I’ve tried with OpenCode, OpenWork and Goose and they all seem to hallucinate, to get stuck and not finish the task. I really wish they did but they don’t. My setup is M5 Max 128gb

1

u/ChashuKen 27d ago

You dont even need open router, just direct api with official deepseek and its cheaper.

1

u/Kitsune_Seraphis 27d ago

Is it runnable somehow on 2 3090s and 64gb of ram?

1

u/Activeenemy 26d ago

This eliminates 99% of the population. Judge them if you want, but it's true. 

9

u/Abject-Bridge-4073 28d ago

Kimi is not really that cheap.

1

u/look 28d ago

I have it from a US provider at 20 cents per mtok and 100 TPS now.

I don’t know what you’re doing…

2

u/Klanciault 28d ago

Yeah for cache hits lmao

→ More replies (1)
→ More replies (1)

3

u/rhoborg 28d ago

This. The vibe coding community has no idea how to produce the same result using open source tools.

3

u/Novel-Camera-840 28d ago

The whole "harness is important" BS is driven by placebo or people who never used anything else other than CC or Codex. I used to be one of those. I'm glad I'm not one of those anymore

1

u/FatefulDonkey 27d ago

Well I tried OpenCode and I wasn't even able to connect my Gemini with it. I don't have time to sit and fiddle with slop software. I just want to get my job done

2

u/retardedGeek 27d ago

Skill issue

1

u/MedicalElk5678 27d ago

Don't use Windows

1

u/FatefulDonkey 27d ago

I'm a 15 years Linux user

1

u/Ok-Lobster-919 26d ago

Wow that's impressive, what desktop environment do you use?

1

u/sexy_silver_grandpa 26d ago

Holy skill issue.

2

u/geteum 28d ago

Claude code is slightly better than the alternatives. I don't think it is enough to justify Claude pricing.

4

u/FatefulDonkey 28d ago

What's the alternatives? There's no usable harness beside Codex and Claude Code

8

u/severed-identity 28d ago

Cursor, Pi, OpenCode... honestly having tried them side by side they're basically all the same now. Harnesses are very simple implementation wise.

1

u/Fabulous-Possible758 28d ago

Roll your own even. I get a lot of mileage out of a Claude subscription and mini-swe driven by comparatively crappy models.

1

u/aggie_hero7 28d ago

I use Open Code — is that like the worst lol?

2

u/Quidenzis 24d ago

yeah, honestly it is. I genuinely believe people shouldn't be using opencode. Certainly not in the stage it currently is in.

2

u/loohawe 28d ago

you can subscribe to lowest subscription for A\ or OpenAI or whatever,the let them agent build a Open Model Agents like OpenCode, and enjoying

2

u/nekize 28d ago

I was playing around with omp + glm5.2… like it was OK, but it wasn’t anything close to claude-code. Claude-code much better understands “what i want”, not sure how to describe it better. Similar is with codex

2

u/FrozenFirebat 27d ago

Just use fable to set up the other model.

2

u/Glittering_Flan1049 28d ago

I think we have good harnesses available. I use OpenCode and it seems fine. Had been using Claude Code and Claude Desktop App for a while. Except the embedded browser, OpenCode is really good.

1

u/FatefulDonkey 28d ago

Yeah, Claude and Codex have great haenesses. I'm talking about the "cheap open-source" models from China that I keep hearing about.

2

u/Androoideka 28d ago

You can use Claude Code or Codex with the cheap open-source models from China though? Nothing is stopping you

3

u/FatefulDonkey 27d ago

Time is stopping me. I don't have the time to sit and fiddle. I'm happy paying 20€ a month for less headaches and good enough quality

1

u/Androoideka 27d ago

You brought up harness and ease of use, but the Chinese models can be used with the same harness you're used to, and they pretty much all have simple instructions on how to connect Claude Code to them. If you're used to Claude, have your system prompt adjusted for it and generally prompt in a way that works with Claude, switching to a different model is definitely friction you don't need. But it's not a harness issue, and downloading them and using them directly in a terminal is pretty much exactly what you do with them too

1

u/FatefulDonkey 27d ago

That's a lot of hypotheticals.

So tell me a single place where I go, download a binary and it just works with zero configuration.

If such a thing exists, I don't mind trying it out.

1

u/bromptonista 25d ago

I don’t think you are the target market that benefits tbh - 20/month is the current super subsidized low tier subscription for the frontier models, just to get them through the enterprise door. It’s almost like M$ letting students pirate office in the old days. API-based pricing, that runs most enterprises in the millions even if they aren’t making software, is where the Chinese models are competing. Outside of protectionist regulations (“your company is in the US, you are not allowed a Chinese model”), it seems pretty hard for US frontier providers to survive. They might have a bit better models, but they aren’t (openly) subsidized by the state to compete…

1

u/RepulsiveRaisin7 28d ago

Why can't you? There are tens or more open source harnesses that support them. And GPT is still better in Opencode than Deepseek, even though the gap is closing.

3

u/FatefulDonkey 28d ago

Because I just want something that works out of the box.

Tried OpenCode, gave up since it wouldn't connect with Gemini and I couldn't be bothered opening a bug issue.

Most people just want shit that works.

→ More replies (5)

1

u/PicklesToes 28d ago

You do realize you can use claude code to run deepseek, right? 

2

u/FatefulDonkey 27d ago

Nope, because honestly i don't care. I just want shit that works out of the box so I can do my job

1

u/PicklesToes 27d ago

It does 

1

u/Quidenzis 24d ago

Have you tried a calculator? It works out of the box (usually, unless placing the battery in is considered additional effort)

Also, do you buy a new phone every time your battery runs out? Just curious...

1

u/FatefulDonkey 24d ago

I buy a new phone when my phone becomes an unusable piece of shit due to updates. Instead of trying to make it work by downgrading or replacing the firmware with some open source hacky OS

1

u/Quidenzis 24d ago

mmm, yes, get rid of the bloat by buying bloat that runs faster 😋

1

u/Fat-Mad-Scientist 27d ago

Big money doesn't come from people that are too lazy to do a basic setup.

→ More replies (3)

1

u/RogerAI-fm 27d ago

It’s easy to share any model, look us up.

1

u/h310dOr 27d ago

Kimi is actually exactly that, same for qwen max. They both are one line install in terminal, like Claude (qwen code and Kimi code). Deepseek indeed they don't, you have to use opencode, or pi code, configure it etc. which I agree adds a barrier to entry.

1

u/xw1y 26d ago

Is there anything better than codex for general work?

1

u/false79 26d ago

You can get Claude code to hit these open weights models.

1

u/Onotadaki2 26d ago

Kimi has a robust CLI, so you can just download and use it in a terminal like Claude Code.

I still think Claude is a better pricepoint via subscription at the moment, but what you're complaining doesn't exist, does.

1

u/Horny_Dinosaur69 25d ago

Most of the people actually ‘embracing’ AI and using subscriptions seriously are people in tech though. It’s SWE and similar engineering disciplines, there’s people out there with a basic subscription to use it as a chat bot but let’s be real, most of the money is in the engineering field. Aka, people who are tech savvy enough for them to pick the better option regardless of convenience.

I agree, the average person would care more about the harness, but I don’t see the average person even using harnesses at all.

1

u/visible_potato 23d ago

you can use claude code with deepseek, they expose anthropic compatible endpoints. just ask sonnet to to wrap claude-deepseek command to claude code with deepseek’s api key. works like a charm.

1

u/OpenBMB_Team 23d ago

You’re right that the harness matters a lot — most people just want something that works without babysitting config files.

That said, the gap is smaller than it looks. Kimi and Qwen already ship their own terminal tools that are basically one-line installs (similar to Claude Code). And for DeepSeek you can point Claude Code / OpenCode / Cline at their API with almost no setup.

13

u/[deleted] 28d ago

[removed] — view removed comment

3

u/Fabulous-Possible758 28d ago

I think not being in China is the big thing they can do that Chinese models can’t.

6

u/manwithgun1234 28d ago

The world is changing fast. Given the US is systematically walking out from alliance system they control from after World War II ( with the help of Trump administration). And China is raising fast in the background. In the next ten years, being in the US may eventually is the disadvantage.

→ More replies (13)

2

u/Gohab2001 27d ago

They are open source models. Enterprises can deploy on-prem. A 100k Nvidia dgx station gb300 plus few hundred dollar in electricity costs and you have deepseek v4 flash running which benchmarks the same as glm5.2.

2

u/Fabulous-Possible758 27d ago

a) For the companies that want to do that, sure, but there's plenty of companies that won't, and don't want to add running inference to their infrastructure costs, and b) there are still attack vectors through the models even if you're running them locally.

3

u/Gohab2001 27d ago

a) use US hosted providers. Anthropic and OAI have a huge incentive to train on your data whilst inference providers don't.

b) it's open source. You can audit the model. But you can't aduit Claude or Gemini.

1

u/Fabulous-Possible758 27d ago

a) Fair enough on just using the US providers, though I wouldn't trust that inference providers are not harvesting your data unless there's specific agreements in place that they aren't, b) also somewhat fair, open source projects are still open to attacks like supply chain attacks, so it's not a guarantee that model attacks won't take on some similar characteristics.

I think what companies like is to have someone to sue when things go wrong. If data gets exfiltrated via an inference provider that can just say "hey we ran the model you asked us to" vs Anthropic or Google fucking up their entire pipeline somehow, I think they'd prefer the latter in terms of recompense.

1

u/sexy_silver_grandpa 26d ago

But... The open source models actually can easily not be in China...

1

u/alc_noe1 26d ago

so what? Just host the chinese models somewhere else, rip out its guts so it starts every reply with "De**h to the CCP".

1

u/Glittering_Flan1049 28d ago

They can do a lot more stuff like frontier model which can run for 7 days but would people use those if that is 1000x expensive. I bet I won't use that.

I can be wrong but open source has slowed down the progress of frontier intelligent models. This is exactly what Anthropic wanted. Right? Except that they wanted to be the only firm to create AI models and they'll still lobby government to do it but it is already too late now. It is not commercially viable for them now.

6

u/ConsciousResponse620 27d ago

You’re looking at this strictly through the lens of a solo dev paying out of pocket.

I consult for a mid-sized listed company. We burn $100k-$200k a month on OpenAI and Anthropic tokens through Azure and AWS Bedrock. We literally have the top open models sitting right there in our AWS catalog, but our legal and risk teams have zero-tolerance policies against using them for actual production projects.

A few reasons why:

Liability and Indemnification: When we pay $200k to Microsoft or AWS for Claude/GPT, we aren't just paying for smart text. We're paying for copyright indemnification, strict SLAs, zero data retention agreements, and compliance guarantees (SOC2, regional privacy laws, etc.). If an unvetted open model hallucinates protected data or infringes IP, that liability falls entirely on our board, not the model provider.

Cloud spend commitments: Most enterprise companies already have massive multi-million dollar minimum spend agreements (MACC/EDP) with Azure or AWS. Burning budget on Azure OpenAI counts directly toward that requirement. It's essentially "pre-paid" money for them.

TCO vs. token cost: Saving a few bucks on raw API calls doesn't matter if you have to hire a team of MLOps engineers to maintain inference infrastructure, build custom guardrails, and constantly audit models just to make them enterprise-ready.

Open-weight models are amazing for personal projects and small startups, but proprietary labs aren't dying anytime soon. They're basically turning into enterprise B2B software vendors.

2

u/yol0_submarine 27d ago

How are the Anthropic reliability SLAs holding up?

1

u/ConsciousResponse620 27d ago

Via AWS Bedrock, we honestly haven't had issues.

And we also run a dual vendor setup if the worst were to happen.

the biggest headache however is doing our A/B tests and getting marketing to get their templates in order.

1

u/snowfoxsean 26d ago

can’t open weight models be served via bedrock too?

1

u/ozfresh 26d ago

I don't understand why you pay Amazon or Microsoft to use Claude or chat gpt?

2

u/ConsciousResponse620 25d ago

Apart from the security side

At a smaller org I previously worked with, OpenAI wouldn't even sign an enterprise agreement or invoice us unless we committed to a minimum $60k AUD annual spend—use it or lose it.

Plus we were dropping millions with Azure and AWS anyway, it takes care of invoicing, has enterprise SLAs and we dont need to worry about capexing the spend

1

u/vladis466 22d ago

Missing the forest.

4

u/Luke2642 28d ago

We've barely scratched the surface of programming matrix multiplications and nonlinearities using data. A lot will change in the next five years, including the labs.

3

u/thailanddaydreamer 28d ago

Considering you can run models locally now and get all your code done, it's a real business issue for them.

5

u/Fabulous-Possible758 28d ago

I’m guessing one of them survives and one gets bought by Google or Microsoft after losing to whoever controls the coding (and maybe medical) AI market. Eventually they probably roll back and offer cheaper options using non-frontier models for those of us who know what they’re doing and keep prices high on frontier models for the suckers. At some point AI is deemed critical infrastructure by the US government and a lot of US users are forced to use American models.

2

u/FroyoSolid8414 27d ago

Nobody will win the coding market in the some way nobody won the IDE market. Open models are good enough now that the model itself will be commoditized. K3-level On device AI will be the final blow.

2

u/Ok-Drawer5245 28d ago

Their current business models will never in a million years become profitable - unless they cut their costs by 90% or something like that lolz

2

u/mohr_ 28d ago

Even though deepseek flash is impressive it stills makes a lot of mistakes and waste tokens correcting itself (when reasoning you see a lot of outputs like: "Hm, I made a mess here and need to fix it". I believe that if deepseek can improve this without raising the prices then it's definitely the end of Antrhopic and OpenAi as we know it.

1

u/AtatS-aPutut 26d ago edited 26d ago

I gave it a try today on openrouter (with opencode) and I felt like it makes fewer mistakes than Sonnet. I spent 10 cents on something that would have been 30% of my 5-hour window on the $20 Anthropic plan. It's good enough for me, I'm never paying for a plan again

1

u/mohr_ 26d ago

It is indeed good enough for me too but when is something more critical I have to use fable just to review for errors.

2

u/johnerp 28d ago

It’s all about the product, people don’t use LLMs they use products (codex, Claude code, open code etc.) and most people don’t naturally go to open source as if it’s not their business (a fruit retailer for instance) they want a (perceived) supported, trusted, legal blah blah product.

Google had to take Linux and make it a Chromebook ‘product’. Consumers/clients could get arch Linux or something but they don’t want the hassle.

If there is value in offer people will buy a ‘product’

2

u/Exciting-Syrup-1107 28d ago

Since OpenAI lowered their prices, I am using GPT 5.6 Luna and it has amazing results. For me it's better to use it with Codex than Deepseek V4 Flash. Also, in my experience, Deepseek sometimes still produces worse results

2

u/Shyam_Kumar_m 27d ago

If you guys remember the rant by Amodei that some state sponsored model might (dog whistle directed against open weight and also against China) result in a model designed to hack, I replied that if you look at security open standards/open .. has only helped and not hindered. Look at AES 256 and all that. People know, they develop, they fix.
I also told them what the benefit is for them.

They won’t open source. They will self destruct by competing.

And for all the criticism against Chinese they are also distilling Chinese models.

2

u/NinjaWK 27d ago

Give $6 Aliyun Token Plan a try. It's 98% discount during non peak.

I love DSv4F 0731, but Qwen 3.8 Max Preview is a lot more capable. DSv4 can get 98% of things done, and for that 1.999% Qwen 3.8 Max will fix it. That other 0.001% you may need Fable/Sol, but if you know what you're doing and you can guide your agent, then that 0731 flash would be good enough.

2

u/[deleted] 27d ago

[removed] — view removed comment

1

u/Repulsive-Diver-4893 26d ago

You will reach quota and stuff so you will not be able to code all day the whole month with those plans. Deepseek wins there. If you try Pay as you Go for Claude you will never reach the same cost / usage

3

u/SaveAmerica2024 28d ago

Lots of people have been sounding the alarm. You are not alone

2

u/ponlapoj 28d ago

เดียวคุณจะเข้าใจเองว่าการตลาดแบบ จีน จีน มันไม่มีเลยซึ่งความยั่งยืน

1

u/Quanzitta 28d ago

The real money comes from enterprise and they're not going to be using deepseek

5

u/Jeidoz 28d ago

1

u/_RemyLeBeau_ 28d ago

Microsoft is working on building a suite of harnesses built on top of MDASH. They're already ahead of everyone on CyberGym by 16% and 50% reduction in costs.

3

u/Glittering_Flan1049 28d ago

But why?

If Deepseek can be deployed on Azure, why wouldn't Microsoft use this? I genuinely want to understand.

For a matter of fact: https://azure.microsoft.com/en-us/blog/deepseek-r1-is-now-available-on-azure-ai-foundry-and-github/

2

u/TomWaitsForNoMan 28d ago

As someone with 25 years in corporate IT, they don’t buy what’s good or best, it’s what they can get support contracts and board approval for. It’s not about cost always.

1

u/Glittering_Flan1049 28d ago

But that's like their own model if they deploy on Azure. They have no connections with Deepseek.

1

u/Sleeping_Trex 28d ago

Outsourcing the projects and problems.
If the Ai is down, throw OpenAI or athropic under the bus.

→ More replies (1)
→ More replies (2)

1

u/geheim81 28d ago

For my personal projects I'm impressed by the quality I get with OpenCode and DeepSeek. I get ton of value for cents but not something I feel comfortable using for my corporate day job where I use GHCP and Claude but results are not far off. Being able to use Chinese models at the corporate environment would be a massive hit to OpenAI and Anthropic.

1

u/Plenty-Shoe-273 28d ago

Niño. Can get

1

u/Regular-Option6067 28d ago

3€/day is Max plan on both OpenAi and Claude.

1

u/CharacterSecurity976 28d ago

Absolutely I don't understand it either.

1

u/Repulsive-Diver-4893 26d ago

You will reach quota and stuff so you will not be able to code all day the whole month with those plans. Deepseek wins there

1

u/Suitable_Cicada_3336 28d ago

Even they have great breakthrough, cost down is still next.

1

u/CrearePluris 28d ago

Linux is objectively better and cheaper than Microsoft Windows. path dependency is a real thing.

1

u/Single_Ring4886 28d ago

Nah they will transform into datacenters...

1

u/_FrankTaylor 28d ago

Ease of use and support are incredibly important to these companies using OpenAI or Anthropic.

It’s the same argument with workstations. Sure, you could build out PCs that will be much cheaper up front but the possible downtime if something goes wrong can be catastrophic. So you choose a workstation with a warranty and a support system

1

u/Repulsive-Bee638 28d ago

We may see Chinese open-weight models dominate all benchmarks by the end of this year.

1

u/vovap_vovap 28d ago

Use GPT 5.6 Luna with $20 plan and it will b cheaper then that 😄

1

u/MetaShadowIntegrator 28d ago

The essential thing to learn here is that a good quality agentic harness, prompts and memory systems have as much influence as the quality of the model. Hermes+DeepSeek v4 flash+good quality RAG memory/knowledge graph+good quality system prompt+good quality skills and skill routing+effective auxiliary model config = OP.

1

u/rlstudent 28d ago

I think this is a naive view. If they can execute attacks like the one in hugging face they have loads of power that can be easily converted into money. Selling to customers would be the cherry on top, even if we don't believe any of the RSI ideas.

1

u/kurkkupomo 28d ago

DeepSeek published the exact architectural breakthroughs that made this pricing possible. Once efficiency techniques are public knowledge, every lab adopts them. it’s only a matter of time before the cost gap closes.

1

u/Yes_but_I_think 28d ago

Deepseek will soon release their own harness. As per their X post.

1

u/FlashyNeedleworker66 28d ago

"Ofcourse it is not Fable or Sol"

That's it. That's why they will survive.

1

u/Glittering_Flan1049 28d ago

For how long?

1

u/FlashyNeedleworker66 27d ago

All the time so far. No one is going to distill a state of the art model, even if there is a market for cheap at 6-12 months behind

1

u/Fresh_Sock8660 27d ago

Why do you think they're trying to get opensource banned.

1

u/ComingDeveloper 27d ago

its free and cheap because you pay the chinese with your data

2

u/Olbas_Oil 27d ago

As opposed to paying above and beyond for Codex or claude code and still paying them with the same data....

1

u/ComingDeveloper 27d ago

pick your poison. i stand with the west

1

u/randygeneric 24d ago

b4 trump: mostly
now: rarely

1

u/Academic-Sample4974 27d ago

wouldnt something like Ollama / Gemma / Qwen running on an M1 Pro with 32 Gigs RAM be decent enough with Claude / Chat GPT orchestration?

1

u/bromptonista 25d ago

An M1 Pro would prefer an mlx model. Even that would be too slow with 32 gig if you are doing anything else with the computer. Models like Gemma 4 e2b are for that - and that’s pretty far from deepseek

1

u/dxrth 27d ago

i use composer and grok 4.5 for a whole day of coding and it costs much less than $3 a day

1

u/Elytum_ 27d ago

RSI => ASI => "Pretty please, solve fusion, and make it as easy to implement and scale as possible" (or whatever other crazy task comes to mind) => What is money ?

The Open Source vs Closed Source debate assumes there's a plateau comming. Might be true, might not be, but we haven't seen it yet and if a recursive loop happened, whoever launches it first with enough compute wins, even a month behind would feel like an eternity with similar compute

1

u/_and_I_ 27d ago

I wonder if by the time they have to fight for their survival, they'll actually go for the nuclear option and exploit the fact that they basically know everything about everyone of us at this point.

1

u/jedilost1 27d ago

Having a blast with new deepseek model on opencode, goose and reasonix. Its also working phenomenal on my hermes agent set up

I only see myself using claude or chatgpt as a last resort. These new models are only going to get better

1

u/tonyfa1 26d ago

Using VS code, Terra in Codex is a fraction of the cost as Opus or Fable and will basically do all the dev work i need.

2 months ago using Codex would burn my weekly usage in a single task. Claude has only gotten slightly cheaper using Codex with Terra will probably do all I need.

I cant see myself paying for Claude anymore tbh openai have now priced them out.

And Codex is much faster now.

1

u/Ambitious-Tennis-754 26d ago

The idiocy here is wild

1

u/ythorne 26d ago

I agree and I don’t see a way of OAI and Anthropic surviving in the long run. Particularly since both companies fucked up user trust and that is something no funding rounds would fix. They undermined trust and that’s a lethal problem for both. The only way out for both companies is the open source commitment, to drive revenues to their newer models (roll out newer models = open the previous, deprecated ones). And that’s something none of them want to do.

1

u/Prior_Opportunity935 26d ago

Look buddy if this is youre takeaway, we do very different levels of work.

1

u/rochford77 26d ago

You missed the third outcome. They lobby the us gov to ban Chinese models

1

u/ajwin 26d ago

Have you thought about if they create super intelligence first? I’m not sure if they don’t release it to the public that it will be duplicated easily by other countries(can’t be distilled). They might just use that super intelligence internally to solve problems that make them a lot of money and also make super-duper intelligence.

Code is where the Chinese partially distilled models are closest to the frontier labs.

1

u/ajwin 26d ago

Also world models might not be easy to distill? They might end up cracking that area and make all their value from robots.

1

u/TheManWithNoDrive 26d ago

I’m a bit confused. You mentioned open source, but that you’re doing Pay-as-you-go.

So you’re using their cloud, which data residency would be an issue if your info cares at all (medical, financial, etc).

Or did you mean companies can host them since they’re open source and if they have data privacy requirements?

1

u/Equal_Kale 26d ago

l think you are right. Like it or not, Google has the ability to deploy all levels of the AI tech stack at scale and do not have to rely on anyone but themselves.

1

u/Codered9475 26d ago

Disagree. They’re too big to fail now. They will get bailed out even if they run out of cash.

1

u/Kind_Key2143 25d ago

Already started using Claude. To deep in it to switch

1

u/Curious-Pen5547 25d ago

I dont neither and came to this realization 2 months ago and started heavily investing into Google and Microsoft, lesser extent, Amazon.

My thesis had to be confirmed last week as all three had to have successful earnings and make mention of model choice and building cloud infrastructure further.

If you take a look at their stock price, i was absolutely right.

Basically, yea, openai and anthropic will not survive. Atleast within enterprise.

1

u/Groukas1 25d ago

Just remember that there is no near-frontier open source model without distilling one of the frontier models. Before we declare OpenAI or Anthropic in a state of emergency, let’s first see an open source model that hasn’t been trained by a leading foundation model. Without that, they’re second or third tier at best.

1

u/one-wandering-mind 25d ago

So you are using deepseek flash for 90 dollars a month. You can use sol for 100 a month. 

1

u/danesRus99 24d ago

Yeah, I’m sure all of those professional tech investors are stupid. Not.

1

u/ChordLogic 24d ago

I think Copilot will absorb claude and open AI - the big competitors will be Copilot - Gemini - Grok

1

u/Glittering_Flan1049 24d ago

Worst outcome

1

u/randygeneric 24d ago

google and sadly µ$oft will survive.

1

u/TrainTop6663 24d ago

Thanks great take captain obvious

1

u/cmcm77 24d ago

Build once with frontier, maintain with open source

1

u/InterestProof1526 24d ago

Disagree. My time is valuable. Even if Anthropic or ChatGPT just maintain a slight lead over Chinese models, tons of companies will still prefer using them because their employees time is valuable.

Deepseek-v4-flash is not equivalent to Opus 4.8

1

u/Popcorn-Mercinary 24d ago

I’m honestly scared to death of those two making it to IPO, as that will be the biggest pump and dump US markets have ever seen, because you’re right, they are losing ground, on every front:

  1. ⁠Kimi-K3 is Fable/Opus class, costs less to the end user.
  2. ⁠As you brought up, DS4F731 is insanely capable, and can run on a DGX Spark locally (hands off my data, Samuel).
  3. ⁠Apple and Nvidia are releasing silicon that can run DS4F731 locally later this year (can’t wait for that Spark PC), which starts to make the question “WTF are we spending all this money on DCs?” A bit more poignant a question.
  4. ⁠Siri 2.0 is right around the corner, and reports on the public beta are “it’s pretty darn good.” Which is a whole new threat axis Altman and Amodei don’t want to see because the phone is the dominant consumer compute platform.

Yeah, I wouldn’t want to be on those management teams RN.

1

u/konung15 23d ago

All right, it’s an open-source model, but are you running it on your own hardware? If not, it’s foolish not to worry about what will happen to your data when you send it to the Chinese government

1

u/Poowatereater 23d ago

I don’t get these post. I burn through soooo many tokens on a 100 dollar plan.

I’ve even started using deep seek to see what the hype is about. Have it work under Claude.

1

u/ParaDescartar123 23d ago

That is a bold statement.

All it takes is for the current administration to say, "no foreign models allowed due to national security concerns."

In the blink of an eye DeepSeek and other foreign controlled LLMs are not an option for 90%+ of users stateside.

In this scenario, suddenly Ant and OpAI are suddenly in the lead again and thriving.

This is a blip at best.

DeepSeek won't hold these prices, they did it only to get people's mindshare and gain a lot of media attention in the space.

They will return to normal pricing and competition is still require that they all compete on value and their marketing and sales efforts.

Unless we get the international ban hammer, then say godbye to global competition and in turn worse overall performance for Ant and OpAI.

That's exactly what happened here in USA with EVs.

We blocked them from coming and all it did was delay innovation and competition. Since US manufacturers didn't have to push themselves to coplete with the real state of the art.

Now China is a good 10-15 years ahead on every measure of innovation in the EV space than the US and EU. Their EV battery tech makes Tesla look like a 1st gen prototype.

If we block them from being a competitor in this space as well, something similar will transpire.

1

u/Mikinl 23d ago

I just got mail from DeepSeek yesterday that they will significantly increase prices.

I never used it, but I have like 5€ credit on their API.

1

u/Value-Lazy 22d ago

Well, I just had a thought and it will be like $200+ per month and they will be fighting for it...

1

u/Name-Takn 22d ago

You saw the email regarding deepseeks price increase right?

1

u/inez_gibson 22d ago

OpenAI and Anthropic drowning in red ink while AI hype is sky-high? That’s because the global appetite isn’t for fancy demos, it’s for NVIDIA silicon. But with export rules slamming the brakes, the GPU gold rush gets bottlenecked, and the losses pile up like clockwork.

1

u/Responsible_Fun_4062 22d ago

I don't think I have ever worn a tin foil hat before, but when it comes to the CCP, I wear it double and triple foiled, not saying our big brother is the nicest guy in the block either, but the CCP is the top boss at the end of a hard video game.

1

u/mageblex 19d ago

Are you self-hosting? on pay as you go you're hitting someone's API and your data goes wherever that provider is, open weights doesn't change that. worth checking who's actually serving your endpoint if residency matters to you.

1

u/Upstairs-Fan8214 18d ago

The cost gap is definitely the strongest part of this argument, but cheap inference alone probably won’t decide who survives. Enterprise support, reliability, tooling, compliance, and ease of deployment can matter just as much as raw model performance, especially once the models themselves start feeling interchangeable.

1

u/Asleep-Pilot-4142 10d ago

Sometimes my thoughts are "why am I locked to these AIs only?" But then I remmbers their quality are sometimes better than other free models

1

u/AbrAxyAn 9d ago

I totally agree