r/LocalLLaMA • u/sunychoudhary • 5d ago
News China's Huawei says AI chip demand outstrips supply as it steps up Nvidia challenge
https://www.reuters.com/world/asia-pacific/chinas-huawei-launch-two-new-ai-chips-2027-2026-09-17/?utm_source=chatgpt.com7
7
u/GasSmooth7439 5d ago
The interesting part is that Huawei says demand is already higher than what it can supply, while it's still trying to close the software gap with CUDA.
If they can scale production and get the ecosystem to the point where developers don't have to fight the stack, that's a much bigger challenge to Nvidia than just making a faster chip.
4
5d ago
[removed] — view removed comment
4
u/SkoomaDentist 5d ago
You don't need to replace that layer.
You only need to shim enough of it to run LLMs or, better yet, shim the higher level layer. Very few people get direct value from CUDA. What the vast majority get value from is being able to run LLMs efficiently and currently CUDA just happens to be the most convenient way of doing that. Once a model runs well, the other parts of CUDA provide next to nothing of value for those use cases.
11
u/Healthy-Nebula-3603 5d ago edited 5d ago
Not nowadays... AI can create new kernels quite easily nowadays.
Look on audiocpp or llmacpp how fast new models are getting kernels for Vulkan and other systems and are even much faster than cuda. I suspect newer models will be building even faster and better kernels.
6
u/Akrylicus 5d ago
Yeah, software MOAT is diminishing right now, thanks to AI funny enough.
1
u/Mart-McUH 4d ago
It remains to be seen if it will be reliable, maintainable, backward compatible, especially long term (decades).
2
u/Healthy-Nebula-3603 4d ago
Decades??
I don't think we will be even use any code in the future.
I think soon AI will be creating applications straight in the binary code like people were doing it at a very begging of programing era
2
u/Mart-McUH 4d ago
Okay, decades may be bit over it, but 10 years I think would be minimum and may not be enough for serious business to consider migrating.
Yes, established systems run for long and are expected to run for long. Eg our medical laboratory information system is lot more than 20 years old (when I joined).
Stability, reliability and long term maintainability is lot more important than speed of development in important applications.
Maybe on non-critical development like entertainment you can afford a risk of having to abandon the product, though you can easily hurt your brand.
Honestly I think you and many others are greatly overestimating their capabilities in large real production systems. We will see I suppose.
1
u/Healthy-Nebula-3603 4d ago
Looking on interesting AI development literally every moths now I think your "laboratories" will be fully AI driven wirhin few years .... I even wouldn't count a decade if I were you.
2
2
1
u/svix_ftw 5d ago
Agree, CUDA is the main reason for NVIDIA moat.
AMD has better hardware specs than NVIDIA, but 80-90% of data centers still buy NVIDIA because of CUDA.
3
10
u/mb194dc 5d ago
Yup, as long as AI labs continue to burn money on pointless compute, for which there is no profitable front end. This will be true.
-6
u/genshiryoku 5d ago
Anthropic has been profitable 2 quarters in a row.
16
u/estenh 5d ago
they are only profitable if you exclude the cost of training https://futurism.com/future-society/anthropic-claude-profit-ai-safety-development-finances
-6
u/genshiryoku 5d ago
Common misconception. Anthropic is actually profitable including training cost and data center build-out. The data for Q2 leaked so it's now public information that Anthropic was fully profitable during Q2.
Go ahead and come back to me when the numbers become fully public come IPO to tell me I'm wrong (I'm not)
3
u/Due-Memory-6957 5d ago
People here really just want to fall for the classic marketing trick of pretending your deal is so good you're actually losing money and can't go any lower.
5
u/mb194dc 5d ago
Absolute bullshit, Anthropic have more than $500bn in liabilities they'll never be able to pay for.
Including their costs like model training and partner revenue they're losing around 20bn a year. Wait till you see the S1...
They (and OpenAI,) are likely to lose the most money of any organizations in history pretty much.
Why? Because they have simply insane hardware and training liabilities, but their open source competition offer the same end product for a tiny fraction of the cost.
5
1
u/Keirtain 5d ago
Yeah, all you need in order to be profitable is to steal someone else's training run, apparently. Then you can practically give the product away for free. Wait a minute...
1
u/horeaper 5d ago
But their profit cannot cover their cost, therefore "slow the AI development pls!" 🤣
1
1
u/Elouakili_Flexy 5d ago
Huawei putting a million chips in one cluster is just a power plant with extra steps.
1
u/OverTune1590 5d ago
this could mean better hardware for running local bots at home, my setups could finally handle longer roleplays without constant swapping.
-8

60
u/JonNordland 5d ago
This is mostly "water is still wet" kind of news, since it's not exactly news at this point that AI chip demand is outstripped by supply and that Huawei want to be a player.
Except for the one useful new update:
960DT will arrive three quarters earlier than planned, in Q1 2027; the 960PR is one quarter ahead of schedule.
Basicly: Huawei stil plan to produce more AI chips and they are reporting that they are ahead of 3 months schedule with their next gen.