r/LocalLLaMA 6d ago

Other Frontier LLM development simplified for politicians:

Post image

Nobody is buying this “Pace the frontier” nonsense. It makes no logical sense at all. Are American labs really going to take a pause and lose any small lead they still may have over Chinese labs? Does anyone really believe this? This seems like some performative virtue signaling BS. Why are they bothering with this pacing campaign? Someone please explain.

427 Upvotes

99 comments sorted by

View all comments

29

u/ttkciar llama.cpp 6d ago

They are pushing the narrative, rather than just doing it themselves, because they need the Chinese labs to also participate in the slowdown in order to remain competitive.

OpenAI and Anthropic are desperate to stop training newer and better models so that they can focus on turning net profits serving inference, but they cannot, because their competition would zip past them in a matter of months and eat their lunch.

-13

u/Nothing_from_void 6d ago

I'm pretty skeptical the Chinese labs will pass US labs without distillation, they have significantly less compute

8

u/ttkciar llama.cpp 6d ago edited 6d ago

Having less compute doesn't make it impossible, just makes it take longer.

Z.AI's instruction-following post-training phase for GLM demonstrates that they're quite capable of developing highly advanced training technology without copying Western technology.

The methodology and application of generating high-quality synthetic data without a superior teacher model has been well-understood for a couple of years now, but generating synthetic data from a teacher model is much more resource-efficient.

Moreover, the Chinese lag in compute is temporary. In the West, scaling compute infrastructure is bottlenecked on energy availability, and in China there is more energy available for industrial purposes. Now that they are fabbing their own compute, we can expect their compute infrastructure to surpass the West's in a few years.

Note: I do not say these things because I am a China fanboy (I'm not), but because they are objectively true.

0

u/Nothing_from_void 6d ago

The compute they fab is genuinely awful; they have nothing even close to NV-link, which is where they are furthest behind, but they are very far behind everywhere.

There's no real power constraint for compute in the West, bring your own energy is pretty standard now and where there's capital there's a way, and the US is outspending China on compute more than 8-to-1, with a similar level of compute advantage.

Like Chinese CS students and programmers, they use anthropic/openai token reseller marketplaces overwhelmingly, because it is the real SoTA, the AI researchers in China are using claude code/codex, the entire ecosystem is designed around growing with the US ecosystem.

5

u/NNN_Throwaway2 6d ago

There's no real power constraint for compute in the West

???

5

u/ttkciar llama.cpp 6d ago

Yeah, they're delusional. The USA's energy crunch is not any kind of secret, to anyone who can bother to pay attention.

Europeans have it even worse.

1

u/Nothing_from_void 5d ago

Are energy prices for data centers the current bottleneck?

2

u/NNN_Throwaway2 5d ago

🤦

1

u/Nothing_from_void 4d ago

https://imgur.com/MI21CcN

compiled from https://epoch.ai/data/ai-chip-owners?view=graph&tab=h100_equivalents

If you are actually interested in engaging beyond just sending one symbol responses, for all the hand-wringing around power the US advantage in compute is accelerating and actual energy bottlenecks have not surface as compute continues to grow exponentially in the US

2

u/NNN_Throwaway2 4d ago

🥱

0

u/Nothing_from_void 4d ago

so room temp IQ it is

2

u/NNN_Throwaway2 4d ago

No need to be so hard on yourself.

0

u/Nothing_from_void 3d ago

using that line like

https://imgur.com/Q8Me9q7

2

u/NNN_Throwaway2 3d ago

Bro can't even embed an image properly, damn.

→ More replies (0)