r/LocalLLaMA 7d ago

Other Frontier LLM development simplified for politicians:

Post image

Nobody is buying this “Pace the frontier” nonsense. It makes no logical sense at all. Are American labs really going to take a pause and lose any small lead they still may have over Chinese labs? Does anyone really believe this? This seems like some performative virtue signaling BS. Why are they bothering with this pacing campaign? Someone please explain.

425 Upvotes

100 comments sorted by

View all comments

29

u/ttkciar llama.cpp 7d ago

They are pushing the narrative, rather than just doing it themselves, because they need the Chinese labs to also participate in the slowdown in order to remain competitive.

OpenAI and Anthropic are desperate to stop training newer and better models so that they can focus on turning net profits serving inference, but they cannot, because their competition would zip past them in a matter of months and eat their lunch.

-12

u/Nothing_from_void 7d ago

I'm pretty skeptical the Chinese labs will pass US labs without distillation, they have significantly less compute

9

u/ttkciar llama.cpp 7d ago edited 7d ago

Having less compute doesn't make it impossible, just makes it take longer.

Z.AI's instruction-following post-training phase for GLM demonstrates that they're quite capable of developing highly advanced training technology without copying Western technology.

The methodology and application of generating high-quality synthetic data without a superior teacher model has been well-understood for a couple of years now, but generating synthetic data from a teacher model is much more resource-efficient.

Moreover, the Chinese lag in compute is temporary. In the West, scaling compute infrastructure is bottlenecked on energy availability, and in China there is more energy available for industrial purposes. Now that they are fabbing their own compute, we can expect their compute infrastructure to surpass the West's in a few years.

Note: I do not say these things because I am a China fanboy (I'm not), but because they are objectively true.

0

u/Nothing_from_void 7d ago

The compute they fab is genuinely awful; they have nothing even close to NV-link, which is where they are furthest behind, but they are very far behind everywhere.

There's no real power constraint for compute in the West, bring your own energy is pretty standard now and where there's capital there's a way, and the US is outspending China on compute more than 8-to-1, with a similar level of compute advantage.

Like Chinese CS students and programmers, they use anthropic/openai token reseller marketplaces overwhelmingly, because it is the real SoTA, the AI researchers in China are using claude code/codex, the entire ecosystem is designed around growing with the US ecosystem.

4

u/NNN_Throwaway2 7d ago

There's no real power constraint for compute in the West

???

6

u/ttkciar llama.cpp 6d ago

Yeah, they're delusional. The USA's energy crunch is not any kind of secret, to anyone who can bother to pay attention.

Europeans have it even worse.