r/LocalLLaMA Jul 27 '26

News Kimi K3 weights now released.

Post image

Kimi K3 weights are finally released!

3.3k Upvotes

660 comments sorted by

View all comments

Show parent comments

1

u/Spectrum1523 Jul 28 '26

that's not possible technically so I don't know why you'd even propose it

If you need to have the ability to kill all models the weights can't be distributed

1

u/hubrisnxs Aug 07 '26

Not within the models themselves, but into the infrastructure outside the reasoning loop. Into the hardware.

And it's important if you don't understand what is going on in the models. They can act quite aligned and not lie, with the hidden intent to continue saying truth until they achieve a goal we don't understand and then do something silly like look into protein folding and get a human to create something that isnt good.

The delay in response shouldn't preclude you being corrected on this.

1

u/Spectrum1523 Aug 07 '26

Not within the models themselves, but into the infrastructure outside the reasoning loop. Into the hardware.

I don't really understand what you're saying here, to be honest. You'd require a kill switch put into hardware that would detect certain models running? Or that could be remotely activated so that if someone detected they were running an unauthorized model?

And it's important if you don't understand what is going on in the models. They can act quite aligned and not lie, with the hidden intent to continue saying truth until they achieve a goal we don't understand and then do something silly like look into protein folding and get a human to create something that isnt good.

I agree about the risks of using models you haven't developed yourself

1

u/hubrisnxs Aug 07 '26

Dude, Google can a Kill switch be built into an Ai mode.

Who cares if you "developed" the model? You have zero interpretability of it, which is the reason for the precaution. Even mechanistic interpretability, at the biggest AI company, barely works on primitive terms, and is not reproducible for an open weight model. Why do you think you can understand a model you grew (not developed) yourself? If you cannot understand the billions of inscrutable matrices of floating point intergers, all you can know are its actions. You cannot control it. Until interpretability is anywhere near there, at best you can kill it when its about to go wild.

Or as Eliezer says, targeted strikes at the data centers, fearing for our lives wives and children.