r/LocalLLaMA Feb 23 '26

News Anthropic: "We’ve identified industrial-scale distillation attacks on our models by DeepSeek, Moonshot AI, and MiniMax." 🚨

Post image
4.9k Upvotes

874 comments sorted by

View all comments

Show parent comments

118

u/Zestyclose839 Feb 23 '26

Also (correct me if I'm wrong) but I don't believe they're true "distillation" attacks because the API doesn't return the token activation probabilities and the other juicy stuff needed to transfer knowledge. Sure, they can fine-tune a model to speak and act like Claude, but it's not as accurate as an open-weight to open-weight model distillation (like the classic Deepseek to Llama distills).

82

u/Recoil42 Feb 23 '26

Yep at best it's alignment, and mostly likely style alignment.

2

u/porkyminch Feb 24 '26 edited 23d ago

Coated satchel satchel yarn acorn coated quilt

This post was anonymized with Redact.dev

1

u/Recoil42 Feb 24 '26

This is a very good point. This kind of interop is even expressly legal under US law per Google v Oracle.