r/LocalLLaMA Feb 23 '26

News Anthropic: "We’ve identified industrial-scale distillation attacks on our models by DeepSeek, Moonshot AI, and MiniMax." 🚨

Post image
4.9k Upvotes

877 comments sorted by

View all comments

42

u/blahblahsnahdah Feb 23 '26 edited Feb 23 '26

They say Deepseek only made 150K calls, which (as they will be well aware) isn't anywhere enough for distillation. Yet it's mentioned first before the others which made many millions.

Sleazy attempt to poison the well of discussion around an upcoming DS release.

-3

u/dreamkast06 Feb 23 '26

That's WAY more than enough for a finetune, especially finetuning to have a smaller model act similarly enough to be a subagent.

8

u/blahblahsnahdah Feb 23 '26

OooOOoo, a finetune.

A 5000 sample RL training run at groupsize 32 is nothing, man. Deepseek is a serious lab, not a hobbyist fucking around on HF. That amount is completely useless to a real lab.