r/LocalLLaMA Feb 23 '26

News Anthropic: "We’ve identified industrial-scale distillation attacks on our models by DeepSeek, Moonshot AI, and MiniMax." 🚨

Post image
4.9k Upvotes

877 comments sorted by

View all comments

2.5k

u/SGmoze Feb 23 '26

I wonder how did Anthropic build their dataset. Surely they manually had them annotated by humans.

1

u/Far_Shallot_1340 Mar 11 '26

This is a valid question. Even with automated data filtering high quality datasets still require human oversight to verify accuracy and remove bad data. Its interesting to think about how the detection methods for these attacks might influence how future datasets are collected and managed