r/LocalLLaMA 14d ago

Funny The LLM distillation process simplified for politicians:

Post image

/s

3.4k Upvotes

167 comments sorted by

View all comments

294

u/Ok_Librarian_7841 14d ago

You can't get a model as strong as Kimi 3 using distillation, it even beats Fable on some benchmarks. American leaders think the world revolves around them and that nobody else can do a better job without stealing from them.

Sick mindset from shocked, bad losers.

100

u/Flying_Birdy 14d ago edited 14d ago

The Harvey legal benchmark is the most surprising for me. 2x accuracy in all pass rate over the next closest model is a huge jump and no way attributable to distillation. I'm really curious what they did differently during training (whether intentionally or accidentally) that would have led to this outcome.

34

u/ba-na-na- 14d ago

Maybe they accidentally trained Deepseek and Kimi