| ▲ | bluegatty 3 days ago | |||||||||||||||||||||||||||||||||||||||||||
Distillation is absolutely - and uncontroversially - a valid term for what is happening here. This isn't really a debate, I'm not making a fine point - just check with all of the various defintions of the term. Moreover - the 'reasoning traces' are not required for distillation at all. Finally - it's entirely possible for them to have used Fable for later stage fine tuning. It's fair to be skeptical of Anthropic (and everyone else) - but this is 'distilling'. | ||||||||||||||||||||||||||||||||||||||||||||
| ▲ | HarHarVeryFunny 2 days ago | parent | next [-] | |||||||||||||||||||||||||||||||||||||||||||
Words have meaning - you cant just redefine them because you want to. Are reasoning traces required for distillation? Well they are if what you are trying to distill is reasoning, such as coding expertise. Do you need reasoning traces for "LLM as judge"? No, but it would be highly perverse to call that distillation when there is a more accurate name for it - LLM as judge. If you want to call use of Anthropic's redacted model outputs in any fashion that violates their terms of service (using them them to help develop anything that competes with Anthropic) as "distillation" then I can't stop you, but it reduces their claims to a joke. Finally, as noted, OpenAI (who are just as anti-Chinese as Anthropic) said that Kimi 3 can't be explained via distillation (even true distillation!!), or even "anything like it". But random internet guy, you, disagrees. OK. | ||||||||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||||||||
| ▲ | throw10920 a day ago | parent | prev [-] | |||||||||||||||||||||||||||||||||||||||||||
The GP (HarHarVeryFunny) is not operating in good faith. They've repeatedly lied about my own words to me, and are making up definitions that people who actually work at a frontier lab would disagree with. | ||||||||||||||||||||||||||||||||||||||||||||