| ▲ | wasfgwp 2 days ago | |||||||
Well there were observed cases of Claude calling itself Deepseek or Qwen. So pot calling the kettle black? To be fair I find it hard to take this too seriously, shouldn’t it be trivial to just replace “Claude” with any other string in your “distillation” dataset? | ||||||||
| ▲ | throw10920 2 days ago | parent [-] | |||||||
> Well there were observed cases of Claude calling itself Deepseek or Qwen. So pot calling the kettle black? There are open-source Deepseek and Qwen models - "distilling" doesn't involve breaking terms of service or hitting an API because you can literally run local inference or even just inspect the weights directly, and that's intended because they're open source. It's categorically different for a nation-state to build massive illicit networks of fraudulent identities to do distillation over tens of thousands of accounts to intentionally bypass providers' terms of service, intention for their models, and business model that very explicitly proprietary and not open source. https://www.chinatalk.media/p/how-to-buy-cheap-claude-tokens... If Claude did distill on proprietary PRC LLMs - then fine, shame on them - I condemn that and I expect others to do the same. But there are no open-source Claude models. The only way for PRC models to have those responses is if they distilled Anthropic's models from their APIs. > To be fair I find it hard to take this too seriously, shouldn’t it be trivial to just replace “Claude” with any other string in your “distillation” dataset? ...and what would happen when it read all of the books and articles about Anthropic and replaced "replaced Claude Opus" with "replaced Qwen Opus"? Did you give any thought to this at all before saying it? | ||||||||
| ||||||||