Remix.run Logo
hashstring 2 days ago

We also have evidence that Anthropic distilled all human info they could get their hands on for the development of all their models.

Distillation should be fair game given the (current) game of LLM training. Yes, as a model creator you probably want to protect against it, but it does make you a hypocrite.

> The developer OpenAI has said it would be impossible to create tools like its groundbreaking chatbot ChatGPT without access to copyrighted material, as pressure grows on artificial intelligence firms over the content used to train their products.

kevincox 2 days ago | parent [-]

If anything this distillation is more ethical than the original as it is on machine-generated content rather than copyrighted human-labour products.