| ▲ | yorwba 2 days ago | |||||||||||||||||||||||||
AI companies legally acquiring books have indeed been in the news: https://news.ycombinator.com/item?id=49330742 And where are you getting the idea that Mistral doesn't train on copyrighted data? There's not a lot of code written by people who've been dead for more than 70 years, but somehow Mistral has been able to release coding models anyway. | ||||||||||||||||||||||||||
| ▲ | spwa4 2 days ago | parent [-] | |||||||||||||||||||||||||
But they have been training on copyrighted data since GPT-2 at least. 2019, and that's when it came out, so before that of course. | ||||||||||||||||||||||||||
| ||||||||||||||||||||||||||