| ▲ | Do newer coding models end up training on the AI slop generated by older models? | |||||||||||||||||||||||||
| 6 points by tbharath 7 hours ago | 6 comments | ||||||||||||||||||||||||||
As old coding models were writing and pushing code to public, does the new coding models train on them because generated code by old coding models were not so good? | ||||||||||||||||||||||||||
| ▲ | OsrsNeedsf2P 7 hours ago | parent | next [-] | |||||||||||||||||||||||||
Yes, that's why good datasets are important. But usually the difference between model generations is architectural improvements | ||||||||||||||||||||||||||
| ||||||||||||||||||||||||||
| ▲ | verdverm 7 hours ago | parent | prev [-] | |||||||||||||||||||||||||
How much do we assume human written code for training the original models is free of human made slop? We all know we all take shortcuts and have code we are not proud of. If they are I deed averaging machines, is it possible their output is the average of human output? I'd argue that the way attention works plays into the slop too | ||||||||||||||||||||||||||
| ||||||||||||||||||||||||||