| ▲ | HarHarVeryFunny 2 hours ago | |
> In the end, shaming people for writing that gets flagged as AI can lead people to sidestep structures the model has learned from us It's interesting why LLMs generate constructions like this more frequently than they presumably exist in the training set. I wonder if this is some sort of mode collapse caused by post training, and/or maybe because they are training on synthetic data so these things become self-perpetuating and self-amplifying (a feedback loop)? The lesson for humans worried about being falsely identified as AI is just learn to write better! It doesn't matter where your repertoire of phrasing comes from (copying AI or not), but one of the basic rules of writing is not to repeat yourself unless you are doing so deliberately for a purpose. Go ahead and use "It's not just X. It's Y" if you want to, but if you use it multiple times in the same short piece of writing, then you may deserve to be called out for poor style, if not for being an AI. | ||
| ▲ | Maxatar 2 hours ago | parent [-] | |
Its not model collapse nor does it have anything to do with training data frequency. It's simply RLHF where the humans hired to tune the conversational style of these LLMs preferred certain idioms over others and so the reward function for these LLMs gravitated toward using them. If LLMs generated text based on training data frequency they'd likely be some of the most vulgar and hostile things ever created. The internet is full of insults, profanity, and low effort content. The repeated phrases are a side effect of reward optimization rather than some kind of model collapse. | ||