Remix.run Logo
dandiep 7 hours ago

Why are LLMs producing this super tense text more often now? Is it because they are being optimized to use fewer tokens?

vanviegen 7 hours ago | parent | next [-]

I think they're overcorrecting for being too long-winded in previous generations (and still, in some cases). I guess this is a hard balance to get right.

Davidzheng 7 hours ago | parent | prev | next [-]

I'm pretty sure the other answers are wrong and it's a side effect of RL (see thinking machines post about inkling training). It's also exacerbated in fable and sol--I think it's token efficiency effect--bc it's about to reason with fewer tokens the density of the token information goes up.

conception 7 hours ago | parent | prev | next [-]

I figure they are being optimized to write code/functions and not prose so all text is getting more code like.

zahlman 7 hours ago | parent | prev [-]

ChatGPT is still plenty verbose by default IMX.