Remix.run Logo
d_tr 4 hours ago

How and why do they get nerfed? To save money?

johnfn 3 hours ago | parent | next [-]

Models do not get nerfed. There has never been evidence of this. This would be trivial to prove if it were true, and such a proof would be a huge story and scandal to a news market hungry for a shred of a signal on AI's downfall.

This is the "your iPhone is listening to you and serving ads based on what you say" of the 2020s.

4chandaily 2 hours ago | parent | next [-]

> This is the "your iPhone is listening to you and serving ads based on what you say" of the 2020s

Of course, it turned out that this wasn't actually completely BS. We just were accusing the wrong vendor. Not disagreeing with you on models.

johnfn 11 minutes ago | parent [-]

I'd be happy to read a source as I am fairly confident that audio transcription -> facebook (or other) ads has never been true.

esafak 2 hours ago | parent | prev [-]

Yes, they can, through quantization. Many providers of open source models openly serve quantized versions; check openrouter.

ricardobeat an hour ago | parent | prev | next [-]

One theory is that they start serving at full precision, then quantize to save on costs as adoption grows. It's kind of a conspiracy theory atm, but I have definitely felt it - I had a large project done on Opus 4.8 release day, a week later it was struggling to complete partial tasks in the same area.

varispeed 3 hours ago | parent | prev [-]

Yes. They save on compute and customer has to use more tokens to achieve their goal which means more profit.