| ▲ | OliveronData 4 hours ago | |||||||
> ... the pace of progress has continued on it's exponential trajectory since ChatGPT first came to the public's attention. Did it? Model wise? I would understand agents wise, sure. But model wise? The attention to detail from the model? The ability to recall minute things? Improvements are there, yes, but mostly on Fable and Astra. Opus still isn't as attentive as Fable in long term writing for example. Sure, Opus 5.5 benchmarks better than Fable. Sure. But is that the model, or is that the RL for agentic work? From where I'm standing, the model work has not been exponential at all, and more and more it looks like the latest and greatest is getting too expensive too fast. Both 5.5 and 5.6 chat models got nerfed, actually nerfed not the tea leaves kind. In mid 5.5 cycle the chat model lost the ability to substitute names if given an outline. 5.6 cycle the chat model lost the ability to use paragraphs after a few hundred words (coinciding with Chat/Work split). There's a race from OpenAI to serve dumber models on chat. I'm not even sure who they are racing against, but the fact that Astra, Sol 6.0, and now Sol 6.1 not being available for chat, should tell you that those models are expensive, and not the kind of models that can be freely "chatted" with on a subscription. OpenAI much prefers you use Work and limit the chat usage, much like Grok and Claude. I'm guessing they will announce that later during the dev days. That could be cost cutting too, true, but really? That's the only explanation? And nothing else? Sure, the progress did not stop. But it is nowhere near close being exponential when it comes to LLMs themselves. Agents are separate. | ||||||||
| ▲ | luma 4 hours ago | parent [-] | |||||||
I didn't use the word LLM. I'm talking AI capability, you're focused on this or that current approach to AI. I think it's fair to assume that the approach will change as new ideas are learned, new and more hardware will be purchased and applied to the problem, and then capabilities will (for now) continue on their exponential curve, same as it has gone for the past several years. These things are knocking down Millennium Prize problems while a substantial subset of commenters here are still thinking about stochastic parrots. | ||||||||
| ||||||||