| ▲ | jstummbillig 5 hours ago | |||||||
Eh. What? Is this common sentiment? I mean Opus 5.5 is absolutely fantastic, unreasonably and unexpectedly so, but Astra was great and as far as I can tell SOTA until, when was it, 3 days ago, no? (Sol 6 idk, have not used it much for coding really. Seemed to work just fine when Astra used it in Codex as subagents.) | ||||||||
| ▲ | phoghed 5 hours ago | parent | next [-] | |||||||
In my experience, no. There’s no way to know though. The whole conversation and industry are a combo of benchmaxing, faith, and mysticism. Since like last December I haven’t had any issues getting work done with whatever the latest Anthropic or OpenAI models at the time were. Tooling and models have only gotten better since then. | ||||||||
| ▲ | copperx 5 hours ago | parent | prev | next [-] | |||||||
Opus 5.5 is so good that I don't want it to be replaced anytime soon. Stop training models, Anthropic, and just serve this thing without regressions for a year or three, can you? | ||||||||
| ||||||||
| ▲ | Eridrus 5 hours ago | parent | prev | next [-] | |||||||
Sol 6 definitely feels kind of dumb and worse than 5.6 Astra seems better though. Showing one potentially saturated benchmark doesn't necessarily fill me with a lot of confidence in the coding results. | ||||||||
| ▲ | nicce 5 hours ago | parent | prev | next [-] | |||||||
When GPT 6 Sol & Luna were released, everything went down. I have been running Sol at max thinking and it is about the same as old Luna with max thinking, give or take. Sometimes feeling even dumber. I can't trust it to do anything big alone anymore without babysitting. | ||||||||
| ▲ | the_duke 5 hours ago | parent | prev [-] | |||||||
On r/codex the sentiment seems to be quite wide-spread. | ||||||||