| ▲ | mchusma an hour ago | |
Coding, agentic flows like logging into my accounts and gathering data, grok bot. 4.6 made more mistakes than SOL or Opus overall. Gave up a lot. And in my opinion, the rate of mistakes is kind of more important than how brilliant it is. I think 4.7 may still be better, but I was hoping for clearly Sol/Opus level and so far it just isn't there for me. | ||
| ▲ | the_sleaze_ 20 minutes ago | parent [-] | |
I didn't find 4.6 any better than Composer 2.5 - which remains incredible and honestly nothing else compares for me. Make a galaxy model search the space, create a document, argue and defend decisions, then hand it to 2.5 to implement. 4.6 was a slower less enjoyable version of that. 4.7 is better at "I want the button to cancel the jobs, dont make any mistakes" but honestly that's not what I use it's class for. | ||