| ▲ | wkcheng 2 hours ago | |||||||||||||||||||||||||||||||||||||
The cost / performance chart shows that in almost all configurations, it looks worse than Opus. Why would you use Sonnet 5.5 on xhigh if you would get better results (higher score, cheaper cost) on Opus 5.5 high? Is there a good use case? This isn't like Luna where it's much cheaper/effective just to use Luna in certain situations. | ||||||||||||||||||||||||||||||||||||||
| ▲ | usaar333 an hour ago | parent | next [-] | |||||||||||||||||||||||||||||||||||||
Per the charts, there is largely no point to using Sonnet 5.5 at high+ as opus low generally will give similar performance at similar or lower cost. But Sonnet 5.5 at medium and below gives you a cheaper option at a performance worse than the lowest thinking Opus (low), which may be viable for "low intelligence" use cases. | ||||||||||||||||||||||||||||||||||||||
| ▲ | ricardobeat 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
At low and medium effort it is 1/3 cheaper, at high it’s a step above Opus/low. It only looks worse at xhigh. | ||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||
| ▲ | delillos 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
There's a sort of magical thinking needed to answer a question like that. You might say it comes down to "feel" of the model; i.e., the indefinable differences in the way that they speak to the user and approach problem solving. Perhaps Opus is suited for tasks that tackle new ground, while Sonnet might be better at tasks that are more grounded in the code. Ultimately it's slightly ridiculous to define model capability on a single axis. It's like a standardized test. Sure, you can line people up by their ACT score, but that doesn't mean a doctor and a brilliant artist who both do well on the ACT have an identical intelligence or approach to life. It just can't be captured. | ||||||||||||||||||||||||||||||||||||||
| ▲ | RussianCow 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
It appears, at least from a quick look, to be noticeably faster than Opus. If true, and you don't need xhigh/max reasoning for your use case (like a well-defined set of code changes), Sonnet might get the job done much more quickly. With that said, at that point, I'd probably use something like DeepSeek V4.1 Flash, which is way faster and significantly cheaper, and probably not noticeably dumber for most use cases. | ||||||||||||||||||||||||||||||||||||||
| ▲ | Jcampuzano2 an hour ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
I'm honestly not sure where they're getting their 30% numbers from at all. In every single chart that they chose to display except for one, it costs similar or more than Sonnet 5, while also being comparable in price to Opus. Maybe it's buried within their system card but I think that this would be one of the first things they'd want to show in the announcement article and they fail to do so. I really don't know who does Anthropic's marketing but they always seem to a pretty terrible job in their announcements from my perspective. | ||||||||||||||||||||||||||||||||||||||
| ▲ | SubiculumCode 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
t/s maybe? IDK, because their token speed comparison was against Sonnet 5. | ||||||||||||||||||||||||||||||||||||||
| ▲ | solenoid0937 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
It literally does not? | ||||||||||||||||||||||||||||||||||||||
| ▲ | dominotw an hour ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
just shows you how little control of output these labs actually have. They are training two models that kind of ended being the same so whatever they were doing specifically didnt make much difference. | ||||||||||||||||||||||||||||||||||||||
| ▲ | quatotor 2 hours ago | parent | prev [-] | |||||||||||||||||||||||||||||||||||||
[dead] | ||||||||||||||||||||||||||||||||||||||