Remix.run Logo
benjiro29 2 hours ago

Also the cost per task.

https://www.vals.ai/benchmarks/vals_index

!!! Vals !!!

Vals Index Opus 4.8 > 5.0 goes from $2.90 to $8.54, for 4% gain ... That is a massive cost increase. Sure, 20% cheaper then Fable, but that is a 3x price increase compared to Opus 4.8 in that test.

https://artificialanalysis.ai/models/claude-opus-5 https://artificialanalysis.ai/models/claude-opus-5#price-cos...

!!! artificial analysis !!

Cost per task is second highest, right below Fable.

* Fable: $2.75

* Opus 5.0: $2.03

* Opus 4.8: $1.80

* GPT 5.6 Sol: $1.04

* Kimi K3: $0.95

Looks like interest levels of cherry picked cost in their report. Cheaper model, clearly NOT. More expensive in both benchmarks.

spider-mario an hour ago | parent [-]

Your numbers are for “max”. Opus 5.0 “max” is $2.03. Opus 5.0 “high” (competitive with Claude 4.8 “max” on that index) is $1.06, less than the $1.80 you are quoting for 4.8 max.

That the most expensive variant is expensive doesn’t really tell us much.

benjiro29 34 minutes ago | parent [-]

Same answer i gave to somebody else up here...

If you start to drop effort levels, you need to compare to the competition models. So GPT models on the same ~intelligence level, are then 50% cheaper.

You see the issue? Its still a expensive model, and from my understanding, it still uses the old tokenizer.

Going to be interesting to see when GPT 6 comes out (very soon).