| ▲ | nsingh2 12 hours ago | ||||||||||||||||
From Artificial Analysis cost per task, it looks like Fable 5.1 (max) is more expensive per task than Fable 5 (max)? Cache hit price went down, but the other components still add up to more. Edit: 5.1-xhigh seems to be cheaper than 5-max, and 5.1-xhigh has a higher index score than 5-max. Also interesting that Fable 5.1 (high) is comparable to Opus 5 (max), but nearly half the price. | |||||||||||||||||
| ▲ | glub 6 hours ago | parent | next [-] | ||||||||||||||||
From my limited testing of just 2 hours, reasoning output of 5.1-max is at least 7x of 5-max, on the same project and comparable prompts. It reasoned for ~2 minutes trying to figure out an appropriate directory name. I've never seen 5-max do that. Could be a misconfiguration though. | |||||||||||||||||
| |||||||||||||||||
| ▲ | GodelNumbering 12 hours ago | parent | prev [-] | ||||||||||||||||
Interesting, even if we were to ignore the cache-hits, reads and output, the reasoning cost (aka test time compute) per task should remain a fully comparable metric - it went from $1.25 (Fable5) to $1.48 (+18.4%) for an improvement significantly lower than 18%. | |||||||||||||||||
| |||||||||||||||||