Remix.run Logo
dumberquestions 7 hours ago

Token price doesn't tell you much without knowing token efficiency.

user43928 6 hours ago | parent | next [-]

Their leading benchmark with cost per task shows a tough sell compared to Fable 5.1 Low and doesn't reach the performance of Fable 5.1 Medium.

How representative that is of real world usage, I don't know.

In their benchmark GPT 5.6 Sol performs suspiciously poorly compared to the former models.

7 hours ago | parent | prev [-]
[deleted]