| ▲ | bermudi 5 hours ago | |||||||
I honestly can't believe serious people are making this argument on a straight face. Gemini 3.7 flash outputs so many tokens per answer it doesn't matter how fast its TPS is, sol will end up being both cheaper and faster than Gemini. So ppl are paying more for a given task, waiting longer and using a dumber intelligence because "TPS number shiny". Gemini 3.8 outputs 11k more tokens PER TASK on average in AAII than 3.7 putting it dead last in output tokens per task in the leaderboard. | ||||||||
| ▲ | gundmc 4 hours ago | parent | next [-] | |||||||
There are numerous benchmarks that measure cost per task, which factors out tokens entirely. Gemini 3.8 flash is significantly lower than Sol on basically all of them https://artificialanalysis.ai/#cost-tabs That said, Luna is the undisputed king here at the moment and is what I use as my workhorse model. | ||||||||
| ||||||||
| ▲ | WarmWash 4 hours ago | parent | prev [-] | |||||||
AA isn't the only benchmark | ||||||||