| ▲ | gundmc 4 hours ago | |
There are numerous benchmarks that measure cost per task, which factors out tokens entirely. Gemini 3.8 flash is significantly lower than Sol on basically all of them https://artificialanalysis.ai/#cost-tabs That said, Luna is the undisputed king here at the moment and is what I use as my workhorse model. | ||
| ▲ | NicoJuicy 2 hours ago | parent [-] | |
It's so funny how many people diverge on the same model. Ps. For the last week I diverged to Luna too, still need to check 3.8 flash. But 3.6 flash was my go-to model 3 weeks ago and before it was deepseek flash/pro for a while. None of the claude models seemed cost effective though. | ||