Remix.run Logo
▲ system2 6 hours ago

Who in their right mind would use haiku while Mimo or GLM cost 10% of what they are charging with much smarter models?

▲nharada 2 hours ago | parent | next [-]

Isn't the point of this release that it's comparable?

AAI Index // Input // Output

Haiku 5.5: 43 // $0.10 // $0.50

Mimo 2.6 Pro: 46 // $0.43 // $0.87

Mimo 2.6 Flash: 38 // $0.10 // $0.28

Seems competitive to me? Plus then I don't have to manage multiple providers

▲mrngld 5 hours ago | parent | prev | next [-]

That's not what any benchmarks that look at cost per task or similar says in terms of cost. The Chinese models, generally speaking, might be cheaper per token but need a lot more tokens to get there.

▲RussianCow 2 hours ago | parent [-]

Except for the new MiMo V2.6 models, which appear to give some of the best value right now, at least on paper. (I haven't tried them so I can't speak from experience.)

▲wyrdcurt 5 hours ago | parent | prev | next [-]

Some people/organizations are ideologically opposed to using Chinese models. Not me, I use GLM-5.3-Flash for almost everything (the subscription-subsidized pricing on a legacy Z.ai plan makes it the best value model by a wide margin), along with some MiMo and DeepSeek. Still, I use Luna for certain tasks where speed is more valuable than performance; I can see this new Haiku displacing Luna for those. If you mean Haiku 4.5 though I agree, that model was a waste of time and money.

▲girvo 3 hours ago | parent | next [-]

I’m on the Legacy v2 plan and same: nothing comes close to 5.3 Flash’s value on it. It’s crazy, no wonder they discontinued them!

▲pimeys 3 hours ago | parent | prev [-]

Luna is not really the fastest. You need to use it in high/max to get the good output for what it is good for: summarizing. And that is already close to two minutes per task...

▲usef- 4 hours ago | parent | prev | next [-]

On subscription pricing a $20 Anthropic subscription gives >$500 equivalent tokens, which is not so different, and you get smarter models. API pricing has decent margins.

And Opus 5.5 is really good.

▲user43928 6 hours ago | parent | prev | next [-]

Presumably everyone who doesn't bother integrating a third party API key into their harness, which would probably be most of the Claude Code users.

▲pkulak 4 hours ago | parent | prev | next [-]

Where do you get this 10% number? Checking providers I know/respect, and GLM 5.3 flash is $0.15/m. Haiku is $0.10/m.

▲skeledrew 5 hours ago | parent | prev | next [-]

Well, unless you're using OpenCode Go, it's per-token costs (even if already super low), while Haiku falls under the Claude sub. It's just more straight forward and you aren't feeling a "loss" with the sub.

▲aesthesia 5 hours ago | parent | prev | next [-]

There really aren't any models at 10% of the price of Luna or Haiku.

▲ray_kay777 4 hours ago | parent | prev [-]

People who are stuck using Bedrock in-geo due to their company policy (me).