Comment by system2

6 hours ago

Who in their right mind would use haiku while Mimo or GLM cost 10% of what they are charging with much smarter models?

That's not what any benchmarks that look at cost per task or similar says in terms of cost. The Chinese models, generally speaking, might be cheaper per token but need a lot more tokens to get there.

  • Except for the new MiMo V2.6 models, which appear to give some of the best value right now, at least on paper. (I haven't tried them so I can't speak from experience.)

Some people/organizations are ideologically opposed to using Chinese models. Not me, I use GLM-5.3-Flash for almost everything (the subscription-subsidized pricing on a legacy Z.ai plan makes it the best value model by a wide margin), along with some MiMo and DeepSeek. Still, I use Luna for certain tasks where speed is more valuable than performance; I can see this new Haiku displacing Luna for those. If you mean Haiku 4.5 though I agree, that model was a waste of time and money.

  • Luna is not really the fastest. You need to use it in high/max to get the good output for what it is good for: summarizing. And that is already close to two minutes per task...

  • I’m on the Legacy v2 plan and same: nothing comes close to 5.3 Flash’s value on it. It’s crazy, no wonder they discontinued them!

Isn't the point of this release that it's comparable?

AAI Index // Input // Output

Haiku 5.5: 43 // $0.10 // $0.50

Mimo 2.6 Pro: 46 // $0.43 // $0.87

Mimo 2.6 Flash: 38 // $0.10 // $0.28

Seems competitive to me? Plus then I don't have to manage multiple providers

Presumably everyone who doesn't bother integrating a third party API key into their harness, which would probably be most of the Claude Code users.

Where do you get this 10% number? Checking providers I know/respect, and GLM 5.3 flash is $0.15/m. Haiku is $0.10/m.

On subscription pricing a $20 Anthropic subscription gives >$500 equivalent tokens, which is not so different, and you get smarter models. API pricing has decent margins.

And Opus 5.5 is really good.

Well, unless you're using OpenCode Go, it's per-token costs (even if already super low), while Haiku falls under the Claude sub. It's just more straight forward and you aren't feeling a "loss" with the sub.