Comment by madduci

5 hours ago

Because Qwen4 has been announced!

and to squeeze anthropic, and other research innovations, not just chinese models but those help bring price down

And GLM 5.3 works great

  • Except for the censorship. We use it for massive data crunching, and roughly 5-8% (depending on the day) gets censored and doesn't get a response. We switched to Mimo 2.6, which is relatively better. For censored stuff, we use Sonnet and OpenAI Nano models.

    Also Mimo 2.6 is roughly 30% cheaper. Without batch.