← Back to context

Comment by lwansbrough

3 hours ago

Anyone else more excited about Chinese models than American models these days? Big thing for me is affordability.

Absolutely! Chinese models are both cheaper and more capable in many cases, compared to the American models and their makers continuously fumbling or reducing model capability with each update. Deepseek decreased costs when they released Flash 4.1 you would not see any American company do this, in reverse they would try charge you more.

  • OpenAI decreased prices with the 5.6 model family.

    And later they further cut Sol and Terra pricing by 20% (maybe only in the API) and Luna by 80%.

    In fact Luna still outperformed DeepSeek Flash 4.1 in cost per task on Artificial Analysis when I last checked.

    However, Luna is slightly less intelligent. I have a feeling that it's pretty dumb and prone to hallucination unless running at xhigh or max effort, where it somehow manages to work quite well.

    I did not personally test the open weight models beyond the old Qwen 3.6 27B, which produced unusably bad results for me.

    The competition is great, and I hope Chinese models will continue to force leading US labs to offer models at a low price point.

    That said, I don't think the Chinese labs have anything over OpenAI and Anthropic when it comes to capability or efficiency - I have no reason not to believe the US labs have even lower cost to serve the models.

    • OpenAI had to cut costs because of Anthropic. I also do not trust the benchmarks when it comes to models anymore. I have tried both Claude and OpenAI models and while it is true that the 5.6 series is smarter than Deepseek (at the time i tested it against 4.0) at that price it is still not worth it and sometimes randomly refuses to do tasks or stops midway etc.

      Do also remember China is this far in the AI race despite all chip restrictions from America. If they were in equal standards I truly think Chinese models would have long surpassed American ones. Also would like to remind how Anthropic CEO is being hostile and blaming Chinese models with distilling meanwhile their own models claimed to be Qwen¹ and their stance against open models is negative² and they still keep blaming China for it.

      1- https://www.anthropic.com/news/position-open-weights-models

      2 replies →

    • > I did not personally test the open weight models beyond the old Qwen 3.6 27B, which produced unusably bad results for me.

      So you don't have much perspective on things, it seems. Let me introduce you to the GLM 5.2 and then 5.3/5.3 flash series of... "oh, wow, I should have bought some RTX PRO 6000's while they were 'cheap'" stage of progression.

      As someone carrying multiple max subscriptions to both claude and codex - primary workhorse is glm 5.3 flash running on rented GPUs for less than a latte/hr.

      I also found qwen 3.6 27B nearly useless for my own needs. DS4 flash 0731 and then 4.1 have been nearly as eye opening as glm 5.3 flash, but have their own warts.

      2 replies →

  • > Deepseek decreased costs when they released Flash 4.1 you would not see any American company do this, in reverse they would try charge you more.

    OpenAI reduced prices and Anthropic increased weekly usage limits.

Absolutely! DeepSeek-V4-Flash-0731 has become my daily driver. It's pretty amazing what it can do for what it costs at deepinfra.com (I don't use deepseek as a provider since they train on your data [at least their honest about it]). GLM-5.1 was my daily driver before that and Kimi K2.5 before that.

  • How does it compare to 4.1 flash? Curious why folks don’t use the more “modern” one.

    • I haven't tried 4.1 flash as I'm assuming its a preview. I did not get good results from the preview version of 4.0 flash (i.e. the one that did not include the month and date of release in its name).

      1 reply →

  • Are you finding DS better then kimi k3 and glm-5.3? Do you mind sharing your primary use case?

    • I've used Kimi K3 for a few months as my main model and DeepSeek 4.1 is as fast and about 10x cheaper.

      I just had like four big sessions going today, paid about $8 in tokens. I see no reason to pay more, this is more than I need for intelligence.

    • My primary use is AI coding agent. Its vastly cheaper than Kimi K3 and I haven't found a scenario where I really need Kimi K3 versus smaller models. GLM-5.3 Flash is good but there is series of bugs in the vllm middleware that prevent GLM models from getting all of their reasoning content returned to them that impairs inference quality. A lot of inference providers use vllm which makes it hard to find a good provider for GLM. I've been using friendli.ai but using GLM-5.3 Flash from them is more expensive then using DS V4 Flash from deepinfra.com simply because deepinfra.com is so cheap. The DS V4 Flash cost at together.ai is similar to the GLM-5.3 Flash from friendli.ai or at least that's what I found in my benchmarks a week ago: https://www.linkedin.com/posts/joshheitzman_i-ran-a-fuller-r...

Yep, I'm trending in that direction, and I'm someone with Claude stickers all over my laptop. My main app dev work is still going to Claude, but everything else is going to China even at API rates now.

One simple task: I needed an LLM to go through and clean up a few thousand page descriptions and titles in my personal search engine index, where the human web page authors had put in no effort sigh. I did a shoot out between Claude, Luna, GLM 5.3 Flash and Deepseek. Despite the high cost, Claude's descriptions were terrible, and even Opus warned me that the descriptions coming back from Haiku were "generalized, not accurate". I expected I would choose Luna because of price, and occasionally it did have wonderful descriptions (one captured emotion in a way no other model did). But in the end, the GLM 5.3 Flash descriptions were the easiest to read, they flow well while also being accurate & including necessary keywords, and being highly affordable. So it won out. It's a task that is nowhere near frontier, but a task where somehow China is better than frontier.

  • API rates still aren’t quite competitive with the OpenAI x20 accounts, but they are definitely getting close with deepseek 4.1 flash. I spent a few days with only 4.1 and was very impressed.

I have a contrarian opinion that China passing America in Ai is the Sputnik moment we need to leave the hubris behind and get our mojo back

debatable if a turn around is possible before '29

  • The analogy makes little sense. The USA was not in front of the USSR and Sputnik merely showed that. It is at this point that the Americans woke up, put a lot of effort and finally were able to surpass the Soviets during the Apollo missions.

    China was never ahead of the USA in AI. So perhaps a more proper analogy is the Moon landing. In real history the side that lost the race never got its mojo back...

    • I'm looking forward, towards the future, when I use "passing ... we need", need being key here as it implies something we don't yet have

      I expect this to happen within 12-18 months, the differentiation has shrunk, many models are now sufficiently capable for most tasks

      1 reply →

Yes, an expensive American LLM has zero capabilities as far as I’m concerned because I’m never going to pay for it.

No, because I'd rather not support our economic and military rivals.

  • I'm Canadian so this sentiment has little value in 2026 unfortunately.

    • As much as the US has been easy to hate lately, I don't hesitate to say Xi Jinping as the most powerful man on Earth would be much, much worse.

    • Also, frankly, as a fellow Canadian it's pretty clear that the biggest "rival" the US has right now is itself. Just passed out in the corner puking on itself shouting about all the foreigners who won't talk to it.

      1 reply →

    • I'm from Europe and I hate America way more than China now. Used to be about equal but then Trump started extorting Ukraine, threatening their own allies and sending billions to Israel to help with a genocide. I think that exposed America for what it really is.

      5 replies →

  • Does it count as supporting a rival if your an American using an American inference provider self-hosting an open weight model from a Chinese lab?

  • Agreed, and also because I support freedom of speech!

    • Neither the US nor the Chinese companies are on your side then. They both censor, just different topics.

      But at least I can run Chinese models locally, and strip a lot of that censorship/refusal.

    • As long as that speech doesn't come from CNN or criticise Charlie Kirk, Israel or Trump? I'm sceptical about how much the US really values free speech