← Back to context

Comment by ngl999

4 days ago

They won't. China doesn't have the training capacity and the quality data sets, the former because of access to fabs, the latter because of internal censorship that is getting worse by day.

Simple as that.

The current panic is happening precisely because Chinese open weight models like Kimi and Qwen are demonstrably competitive with SotA models like Fable, but far cheaper.

https://fireworks.ai/blog/kimik3-fable

  • No, the Chinese tech companies routinely super optimize for a given 'famous' or 'established' or 'mainstream' benchmarks. Any other concern is secondary.

    In other words, they look good on surface, but suck whenever anyone put them to any serious use. That's why they are cheap, they have to be cheap because they suck.

    • how can you say that with a username of "ngl999", that doesn't make sense

      (IIUC, -ngl [NUM_LAYERS] specify number of layers to offload to GPU in llama.cpp, 999 on most use cases might as well be -1)

      1 reply →

> the latter because of internal censorship

Having limited amounts of high quality Chinese data to train on doesn’t really decrease or affect the appeal of these models to the rest of the world. I don’t think the CCP is restricting or limiting the datasets that models can be trained on.

Also the models themselves don’t even seem to be censored or restricted that much. If the Chinese government is fine with the guardrails being on the API level and maybe even restricts the usage of external providers that’s again a net benefit to everyone else since they would have little incentives to force Chinese companies to lobotomize their models.

You are probably about capacity so we can only hope that Huawei and others can catch up and break Nvidia’s monopoly (since Intel and AMD suck too much too much to accomplish anything useful that’s the next best thing)

Also Chinese labs are actually still publishing research publicly which alone would accelerate and increase the competitiveness of open AI models developed in the US or even Europe.

  • > doesn’t really decrease or affect the appeal of these models to the rest of the world

    Anyone who tried using Chinese models for serious programming tasks ditch them quickly if they have a choice.

    > I don’t think the CCP is restricting or limiting the datasets that models can be trained on.

    What I meant was censorship limits the amount of high quality training sets by limiting the amount of all training sets from which high quality sets grow out from, e.g. the entire set of posts on Baidu forums before 2017 is gone forever.

    > so we can only hope that Huawei

    Chinese companies do not have access to commercial-scale advanced nodes. The impact is at best negligible.

    • > Anyone who tried using Chinese models for serious programming

      Perhaps, but that’s because they are just objectively worse for many tasks than GPT/Claude. I don’t see how that ties to censorship inside of China.

      > censorship limits the amount of high quality training sets by limiting the amount of all training sets

      Well again, I don’t really understand how is the lack of high quality Chinese datasets matter much if they have access to English/etc. datasets. That’s only an issue to their users inside of China.

      > not have access to commercial-scale advanced nodes

      Well the gap was way bigger a few years ago. If Chinese companies can at some point produce GPUs that are competitive cost wise i.e. they are willing to tolerate significantly lower margins than Nvidia (whose margins are obscene) that’s not a huge issue.

      2 replies →