Comment by dyauspitr
1 day ago
What I’ve seen is coding is usually done with US frontier models and anything that is part of a feature on an app and runs at scale on the API is a Chinese model because they are dirt cheap.
1 day ago
What I’ve seen is coding is usually done with US frontier models and anything that is part of a feature on an app and runs at scale on the API is a Chinese model because they are dirt cheap.
That has been my experience too.
It will cost you more than it saves to use smaller Chinese models to code; because of the repeated work. That has been slowly changing recently, but with much larger Chinese models, however those models are so expensive they're much more price-competitive iwth the US competition.
But for actually providing end-user AI features, particularly simpler ones, the US isn't even in contention. The costs and limitations just outright kill those features conceptually.
That's what I lean towards with the exception that Gemma is also good on a lot of tasks and cheap, while not being Chinese.
The problem is Gemma is actually not that cheap compared to deepseek for example