← Back to context

Comment by gorgmah

11 hours ago

We already know that competition brought GLM 5.2 prices down roughly 45% since its release on June 16th (1.5 months ago), and the price downward slope is probably still going (I've been checking regularly and new providers keep fighting on price, I don't think prices have settled yet). For reference : https://openrouter.ai/z-ai/glm-5.2#providers

I saw arguments like "Providers cannot price less than their costs" in other comments. In economics, it's generally admitted that they shouldn't price less than their marginal costs, i.e. in their case roughly the cost of electricity, since a lot of these datacenters are not at capacity in terms of graphics cards usage (speculation since it's very easy to rent a GC for a couple hours on some providers). My guess is that someone will be selling tokens at less than electricity + depreciation of GCs soon, since there's a lot of competition and "smaller" data centers have overcapacity? This is speculation, correct me if I'm wrong

> My guess is that someone will be selling tokens at less than electricity + depreciation of GCs soon, since there's a lot of competition and "smaller" data centers have overcapacity? This is speculation, correct me if I'm wrong

My guess is they are selling you the tokens, then selling your tokens (data) onto someone else.

  • I see these conspiratorial arguments all the time and I think people massively overestimate the value of the average users tokens.

    The problems with frontier models (design taste, ability to solve novel/difficult problems, etc) cannot be solved by throwing more slop from the average user at it.

    Actually, most of the main deficiencies in current models stem from the fact that their data sets aren’t curated and specialized enough.

    • I don't think the goal of this data is necessarily model improvement.

      I think it's marketing, advertising, and product refinement.

      Ex: all the things Google wants your search data for.

      It's somewhat silly to think the value of that data has changed much. Advertisers want to know what's popular and getting clicks and attention. Competitors want to know what features are getting used in their markets.

      In the simplest case, think of this data as improving the harness, not the model.

      1 reply →

    • The prompts contain sensitive personal data.

      That would be valuable to advertizers for example.

Press x to doubt on the 45% number. The cheaper providers on open router are fp4 vs fp8 for official zai. There are some cheap fp8 ones (like novita) but the ui makes it seem like it's a temporary promotion, with their normal prices being almost equal to official zai (idk much about open router so not really sure what's going on with these discounts)

  • Yes it's true that it's not super clear whether these prices are permanent or short term promotions. On the other hand, there are so many providers making promotional offerings that you could probably easily switch from one to another should their prices go up?

  • If you click the provider it shows the precision, 45% off at AkashML shows FP8. The drawback is the small context window, at 96k.

    Then there's 43% off at StreamLake with FP8 precision and 1M context window.

  • There is nothing to doubt, the cheapest price on openrouter is ~45% lower than when GLM5.2 was released.

Seminanlysis is estimating sub $1 cost per MT for ~2Trillion models. The numbers change based on throughput and quant, but it is conceivable that provider costs at scale are low enough that even $2.42 per MT on GLM 5.2 (current best price) is margin positive by a wide margin.