← Back to context

Comment by MikhailTal

8 hours ago

> Google DeepMind: We are building strong momentum: Flash is in high demand, our Cyber model is live, and Gemma models have surpassed 900M+ downloads

Considering these are the best stats they could find, gemini usage+general situation must be really, really bleak.

High demand means nothing. A model being live is nothing to brag about. And gemma downloads also can be from auto CI pipelines etc. Nothing concrete

> Flash is in high demand

I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?

  • That said though, Flash isn't it. The prices on the latest flash models put Sonnet and Terra to shame.

  • Does everyone want AI?

    • For what? There are some things I want AI for because it does it well. There are some things I don't want AI for because it just makes a mess (hallucinations). Maybe the next AI will be different and we will have the conversation again.

      2 replies →

    • I guess it should be said, of everyone who wants AI, they don't want to pay the expense for it to the level they want to use it.

  • > I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?

    I keep seeing this but this line of thinking doesn't make any sense. What does it really mean?

    There are expensive models that increase the probability of you doing your task under a lower cost. That means you can't use Gemma for coding your new compiler - it would just be overall costlier.

    Heavier models are cheaper at more complicated tasks because they use fewer turns and fewer mistakes.

    Cheaper models are more likely to be cheap at less complicated tasks. Like if you just ask Gemma "Hi" it would probably be cheaper than asking Opus.

    So what does this statement really mean? People don't want to pay the extra for a more costly model? Why wouldn't you? It reduces your overall cost!

    • > Why wouldn't you? It reduces your overall cost!

      Because real Fable usage starts at $20/month, and has oppressive usage limits even at that (ridiculous) monthly price.

      Compared to my $3/month GLM-5.2 subscription, I have never felt like I was leaving capabilities on the table by refusing to cough up $20 for 15 minutes of Fable use per day.

      6 replies →

> And gemma downloads also can be from auto CI pipelines etc. Nothing concrete

I have always found NPM download numbers truly suspect. Is no one caching? Are they estimating true number of downloads base on some estimate of cache hits?

  • Absolute download numbers are a pure vanity metric. Relative download numbers compared to other models tell you a little bit.

  • > And gemma downloads also can be from auto CI pipelines etc. Nothing concrete

    Thank you for adding some clarity to this. When I calculated 900 million downloads divided by 8.3 billion people in the world, I came with a number that made it look like about one person in 10 were downloading this model.