Comment by pookieinc

13 hours ago

"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pricing on cache reads (where the model reads inputs that have already been processed and stored). For highly agentic work, the savings will often be much larger—up to approximately 45%."

Glad to see this!

This is just cache reads. In real usage it costs 15% more than Fable 5 -- all for marginal gains.

https://artificialanalysis.ai/

  • Cache reads dominate in modern workflows (coding CLIs and modern web clients such as ChatGPT Work and Claude Cowork (web)).

    • Output tokens are 5x more expensive than input tokens, so I'm not sure "dominate" is entirely correct.

      A conversation with 20 turns, 50k tok growth per turn, 1m tok context at end would price out like this:

      Fable 5 ($1/M cache reads) ; cache reads 9.5M tok × $1.00 = $9.50 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $72.00

      Fable 5.1 ($0.25/M cache reads) ; cache reads 9.5M tok × $0.25 = $2.38 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $64.88

      So yes, cheaper, but not massively.

The big issue they face right now is that vastly cheaper open models are proving capable for more and more uses at cents on the dollar.

This is the right direction, but they aren't going to get there fast enough.

They will list, investors who don't know anything about tech will buy, the world will realise that China just put out a model that is good enough at a fraction of the price, they will crater.

I hope that this also applies to the Subscription usage. As that can then stretch out Fable usage by a lot more.