← Back to context

Comment by simonw

14 hours ago

Bit of a discount if you're using caching:

> same input and output prices, with cache reads at a quarter of the cost

This should impact any long-running agent since subsequent calls can benefit from cached reads for previous transcripts.

~30% reduction in real-world task cost vs. Fable 5 in our evals at viktor.com ! Caching goes a looong way

And yet, despite this, the quota limits went down by 17%.

  • In my opinion, this is a bit disingenuous.

    They were _temporarily_ increased in May by 50% [1]. They continued to extend them through July and August (admittedly, their messaging around this has just been a complete mess and they frequently pushed the deadline back as it approached).

    So, now they are giving you a 25% quota increase compared to where things originally stood in May.

    So, let me ask you this: assuming you knew that the 50% quota increase was temporary all along, would you then have complained about Anthropic restoring things back to the original limit?

    [1] https://www.anthropic.com/news/higher-limits-spacex

    • Yes, some people will complain about anything (and everything) related to AI. And relentlessly push the most negative interpretation of any datum.

    • On the contrary, you and Anthropic are being disingenuous by pretending that a usage reduction is actually an increase. Especially when the 20x max plan isn't actually anywhere near 20x, as people have recently realized.