Comment by simonw
14 hours ago
Bit of a discount if you're using caching:
> same input and output prices, with cache reads at a quarter of the cost
This should impact any long-running agent since subsequent calls can benefit from cached reads for previous transcripts.
~30% reduction in real-world task cost vs. Fable 5 in our evals at viktor.com ! Caching goes a looong way
And yet, despite this, the quota limits went down by 17%.
In my opinion, this is a bit disingenuous.
They were _temporarily_ increased in May by 50% [1]. They continued to extend them through July and August (admittedly, their messaging around this has just been a complete mess and they frequently pushed the deadline back as it approached).
So, now they are giving you a 25% quota increase compared to where things originally stood in May.
So, let me ask you this: assuming you knew that the 50% quota increase was temporary all along, would you then have complained about Anthropic restoring things back to the original limit?
[1] https://www.anthropic.com/news/higher-limits-spacex
Yes, some people will complain about anything (and everything) related to AI. And relentlessly push the most negative interpretation of any datum.
On the contrary, you and Anthropic are being disingenuous by pretending that a usage reduction is actually an increase. Especially when the 20x max plan isn't actually anywhere near 20x, as people have recently realized.