4 ms·
Bit of a discount if you're using caching: > same input and output prices, with cache reads at a quarter of the cost This should impact any long-running agent
by simonw 1mo ago
Bit of a discount if you're using caching:
> same input and output prices, with cache reads at a quarter of the cost
This should impact any long-running agent since subsequent calls can benefit from cached reads for previous transcripts.
- Twixes 1mo ago~30% reduction in real-world task cost vs. Fable 5 in our evals at viktor.com ! Caching goes a looong way
- behnamoh 1mo agoAnd yet, despite this, the quota limits went down by 17%.
- davely 1mo agoIn my opinion, this is a bit disingenuous. They were _temporarily_ increased in May by 50% [1]. They continued to extend them through July and August (admittedly, their messaging around this has just been a complete mess and they frequently pushed the deadline back as it approached). So, now they are giving you a 25% quota increase compared to where things originally stood in May. So, let me ask you this: assuming you knew that the 50% quota increase was temporary all along, would you then have complained about Anthropic restoring things back to the original limit? [1] https://www.anthropic.com/news/higher-limits-spacex https://www.anthropic.com/news/higher-limits-spacex
- Petersipoi 1mo agoOn the contrary, you and Anthropic are being disingenuous by pretending that a usage reduction is actually an increase. Especially when the 20x max plan isn't actually anywhere near 20x, as people have recently realized.
- anthonyrstevens 1mo agoYes, some people will complain about anything (and everything) related to AI. And relentlessly push the most negative interpretation of any datum.