4 ms·
"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pric
by pookieinc 26d ago
"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pricing on cache reads (where the model reads inputs that have already been processed and stored). For highly agentic work, the savings will often be much larger—up to approximately 45%."
Glad to see this!
- benjiro29 26d agoI hope that this also applies to the Subscription usage. As that can then stretch out Fable usage by a lot more.
- Narretz 26d agoSounds like it specifically does not apply to subscription usage.
- re-thc 26d agoThat's load bearing!
- LtdJorge 26d agoBut does it fail open or close?
- fearmerchant 26d agoThe gate is green
- aprilnya 26d agoMy understanding is subscription usage generally has free cache reads, but I'm not sure if maybe Fable was different in that regard.
- MitziMoto 26d agoThe fact that we don't know is part of the problem. Subscription usage has always been pretty opaque.
- scrollop 26d agoAnd then you see this: https://artificialanalysis.ai/models#cost-tabs https://artificialanalysis.ai/models#cost-tabs
- ActionHank 26d agoThe big issue they face right now is that vastly cheaper open models are proving capable for more and more uses at cents on the dollar. This is the right direction, but they aren't going to get there fast enough. They will list, investors who don't know anything about tech will buy, the world will realise that China just put out a model that is good enough at a fraction of the price, they will crater.
- victor9000 25d agoUS enterprise customers aren't going to convert to overseas models, average consumers might though.
- ActionHank 25d agoAt 1/10th the price they will. Claude is way overpriced for most peoples needs.
- george_max 26d agoThis is just cache reads. In real usage it costs 15% more than Fable 5 -- all for marginal gains. https://artificialanalysis.ai/ https://artificialanalysis.ai/
- edg5000 25d agoCache reads dominate in modern workflows (coding CLIs and modern web clients such as ChatGPT Work and Claude Cowork (web)).
- fastball 25d agoOutput tokens are 5x more expensive than input tokens, so I'm not sure "dominate" is entirely correct. A conversation with 20 turns, 50k tok growth per turn, 1m tok context at end would price out like this: Fable 5 ($1/M cache reads) ; cache reads 9.5M tok × $1.00 = $9.50 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $72.00 Fable 5.1 ($0.25/M cache reads) ; cache reads 9.5M tok × $0.25 = $2.38 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $64.88 So yes, cheaper, but not massively.
- edg5000 25d ago[dead]