8 ms·
"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pric
by pookieinc 16d ago
"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pricing on cache reads (where the model reads inputs that have already been processed and stored). For highly agentic work, the savings will often be much larger—up to approximately 45%."
Glad to see this!
- benjiro29 16d agoI hope that this also applies to the Subscription usage. As that can then stretch out Fable usage by a lot more.
- Narretz 16d agoSounds like it specifically does not apply to subscription usage.
- re-thc 16d agoThat's load bearing!
- LtdJorge 16d agoBut does it fail open or close?
- fearmerchant 16d agoThe gate is green
- aprilnya 16d agoMy understanding is subscription usage generally has free cache reads, but I'm not sure if maybe Fable was different in that regard.
- MitziMoto 16d agoThe fact that we don't know is part of the problem. Subscription usage has always been pretty opaque.
- scrollop 16d agoAnd then you see this: https://artificialanalysis.ai/models#cost-tabs https://artificialanalysis.ai/models#cost-tabs
- ActionHank 16d agoThe big issue they face right now is that vastly cheaper open models are proving capable for more and more uses at cents on the dollar. This is the right direction, but they aren't going to get there fast enough. They will list, investors who don't know anything about tech will buy, the world will realise that China just put out a model that is good enough at a fraction of the price, they will crater.
- victor9000 16d agoUS enterprise customers aren't going to convert to overseas models, average consumers might though.
- ActionHank 16d agoAt 1/10th the price they will. Claude is way overpriced for most peoples needs.
- george_max 16d agoThis is just cache reads. In real usage it costs 15% more than Fable 5 -- all for marginal gains. https://artificialanalysis.ai/ https://artificialanalysis.ai/
- edg5000 16d agoCache reads dominate in modern workflows (coding CLIs and modern web clients such as ChatGPT Work and Claude Cowork (web)).
- fastball 16d agoOutput tokens are 5x more expensive than input tokens, so I'm not sure "dominate" is entirely correct. A conversation with 20 turns, 50k tok growth per turn, 1m tok context at end would price out like this: Fable 5 ($1/M cache reads) ; cache reads 9.5M tok × $1.00 = $9.50 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $72.00 Fable 5.1 ($0.25/M cache reads) ; cache reads 9.5M tok × $0.25 = $2.38 ; cache writes 1M tok × $12.50 = $12.50 ; output 1M tok × $50 = $50.00 ; total = $64.88 So yes, cheaper, but not massively.
- edg5000 16d ago[dead]