2 ms·
I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04/M, that is $8-12 per day
by sieve 11d ago
I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04/M, that is $8-12 per day only for cached reads.
- ignoramous 11d ago> At $0.04/M Unless you meant step-3.7-flash, the input cache hits are $0.05 per mil for step-5-preview. > $8-12 per day only for cached reads Pretty decent "API" rates for ~500M+ tokens on Step Fun 5, a Kimi K3 / GLM 5.3 level model? Their "Step Plan" is ridiculous, by comparison: ~$60 usage on $6.99/mo; ~$220 on $9.99/mo. https://platform.stepfun.ai/docs/en/step-plan/overview https://platform.stepfun.ai/docs/en/step-plan/overview
- sieve 11d agoYes, I meant the Flash version. I have used Kimi 2.5 and GLM 5.3 (& 5.3 Flash). Do not need them for what I do outside of spec hardening (basically, a lot of chatting). I tend to know exactly what I want and most of the weaker models are enough to get me there. I have mainly been using MiMo, DeepSeek V4 Flash and MuseSpark Contributor over the last month or so.
- esafak 11d agoNot generous enough? How token efficient and fast is it compared with American models?