3 ms·
"Cache reads now cost 75% less, or $0.25 per million tokens." For me, at a typical 95% cache hit rate, I think my optimal context window size before autocompact
by seaurchinzee 1mo ago
"Cache reads now cost 75% less, or $0.25 per million tokens." For me, at a typical 95% cache hit rate, I think my optimal context window size before autocompaction goes from ~200K to ~400K tokens. Great for longer horizon tasks.
- cute_boi 1mo agolooks like it is only for api.....
- kingstnap 1mo agoDo API prices not affect usage limits for subscriptions? They do in Codex.
- seaurchinzee 1mo agoOh dang, that's really unfortunate, nice catch. At least Claude subscription users got a usage reset. But yeah, I can't help but feel Codex is far more generous with their subscription quota at the moment. I've been using Fable to orchestrate GPT Sol Max and Sol Ultra agents all day, and I've barely made a dent.
- eaf7e281 1mo agomay i ask where did you get this? i try to look through the docs, but i didn't find where they said its only for API is it in the system card? really hope not, that change the only positive part in this release
- demibabs 1mo agoThey specifically said it in the press release. I don’t see why they wouldn’t have mentioned it if it also applied to subs
- skiph 1mo ago[flagged]