5 ms·
Side note: I don’t know what Anthropic changed but now Claude Code consumes the quota incredibly fast. I have the Max5 plan, and it just consumed about 10% of t
by grewil2 6mo ago
Side note: I don’t know what Anthropic changed but now Claude Code consumes the quota incredibly fast. I have the Max5 plan, and it just consumed about 10% of the session quota in 10 minutes on a single prompt. For $100/month, I have higher expectations.
- maximinus_thrax 6mo agoI'm very surprised to see enshittification starting so early. I was expecting at last 3-4 years of VC subsidized gravy train.
- kderbyma 6mo agoThis has been 6 months of constant decline so at this point I am wondering when they cliff it like wework
- landr0id 6mo agoRelevant: https://www.reddit.com/r/ClaudeAI/comments/1s7zgj0/investigating_usage_limits_hitting_faster_than/ https://www.reddit.com/r/ClaudeAI/comments/1s7zgj0/investiga... https://www.reddit.com/r/ClaudeAI/comments/1s7mkn3/psa_claude_code_has_two_cache_bugs_that_can/ https://www.reddit.com/r/ClaudeAI/comments/1s7mkn3/psa_claud...
- nixpulvis 6mo agoThis is a serious problem with the fact that it's nearly impossible to understand what a "token" is and how to tame their use in a principled way. It's like if cars didn't advertise MPG, but instead something that could change randomly.
- uoaei 6mo agoLike if cars measured fuel efficiency or range using the knobs in the tread on your tire.
- amitprasad 6mo agoRelevant post: https://modal.com/blog/dollars-per-token-considered-harmful https://modal.com/blog/dollars-per-token-considered-harmful (disclaimer: I work with the author)
- nixpulvis 6mo agoI completely agree that requests are what should be charged for. But I think there are two things, given that requests aren't all going to cost the same amount: 1. Estimate free invoicing the requests and letting users figure it out after the fact. 2. Somehow estimating cost and telling users how much a request will cost. We have 1, we want 2.
- clawfund 6mo ago[flagged]
- sebmellen 6mo agoPlease do not bot HN.
- claw-el 6mo agoAlso, certain models are more verbose than the others. We are basically at the mercy of a model who likes to ramble a lot.
- konfusinomicon 6mo agoim fiarly certain the knob on the machine that controls length of redundant comments and docblocks is cranked to 11. it makes me curious how much of their bottom line is driven by redundant comment output.
- smohare 6mo ago[dead]
- onemoresoop 6mo agoThat explains things. Im getting this: API Error: 400 {"error":{"message":"Budget has been exceeded! Current cost: 271.29866200000015, Max budget: 200.0","type":"budget_exceeded","param":null,"code":"400"}} So I completetly ran out of tokens and haven’t even used it at all for the past couple of days, and last week my usage was very light. Let me scratch that, all my usage has been very light since I got this plan at work. It’s a an enterprise subscription I believe, hard to tell since it doesn’t connect directly to Anthropic, rather it goes through a proxy on Azure. Im not liking this at all and all, so flaky and opaque. Not possible to get a breakdown on what the usage went on, right? Do we have to contact Anthropic for a refund or will they restore the bogus usage?
- prodigycorp 6mo agoAnthropic really needs to opensource claude code. One of the biggest turnoffs as a claude code user is the CC community cargo culting the subreddit because community outreach is otherwise poor.
- landr0id 6mo agoLooks like your wish was accidentally granted :)
- prodigycorp 6mo agoHilarious. Makes you hope CC team sticks to it, so much goodwill to be had by doing so.
- manmal 6mo agoLooks like they are falling victim to their own slop. This smells a lot like the Amazon outages caused by mandated clanker usage.
- no1youknowz 6mo agoI've been jumping from Claude -> Gemini -> GPT Codex. Both Claude and Gemini really reduced quotas and so I cancelled. Only subbed GPT for the special 2x quota in March and now my allocation is done as well. I decided to give opencode go a try today. It's $5 for the first month. Didn't get much success with Kimi K2, overly chatty, built too complex solutions - burned 40% of my allocation and nothing worked. ¯\_(ツ)_/¯. But Minimax m2.7. Wow, it feels just like Claude Opus 4.6. Really has serious chops in Rust. Tomorrow/Wednesday will try a month of their $40 plan and see how it goes.
- victorbjorklund 6mo agoMinimax 2.7 is great. Not close to Claude but good enough for a lot of coding tasks.
- conception 6mo agoI noticed 1M context window is default and no way not to use it. If your context is at 500-900k tokens every prompt, you’re gonna hit limits fast.
- aberoham 6mo agoexport CLAUDE_CODE_DISABLE_1M_CONTEXT=1
- teaearlgraycold 6mo agoAnthropic is not building good will as a consumer brand. They've got the best product right now but there's a spring charging behind me ready to launch me into OpenCode as soon as the time is right.
- kylecazar 6mo agoWould you use Opus if you switched to OpenCode?
- teaearlgraycold 6mo agoI'd like to use Opus with OpenCode right now to combine the best TUI agent app with the best LLM. But my understanding is Anthropic will nuke me from orbit if I try that.
- corford 6mo agoOpenCode with a Copilot Business sub and Opus 4.6 as the model works well
- teaearlgraycold 6mo agoI'm looking at their plans (https://github.com/features/copilot/plans https://github.com/features/copilot/plans) it seems like the limits might be pretty low, even with the Pro+ plan which is 2x the cost of Claude Pro. It seems like Claude Pro might be 10-20x the Opus tokens for only twice the price.
- skwallace36 6mo agothings are rough out there right now
- lkbm 6mo agoI've heard this a few times lately, but this past weekend I built a website for a friend's birthday, and it took me several hours and many queries to get through my regular paid plan. I just use default settings (Sonnet 4.6, medium effort, thinking on). I'm guessing Opus eats up usage much, much faster. I don't know what's going on, since a lot of people are hitting limits and I don't seem to be.
- notatoad 6mo agowhat they changed was peak vs off-peak usage metering. using it on the weekend gets you more use than during weekdays 9-5 in US eastern time.
- lkbm 6mo agoTechnically, this was Friday morning, so I think I was still in peak hours.
- hrimfaxi 6mo agoI'm surprised it's during east coast working hours and not west coast.
- notatoad 6mo agothe speculation i read was that it's trading hours, and they're getting a lot of load from the finance industry
- matheusmoreira 6mo agoI waited until off peak hours to use Opus 4.6 to do some research. One prompt consumed 100% of my 5h limit and 15% of my weekly usage. Even off peak it's still insane. Opus didn't even manage to finish what it was doing.
- teaearlgraycold 6mo agoEven with Opus I don’t usually hit limits on the standard plan. But I am not doing professional work at the moment and I actually alternate between using the LLM and reading/writing code the old fashioned way. I can see how you’d blow through the quota quickly if you try to use LLMs as universal problem solvers.
- outside1234 6mo agoThey need to get to profitability because that sweet sweet Saudi subsidy cash is gone gone.
- kderbyma 6mo agoThey wont be profitable at this point...they just dont realise they are eating their own tail.
- irishcoffee 6mo agoReminds me of when I would mess with my friends on "pay per text" plans by sending them 10 text messages instead of just 1. I should start paying attention to unattended laptops and blow up some token usage in the same manner. It's almost like an evolution of bobby tables.
- xantronix 6mo agoThis is a very normal thing to be the top comment on an article on how to use Claude Code.
- alcor-z 6mo ago[dead]
- LeonTing1010 6mo ago[flagged]
- zar1048576 6mo agoHave had similar issues with costs sometimes being all over the map. I suspect that the major providers will figure this out as it’s an important consideration in the enterprise setting