6 ms·
Okay, so anthropic has amazing AI which supposedly writes most of their code and can continuously improve... meanwhile they have outages on a regular basis, and
by aleqs 4mo ago
Okay, so anthropic has amazing AI which supposedly writes most of their code and can continuously improve... meanwhile they have outages on a regular basis, and any kind of long-running work will now consistently hit 'API Error: Server is temporarily limiting requests'. Not sure of this is intentional to force a reduction of token usage, but at this point I need to build around these throttling limits and outages with my own tools to restart/resume sessions. From my experience, in the last 2 weeks, literally 100% of any non-trivial Claude session/work will now be blocked on these issues, requiring manual intervention.
One of my focuses now is my own model-agnostic, harness and workflow orchestration (I know everyone is building these) , baselining on opus, and aiming to transition to Chinese models like deepseek in the short term and hopefully open, self hosted models in the future (which I plan to open source).
The nonstop marketing fluff from anthropic while their service quality and availability noticeably degrades... just continues to destroy my trust in the company.
- ChadMoran 4mo agoBetter doesn't mean perfect.
- jakobnissen 4mo agoTheir outages are probably not due to their code though. It’s probably their infrastructure that can’t keep up. So seeing failures of infrastructure doesn’t really tell you anything about how good or bad Anthropic makes use of their models.
- aleqs 4mo agoThat seems like an assumption based on basically nothing. There is a lot of code at the infra layer, and based on the stack choices for Claude code and based on how buggy and unreliable ~everything from anthropic is, it seems pretty bizarre to claim these issues are not related to their code.
- keeda 4mo agoThere are other indications, however, like Anthropic paying through the nose for compute just months after Dario told Dwarkesh how hard it is to predict demand, or ChatGPT and Codex not quite having the same issues after Altman spent much-publicized years scrounging for trillion-dollars of capacity. While I'm very bullish on Anthropic, I'm a bit wary about their IPO because it seems to me that they're filing now while their financials look good and before other trends like the decline of tokenmaxxing and their compute bills catch up.
- qwery 4mo agoWhoa, first name basis with Dario but not Sam. Ouch. [I actually have no idea who Dwarkesh is and it sounds like a first name to me but that's not a particularly reliable indicator so I won't comment on your relationship with Dwarkesh.] Oh, are they filing now? I think their financials look somewhere in between devastating and criminal, so I'm really looking forward to the IPO!
- keeda 4mo agoOh, not just them -- Satya, Jensen and I are all on a first name basis. They just don't know it yet ;-)
- matthewdgreen 4mo agoThe messed up scrolling behavior I keep getting in Claude Code is definitely due to their code.
- llbbdd 4mo agoThere is a setting that fixes this, I can't remember what it's called off the top of my head
- NichoPaolucci 4mo agoThis concept is so funny to me. Would love a toggle switch... "Oh yeah, just go to Settings > Bugs Enabled and turn OFF text display errors"
- oblio 4mo agoI've tried about 6 of those "settings" and hacks since November 2025 and not much luck.
- ashdksnndck 4mo agoCLAUDE_CODE_NO_FLICKER=1 This is a beta feature where Claude code draws the interface on the terminal’s alternate screen buffer like vim or htop. I believe it’s not the default because there are some potential compatibility issues deepening on your terminal setup. I’ve found it to be a nice improvement. It also fixed the issue where copy-pasting selected text from the terminal creates unwanted line breaks.
- matthewdgreen 4mo agoClaude Code is essentially a terminal emulator that runs on mature OSes with excellent support for this type of application. Why are they having difficulty implementing it?
- j2kun 4mo agoWe all saw their code...
- Melatonic 4mo agoThe whole thing is actually powered by a shitton of hamsters inside a bunch of 4u rack mount cases running on spinning wheels at high speed. Somehow at scale this works. Sometimes they all happen to randomly take a nap at the same time - hence the outages
- aagha 4mo agoAnd don't forget that they have BILLIONS of dollars and can't figure out how to get a decent support or public communications system setup.
- aleqs 4mo agoThey can't even seem to get their usage metering consistent.
- lukan 4mo agoYou mean on some days it goes faster and some other days slower? That is by design. It depends on how much other people are using their services right now and they do communicate it somewhere in the TOS that they do this. Otherwise they could give us a fixed amount of tokens - but they don't because it is not fixed.
- fc417fc802 4mo agoIf they implement demand pricing then they should be transparent about the current rate at any given time.
- thinkingtoilet 4mo agoDon't confuse things. It's not "can't figure out", it's "don't care to figure out". They're not dumb. They just don't care about support.
- contagiousflow 4mo agoCouldn't they just have background agents "figure it out"
- collingreen 4mo agoIf agents can just figure it out, isn't that AGI?
- f311a 4mo agoInfrastructure is a much harder problem. They can't even improve Claude Code, which eats 1GB+ of RAM. Meanwhile, my editor only consumes 80MB of RAM.
- andai 4mo agoTry 64K! https://en.wikipedia.org/wiki/Turbo_Pascal https://en.wikipedia.org/wiki/Turbo_Pascal Also remember when XP was super bloated cause it needed 64MB?
- airstrike 4mo agoThis might explain it, in the opposite way it was meant to: https://fxtwitter.com/trq212/status/2014051501786931427 https://fxtwitter.com/trq212/status/2014051501786931427 > Most people's mental model of Claude Code is that "it's just a TUI" but it should really be closer to "a small game engine".
- applfanboysbgon 4mo agoI hadn't seen that quote before, what an embarrassing thing to go on the internet and write...
- pragmatic 4mo agoSomebody read/watched too much Casey Muratori.
- CamperBob2 4mo ago
- qsort 4mo agoLook, I've never been someone who mindlessly hypes AI companies, as a matter of fact I think they have serious leadership problems across the board, but you people are straw-manning them so badly it actually makes me sympathize with them. They aren't saying they have fully automated luxury AGI, they specifically list the ways models fall short of that bar and caution against people taking the 8x figure as the actual uplift number. At the same time they recognize that 80% of new code is now AI-authored, when two years ago those models were little more than toys. And frankly that checks out: if two years ago you told me we'd have something like Opus 4.8/GPT 5.5 I would have rolled to disbelieve.
- sensanaty 4mo ago> At the same time they recognize that 80% of new code is now Al-authored I can setup a loop that will write a trillion lines of code automatically, how much of it is actually useful? Or are we back to counting LoC because there's no other metric for these systems that anyone can rely on?
- signatoremo 4mo agoI could write a bash script that copies a codebase repeatedly in the pre-AI past as well, but I didn't do that because I wasn't stupid. More than 80% of my code is now AI-generated, and trust me I'm still not stupid. It was 0% only a year ago. Who says LoC is the only metric we should rely on? A software product should first and foremost meet user requirements, functionality and performance. Judging from the sensational rise of Anthropic's user base and revenue I think we can safely says they're in that ball pack.
- jpleyden98 4mo agoIt's 80% of new code they shipped that is AI authored. Would you ship pointless code? I do tend to agree though, it could be that AI solves problems with more code than a human would. What you need to measure is the value the code brings and how much of that is done by AI, hard to get an objective measure of that though.
- 4mo ago
- claudiug 4mo agothose are results of the humans only. not the AI. AI is perfect /s
- rishabhaiover 4mo agoyou're conflating a compute problem with a code quality problem.
- belter 4mo ago[dead]
- asdfman123 4mo agoPersonally at my own job self-writing code is letting us tackle big, long-deferred refactoring projects (like the article mentions), but any sort of refactoring introduces new bugs.
- Quekid5 4mo agoIndeed... why is Anthropic even employing people at all if this AI magic story is true?
- drivebyhooting 4mo agoYou still need wizards to cast the spells..
- killbot5000 4mo agoNot if your spells cast their own spells.
- jimbokun 4mo agoRead the article. They are saying very clearly the models are not casting their own spells…yet. But looking at trends and speculating when they may start doing so.
- emp17344 4mo agoNot if you’re claiming that the spells, once cast, automatically get exponentially spellier until they awaken into a spell god, capable of literally anything, including casting more complicated spells than any wizard is capable of. If that were true, you’d have no need for wizards. The fact that wizards are still around means it’s probably bullshit.
- krapp 4mo agoWhat really happens is the spells only have other spells to draw from and they begin to degenerate over time, eventually turning into chaotic eldritch horrors that randomly add limbs to people or adamantly refuse to discuss goblins or just shriek in gibbering madness. Our Evil Overlord sacrifices the dreams of children to keep the magic sustained and controlled, and soon the people can't even think or speak without the help of magic. And they think they're wizards even though they can't even read a grimoire.
- 4mo ago
- rush86999 4mo agoJust as you expected, I'm throwing in my harness. Please support: https://github.com/rush86999/atom https://github.com/rush86999/atom
- 0xbadcafebee 4mo agoHave you considered just... using OpenAI? They are more reliable, models are just as good, and their subscriptions provide more requests per dollar.
- patcon 4mo agoNot necessarily the parent's fault, but the energy of this thread is not my favourite...
- 0x53 4mo agoThey also don’t have…a login page with authentication . To access the console you get an email link. No passkeys, passwords, 2fa, just an email.
- bluerooibos 4mo agoWell, people keep throwing money at them, including you and investors. So why would they care? It hasn't annoyed you or a large enough portion of users enough to move off their service - because there isn't a better alternative.
- windexh8er 4mo agoOpus 4.8's critical assessment of Anthropic's "When AI builds itself" [0][1]. Because, why not? [0] https://pastebin.com/Vc5Yq9Ai https://pastebin.com/Vc5Yq9Ai [1] https://www.anthropic.com/institute/recursive-self-improvement https://www.anthropic.com/institute/recursive-self-improveme...
- solid_fuel 4mo agoWhat does this add? Everyone in here is perfectly capable of prompting Opus for a writeup. Why don't you, windexh8er, try providing some thoughts of your own instead?
- windexh8er 4mo agoIrony, maybe? Do you not get it? If these models are so great solid_fuel then I guess it wouldn't be interesting that Anthropic's own models can make up ulterior BS as analysis. So why don't you pound sand since that clearly went straight over your head? That would be far more useful than your asinine response.
- prng2021 4mo agoWe’ve got a company of several thousand employees serving hundreds of millions of people arguably the best AI model in the market. Meanwhile you’re asking for a handkerchief for your pool of tears because their product is struggling to do your daily job functions for you, with much of that due to being limited by the worlds supply of silicon, electricity, water, and other resources. Cry me a river.
- z3c0 4mo ago> their product is struggling to do your daily job functions for you So what's the value prop?
- hombre_fatal 4mo agoThis comment is a good example of the double standard laymen have about AI usage: If you use AI, then AI must be expected to solve all problems, even problems that affect everyone like infra scaling. And if perfection isn’t delivered, then of course it wasn’t: you used AI and AI sucks.
- jayd16 4mo agoIt's not a double standard. Its being held up against the marketing.
- AnimalMuppet 4mo agoIf their AI is good enough to write their code, why isn't it good enough to tell them how to fix their infra? That's a different problem space, but it's not harder than the code.
- deleted 4mo ago[deleted]
- hombre_fatal 4mo agoThe software engineer inside us wants to believe otherwise, but scaling infrastructure is much harder than maintaining a TUI.
- weakfish 4mo agoAh, excuse me, I didn’t realize I was a mere layman.
- jatora 4mo agoThis is weird to me because i am using claude code 10+ hours/day 7 days a week, usually multiple sessions, and run into api errors maybe in 1 or 2 sessions per week. And about..2 major outages of 10-20min in the last month. Not terrible and nowhere near what you are reporting. Therefore I dont believe you, because you dont even couch this in terms of it being something that seems particular to you or your region. Obvious dishonestly is fairly bad of you.
- deleted 4mo ago[deleted]
- anjel 4mo agoAnswers the question: how can Anthropic sell more Usage "Credits"
- cookiengineer 4mo agoThe main reason I am building my own agentic environment is that I need full control and reproducibility of what I am building. Post November and post openclaw agentic environments need to be built differently, and for selfhosting models the context size problem really requires a strong harness which intelligently helps reduce context size. Planner/orchestrator architecture, agent to agent summarizer, specification based tools (fck all this markdown memory bullshit btw), tool call shrinking, and workflow management are all really important because of the context size problem. Nobody has enough VRAM for the large K/V caches, and nobody can afford f16/f32 caches in terms of memory, which are also necessary for longer conversations. MoE 30b models have improved so much though, qwen 3/3.6 coder is the real champion doing almost the same things with less than 1/10th the memory requirements. Just think about that in terms of engineering and what your bet is going to be. Haiku pales in comparison. Currently my focus with exocomp is trying to figure out how I can record, replay, restart, and debug workflow sessions of agents in a better manner so that I as a human can understand what's going on. Currently I think that UI will be something like a gantt chart where you have a graph with connections representing agent to agent communication. And yes, that's a lot of fiddling with SVG as it turns out, so I'm not quite there yet. Anyways, in case you're interested. I'm manually building this env and trying to unit test the critical parts. [1] [1] https://github.com/cookiengineer/exocomp https://github.com/cookiengineer/exocomp
- thordenmark 4mo agoGrowing pains of being successful. These are solvable problems and will be. Can they maintain their momentum without pissing off too much of their customer base before these issues are resolved?