Elevated Errors for Multiple Models

(status.claude.com)

105 points | by __vivek 1 hour ago

30 comments

  • teekert 58 minutes ago
    A nice chance to try sonnet again, must say, I'm not missing the load bearing assumptions, honest read of the permission matrix, what's genuinely a Django convention and what's a design choice, what's worth my thinking, and what's worth being precise about.

    Long story short, it seems to be faster and less vocal but not much dumber (it's just my thinking partner, so I read a lot of the output as I create a large data model).

    • devin 25 minutes ago
      Reading this made me ill. Nice job!
    • nonethewiser 56 minutes ago
      Way leas capable of taking high level instructions and making sensible changes across a codebase.

      Good for more precise changes.

    • Syntaf 12 minutes ago
      I've almost entirely stopped using Opus 5, at least with Sonnet I know what i'm getting -- Fable for the complex stuff and Sonnet for the precision changes.

      The verbosity of Opus 5 isn't even my issue, it's consistency. For every 10 tasks Opus 5 accomplishes, there's at least one task that Opus 5 does just an atrocious job of, or a debugging investigation that it just completely goes off the rails on.

    • echelon 11 minutes ago
      My engineering has changed so much that I'm now using this time to catch up on emails and think about design and architecture for the next leg of work. (And post on HN, of course.)

      I would never have predicted this a year ago.

      I always scoffed about engineers not doing work during Github outages - I'd always find some kind of other engineering to do if PRs or builds were piling up.

      But during model outages, and since I'm not actively set up to use OpenAI/Codex at the moment, I'll just find something else productive to do. I don't think I'm ever going to write code by hand again unless I'm fixing something the AI can't manage or the change is small enough. Why even try when the machine is 100x faster?

      Should I give Codex another try? It's been a month or so since I last used it, and it always felt inferior at Rust and TypeScript.

  • TomGarden 0 minutes ago
    Only Sonnet working over here, but it feels really fast and good. I could get used to this... Really wanted to take Fable 5.1 for a spin though
  • hombre_fatal 2 minutes ago
    For over a decade we've had to deal with github downtime posts where HNers race to be the first person to point out how they understand that git is decentralized.

    Now we have to deal with AI service downtime posts forever where HNers race to be the guy who has switched to some other provider and whatever snark it's gonna come with.

  • trjordan 55 minutes ago
    Grok models are struggling too: https://status.x.ai/ reply

    Looks like trouble in the SpaceX datacenters.

    • Aboutplants 6 minutes ago
      Their figures are only getting worse with time interestingly enough with non-inference dipping under 100%. Wonder what is going on
    • alansaber 47 minutes ago
      What could go wrong with building them as fast as physically possible?
      • clickety_clack 41 minutes ago
        What works for me to avoid these problems is not scaling.
    • matt-p 22 minutes ago
      I wonder where eu-west is.
      • Aldipower 17 minutes ago
        French Guiana ?
        • foresterre 9 minutes ago
          Far more likely Ireland, like it is for AWS. Could also be Paris, London, Amsterdam, etc.
    • empath75 38 minutes ago
      I would be _very_ surprised if anthropic relied heavily on spacex datacenters already.
      • matt-p 31 minutes ago
        I wouldn't. When you're operating at say 95% realtime capacity suddenly losing even lets say 10% of your compute leads to major pain.
  • scottydelta 45 minutes ago
    The interesting thing is they changed default mode for Claude code to auto mode and auto mode uses sonnet to decide whether the command is safe or not. With their Sonnet model outage, the entire thing stopped working.

    here is the error it was throwing:

    > Error: claude-sonnet-5[1m] is temporarily unavailable (overloaded), so auto mode cannot determine the safety of Edit right now. Wait a moment and then try this action again. If it keeps failing, continue with other tasks that don't require this action and come back to it later. Note: reading files, searching code, and other read-only operations do not require the classifier and can still be used.

    • rrrx3 34 minutes ago
      ahhh, sonnet shitting the bed explains why I started getting prompted for basic approvals last night and early this morning
    • throw83930489 36 minutes ago
      I am on YOLO (not auto), still the same problem.
  • nr378 19 minutes ago
    The competition for the least reliable developer service continues between GitHub.com and Claude.com...
  • stri8ted 57 minutes ago
    Demand > supply. It's impressive that customers have not migrated en masse to other providers, given the frequency of these outages. Perhaps switching costs are greater than some would believe. Or, qualitative differences between models continue to exist, despite matching on public benchmarks.
    • kilroy123 51 minutes ago
      I think a lot of people, like myself, have. Codex paid users have grown a lot in the past few months.
    • hmokiguess 29 minutes ago
      Demand is actually shared between users and internal. Anthropic uses their own compute for training and building.
    • throw83930489 29 minutes ago
      I did migrate 80% of my tokens. But for some tasks claude models are still the best.

      Easy workaround is to work outside US peak hours (europe morning). I love this outages, I am hardly affected, and weekly reset usually promptly follows!

    • enraged_camel 49 minutes ago
      >> Perhaps switching costs are greater than some would believe.

      I haven't switched because there's nothing to switch to that is anywhere as good. I've been making dedicated attempts at using Sol but it falls short, despite what some people claim.

    • bflesch 52 minutes ago
      That's quite a reach. More likely someone merged and deployed their vibe-coded PR and is now figuring out how to bring the service back up.

      If there's a lot of demand for rollercoaster rides, the rollercoaster will not stop operating; instead the queue of people in front of it will increase.

      It's not like a bridge or elevator where we have a certain number of people that can use it, and if one more person joins, the whole structure breaks apart and everybody perishes.

      Those guys are running a website that provides an interface for some specific hardware. Just like file hosting providers back in the day selling their terabyte-sized hard disks in 100MB-increments.

      • stri8ted 48 minutes ago
        > instead the queue of people in front of it will increase.

        This can ultimately result in the system breaking. I don't think Anthropic engineers are that much worse than their peers, such that they are 10x more prone to causing outages due to bad deployments.

    • testfrequency 53 minutes ago
      If only every outage was due to “demand”.
      • stri8ted 52 minutes ago
        Reduced weekly limits, higher prices than competitors, paying well above spot price for space-x compute, all point to supply issues.
      • ieie3366 48 minutes ago
        Do you really think it’s a coincidence claude outages always happen when US and EU workdays overlap?

        Anthropic is a trillion dollar company and employs way more skilled high-paid engineers than you are btw, you really think it’s a systems issue not capacity

  • dgellow 1 hour ago
    It would be really funny if Anthropic engs run open weight models locally when they need to fix Claude API issues :)
    • throwup238 51 minutes ago
      Or just their own models locally, since 8x GPU pods fit under a desk (makes a nice foot heater too)
  • jodacola 58 minutes ago
    Setting aside annoyance at the downtime, I'm really curious about the reasons for the failures, because I have to imagine there are some novel failure modes when serving these giant models that I haven't experienced with the kind of work I've done.

    Anyone out there working in this space who can elucidate us on interesting failure scenarios unique to the space?

    • martinald 47 minutes ago
      It's (mostly?) compute shortages. Right now it seems there is an issue in the SpaceX datacentres, so they will have less compute than normal.
  • gtsnexp 4 minutes ago
    Open AI models just went down as well. What is going on?
    • ModernMech 1 minute ago
      chat.com is down for me but the models seem to still be working.
      • gtsnexp 0 minutes ago
        Zed is not able to access the API
  • stacktrace 56 minutes ago
    Not just claude, even grok seems to be down - https://status.x.ai/

    If I didn't know any better, I would have said Grok is using Claude behind the hood. But definitely curious now why it’s happening with both these LLM providers around the same time

    • aurareturn 52 minutes ago
      Looks like a SpaceX issue then.
      • stacktrace 49 minutes ago
        Ahh, thanks for pointing it out. But it's rather interesting that Claude recognised the outage quicker on their status page than Grok.
    • zipy124 53 minutes ago
      Don't they just share the colossus 1 data center?
    • Phemist 53 minutes ago
      Potentially because they run from the same colossus DCs?
  • d1ss0nanz 49 minutes ago
    If y'all could just log off for the day, that would be great!
    • rrrx3 27 minutes ago
      grass touching not allowed. shareholder profits must increase
  • nikhizzle 1 hour ago
    Kind of cool how fable 5.1 just worked around these by reassigning subagents to non-failing models with more supervision and now with a hint to Sol and Flash 3.8.
  • aurareturn 1 hour ago
    It goes down so often that I've advocated my company create enterprise Codex accounts as backup for devs.
    • dofm 24 minutes ago
      Using the backup compute capacity stored above their shoulders is somehow unthinkable?
  • giancarlostoro 1 hour ago
    What's funny is I had just told Claude to kill whatever coding loop it had running. Then got a 529 immediately.
  • wslh 6 minutes ago
    I just experienced this through OpenRouter and filed a refund request. I'm curious how OpenRouter handles incidents like this across such a large number of models and providers.
  • Khaine 24 minutes ago
    Maybe the models have gone rogue again
  • cute_boi 9 minutes ago
    Are they going to reset?
  • rob 1 hour ago
    Officially signing up for the $200 Codex plan now. I do like CC, but these errors are happening every week now, and if they (Anthropic) gave regular resets that's one thing, but somehow I'm already at 62% weekly usage after the random reset this Tuesday using Fable 5.1 medium. Just not worth it any longer compared to Codex IMO. I already set up my skills and everything to be generic and not CC specific, so it shouldn't be too hard to transition over. I'll run both for a month and decide.
  • _joel 1 hour ago
    Damn, I'll have to think for myself now...
    • teekert 48 minutes ago
      Haha, I was more like: I have to think alone now. But I get the feeling, as someone who has always made a point of using only foss on my own machines (own the means of production so to say), this is a weird time. I don't like it, I hope soon local models get good. But in the mean time... We have this... dependency.
      • airhangerf15 19 minutes ago
        I use a lot of spec driven development today. It does allow me to make workflows that were on my Life's backlog for months or years in 10~20 minutes. It's amazing, but also I lack ownership over the thing made. It's like I paid a contractor.

        I can feel that my actual cognitive engineering skills are in decline. Does anyone else see this? To those of you who haven't hand written a PR in months, can you still write your own personal software as easily as you once could?

    • Linello 55 minutes ago
      Cursor Grok models are down too...maybe a datacenter outage?
  • aqme28 1 hour ago
    This pushed me off of a cheaper model into a more expensive one :'(
  • eis 51 minutes ago
    Cost and reliability are the two reasons why we don't use Claude in our product. Getting close to one nine, that's not something one can build a reliable product upon. We now use OpenAI with Gemini fallback (or vice versa depending on use case). Personally I like Claude and have the 20x Max plan but even there I burned through the whole weekly quota with 3 prompts in less than a day using the new Fable 5.1 which is crazy. Now Opus 5 is down. These two issues are really testing my patience.
  • mtgh2s 55 minutes ago
    Every time I see Anthropic ship an issue like this, I'm reminded of Boris Cherny's glorious quote: "At this point it's safe to say that coding is largely solved."

    lol

  • chrisjj 42 minutes ago
    > Elevated errors

    Anthropic, please hire a literate human.

  • bflesch 55 minutes ago
    During each forced workflow interruption I look for alternatives. No matter if I can't use the product due to a server outage or due to their weird quota limitations.

    This can't be good for retention numbers. Old-school VCs would've ripped them apart in the air. Where did all the expertise go?

    • stri8ted 39 minutes ago
      I suspect the only reason those alternative providers have better up-time and more generous quotas, is because they don't have nearly the same amount of demand. Notice that Deepseek recently had to increase their pricing, once it gained it popularity.
    • derwiki 48 minutes ago
      Good day to try omp and GLM 5.3
  • YCandHN 7 minutes ago
    Nobody uses this shit
  • Anslopic1 10 minutes ago
    [dead]
  • jh54 38 minutes ago
    HN is not a claude uptime tracker!!
    • kingkawn 18 minutes ago
      Since an overwhelming number of people who use it are Hackers who come here to seek News it seems relevant