‹ BackHN Continuity

Thread

Claude partial outage

171 points · 151 comments · sidcool

  1. netdevphoenix · · focus · HN ↗
    Imagine having a bunch of Claude agents doing time sensitive work and this happens and now a single human needs to pick up several 8k PRs authored by bots that can't talk to you now. Not sure how effective it would be to switch models using OpenRouter and the like given that, as the Earendil guys said, there is a lot of session detail that the provider doesn't share. The uptime is really going to cap their market penetration imo.
    1. psyphy2 · · focus · HN ↗
      can't you say the say with imagine "power goes out" or the "internet goes out"?

      (but yes if AI is going to propagate, they need to fix their uptimes)

      1. shaky-carrousel · · focus · HN ↗
        No, you can't. Because power or internet doesn't go out globally.
        1. ethagnawl · · focus · HN ↗
          Solar flares would like a word.
          1. shaky-carrousel · · focus · HN ↗
            Sure they would but they won't, because they don't affect power globally.
        2. dannyw · · focus · HN ↗
          Claude models don't go out globally either. Bedrock is fine. Azure is fine.
      2. lxgr · · focus · HN ↗
        Neither electricity, nor the Internet is provided by one or two organizations globally, both with a blast radius that quite often takes out all of their services.

        Not sure how long the latter will remain true – centralization is well underway there; try not to think of Cloudflare too much...

        1. 0x457 · · focus · HN ↗
          Globally no, but in the US there probably at most 2 providers covering your area + 5G/LTE providers. Electricity - single provider in the area. So if your job is on-site, then entire company stops working when either goes out. You don't really care that another company on the other side of the globe or street continues working.

          We already got used to "internal going down globally" when AWS was the main cloud provider. Now its 3, AWS, GCP, Cloudflare, but "portion of internet that you care about is down for the entire world" is a casual thing to happen now.

          1. lxgr · · focus · HN ↗
            > You don't really care that another company on the other side of the globe or street continues working.

            As somebody in need of a doctor, on a plane etc. I'd definitely consider it much worse if all hospitals, air traffic controllers etc. stopped working at the same time instead of just a local subset.

      3. PhearTheCeal · · focus · HN ↗
        anyone can buy a backup generator for power
      4. SupLockDef · · focus · HN ↗
        No worries, coding is solved they said in February 2026!

        Isn't that a good news!

    2. mococa · · focus · HN ↗
      They probably have an OpenAI subscription, same for OpenAI SWE, they have a Claude Code subscription
      1. ceejayoz · · focus · HN ↗
        They likely also have dedicated and separate clusters for their own use.
        1. mococa · · focus · HN ↗
          If so - for running Chinese models :trollface:
          1. ceejayoz · · focus · HN ↗
            I'm sure they run every model they can get their hands on in internal clusters.
      2. rcleveng · · focus · HN ↗
        When I worked at one of the early paging and cellular companies, every company had phones and pagers on at least one or two competitors for the oncall folks - you can't get a call to fix your cellular network when you are down.

        Likewise while at Google, they didn't let us put Google Voice numbers down as our oncall phone number. Some folks even had multiple phones on different networks as backups.

        I'm sure they have either other SOTA models or open source models around to help if things to completely pear shaped.

    3. sarjann · · focus · HN ↗
      Never thought of it, but knowing that they encrypt reasoning is it possible that they share the decryption keys with other cloud providers like GCP Vertex / AWS Bedrock?
      1. chinathrow · · focus · HN ↗
        The NSA wants a word.
      2. tekacs · · focus · HN ↗
        They don't, no (last that I looked) – if you switch providers, encrypted reasoning no longer works.
      3. cbhl · · focus · HN ↗
        They don’t even share the reasoning keys between *models in the same model family*.

        Opus isn’t even allowed to decrypt thinking blocks from Fable if you switch models mid-conversation.

        <a href="https:&#x2F;&#x2F;platform.claude.com&#x2F;docs&#x2F;en&#x2F;build-with-claude&#x2F;preserved-thinking#prefix-check" rel="nofollow">https:&#x2F;&#x2F;platform.claude.com&#x2F;docs&#x2F;en&#x2F;build-with-claude&#x2F;preser...

    4. maxcb · · focus · HN ↗
      &gt; now a single human needs to pick up several 8k PRs authored by bots

      PRs shouldn&#x27;t really be reaching 8k diffs no matter who authors them.

      1. reaperducer · · focus · HN ↗
        PRs shouldn&#x27;t really be reaching 8k diffs no matter who authors them.

        There&#x27;s an ocean between what &quot;should&quot; happen and what actually happens in tech today.

        1. [deleted] · · focus · HN ↗

          [deleted]

      2. Iolaum · · focus · HN ↗
        Given that CI is turning out to be a bottleneck - and a real pain for chainedd PR&#x27;s - I have my hesitations on this.
        1. tjwebbnorfolk · · focus · HN ↗
          yea, a lot of &quot;best practices&quot; are no longer best practices
    5. Alifatisk · · focus · HN ↗
      &gt; Imagine having a bunch of Claude agents doing time sensitive work and this happens and now a single human needs to pick up several 8k PRs authored by bots that can&#x27;t talk to you now.

      8k diffs PR is incredible, is it even sane for a human to comprehend that much?

      There has to be a name for such scenario were the handoff from Agent to Developer is so difficult, its not even worth pursuing but to instead wait until the Agent comes back online.

      No matter how good of a developer you are, we are still constrained to our minds working memory. Lots of people I have spoken to are aware of this. Agents produce lots of diffs in a fast pace, our minds don&#x27;t get the time to grasp the changed logic. And then, we an Agent is unavailable all of the sudden, say in the case of it being an outage or usage limit reached, the human now have to go through massive amounts of changes. They cannot continue as usual, there is a spike, they have to learn the new changes and grasp everything before they can continue towards the goal.

    6. dannyw · · focus · HN ↗
      If you&#x27;re not using Codex and Claude at the same time, you&#x27;re missing out heaps. Even without outages, OpenAI models are great reviewers for Claude; and vice versa.

      Other than trivial PRs; everything I do with Opus&#x2F;Fable gets reviewed by Astra; and everything I do with Astra gets reviewed by Opus&#x2F;Fable.

      Using only models from a single vendor, is like testing your website&#x2F;webapp only on Chrome.

      1. dingaling911 · · focus · HN ↗
        There is zero evidence for this &quot;use multiple models&quot; trope, IMO.

        It&#x27;s just feelings.

        1. dannyw · · focus · HN ↗
          There&#x27;s far more than zero evidence, see <a href="https:&#x2F;&#x2F;arxiv.org&#x2F;html&#x2F;2402.08806v1" rel="nofollow">https:&#x2F;&#x2F;arxiv.org&#x2F;html&#x2F;2402.08806v1 ; or the industry-standard practice of using multiple model families as LLM judges; or even 1P implementations <a href="https:&#x2F;&#x2F;code.claude.com&#x2F;docs&#x2F;en&#x2F;advisor" rel="nofollow">https:&#x2F;&#x2F;code.claude.com&#x2F;docs&#x2F;en&#x2F;advisor (which misses most of the benefit; since you want different model families, with different pretrains and posttrains).
          1. dingaling911 · · focus · HN ↗
            Eh that’s about medical diagnoses. Not directly transferable. For regular usage, multiple models just make you feel productive but I bet they aren’t any more productive than just one.
      2. pruzicka · · focus · HN ↗
        In my workflow, I still struggle to employ both providers in a way that makes sense and doesn&#x27;t require me passing data or prompts between the two.

        Could you share a bit more about how you do it?

        1. dannyw · · focus · HN ↗
          For me, commits are the atomic block; and a good commit should be self-contained anyway; where no additional context is necessary. If an agent can&#x27;t figure out what a commit is supposed to do, with only the commit title&#x2F;description and diff, it&#x27;s a bad commit, and this has always been true in software engineering. I do not pass prompts between the two ordinarily.

          After claude or codex finishes a commit, I switch console tabs and ask the other to review it; with the commit ID. I often find it helpful to inject a bit of human knowledge, and callout any areas of attention I see from a quick skim. (But that could just be me wanting to not abstract myself away from software engineering that much :)

          Sometimes, I do ask Codex to read my ~&#x2F;.claude&#x2F;; and vice-versa. But, generally, I try to keep as much knowledge (e.g. investigations, reports, deep dives) inside the git tree as possible; so that is not necessary.

          I don&#x27;t use skills, but I do have ~&#x2F;.codex&#x2F;AGENTS.md and ~&#x2F;.claude&#x2F;CLAUDE.md. These are high-level instructions for what (1) I consider readable, maintainable code and patterns, and (2) workarounds for empirically observed model behavior IOdon&#x27;t like, such as Astra being a bit of a &quot;over-correct over-validation nit-picker&quot;. I keep these human-authored, and update regularly based on what I find annoying.

          I do all of this before I submit a PR; but of course, for trivial stuff (e.g. CSS changes, copy&#x2F;string changes, etc), I don&#x27;t bother.

          My practices and workflows do change over time. Back in the ~Opus 4.5 days I&#x27;d often define a rubric&#x2F;criteria in a markdown file, iterate with AI to improve it, and that&#x27;s the &quot;spec&quot;. I&#x27;ve stopped doing that since GPT 5.6; partly because models have gotten a lot better at understanding high level intent from the context; and partly because nearly all models these days feel &#x27;gradermaxxed&#x27; when working like that.

          Finally, consumer $100&#x2F;$200mo subs get you _so_ far, I get a lot of value from both. I used to have multiple Claude subs for a while, but trying to &#x27;get full value&#x27; made me work on projects just for the sake of it; so 2x$200&#x2F;mo is my cap :)

          1. pruzicka · · focus · HN ↗
            thanks Danny :)
    7. [deleted] · · focus · HN ↗

      [deleted]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.