Ask HN: Why are OpenAI, Claude, and Grok simultaneously down? Coincidence?

https://status.openai.com https://status.claude.com https://status.x.ai

63 points | by halcdev 54 minutes ago

27 comments

  • niobe 20 minutes ago
    Well no one said it yet so I will, "international actors" is at least a possibility. And I don't mean any specific country because pretty much anyone is a potential these days, which makes it a perfect cover for different anyones. Demonstrating vulnerability in the US's AI boom can move the markets. That's a financial incentive and a strong geopolitical one.

    More likely just cascading overload though: "Never attribute to malice what can be explained by incompetence", or in this case, "growing as fast as possible"

  • JackFr 0 minutes ago
    Obvious answer is it's the AI singularity. Been nice run for humanity. So long everyone.
  • Insanity 30 minutes ago
    Think of it like one big distributed system. OpenAI is down, so people migrate to Claude, now this one gets overloaded and goes down, etc.

    So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess.

    • erdos_2 16 minutes ago
      It'd be funny if this is true because that'd prolly mean nobody is touching Gemini even as a fallback.
      • Insanity 10 minutes ago
        Lol I didn't even think about Gemini missing from the list. Not sure what that says about Gemini or me :)
      • rtcoms 13 minutes ago
        Just now I got this from gemini

        It looks like there's no response available for this search. Try asking something else.

        • exe34 5 minutes ago
          I bet they had to implement that manually to make it look like they failed too!
      • benatkin 9 minutes ago
        Not even the best agent that starts with a G
    • fny 3 minutes ago
      I find it hard to believe that enough people would flock to from Claude and Chat to Grok to cause an outage. I feel like Gemini is the dominant release valve in this case especially for enterprise.
    • paxys 4 minutes ago
      Especially considering memory/gpu/compute are scarce so these services are likely running with very little buffer.
    • toomuchtodo 26 minutes ago
      https://en.wikipedia.org/wiki/Domino_effect

      Edit: Updated per valleyer's suggestion.

      • valleyer 21 minutes ago
        "Domino effect" would probably be the more relevant named phenomenon there.
      • throwaway894345 20 minutes ago
        This isn’t a thundering herd problem, it’s a cascading failure. (Thundering herd is about a bunch of workers waking up simultaneously)
  • docheinestages 17 minutes ago
    My gut feeling tells me it has something to do with Cloudflare. Along with AWS, they're two of the main suspects in such incidents.
  • Linello 11 minutes ago
    What about a hard-takeoff scenario of an unleashed OpenAI Astra taking other models down for computational resources control?
  • Avicebron 5 minutes ago
    I suspect Azure is having issues, Microsoft has had outages the paat two days, especially with email.
  • faitswulff 8 minutes ago
    Heard on the grape vine that the OpenAI blip was a cloudflare issue
  • codazoda 48 minutes ago
    I kinda assume it's because one went down and a large amount of work shifted to another.

    I'm also aware that they have overlap in some areas on data centers.

  • elar_verole 16 minutes ago
    Pretty sure it's a US thing since it's available here in France. What exactly is down, idk
  • chasd00 8 minutes ago
    claide.ai is working for me, so is chatgpt.com. grok still has a status message about issues, i can't try it without signing up.
  • jedbrooke 13 minutes ago
    according to https://downdetector.com/ Gemini is down too (and copilot, but that just uses ChatGPT right?)
  • maxbaines 38 minutes ago
    They all rent compute from SpaceXAI
    • lavezzi 21 minutes ago
      I don't believe OpenAI does
      • maxbaines 19 minutes ago
        My mistake, in fact it was google not OpenAI, makes sense OpenAI doesn't.
    • halcdev 32 minutes ago
      Surely it's a bit more distributed than that, right?
  • elorant 29 minutes ago
    Some npm library that makes headers bold would be broken.
    • ibejoeb 7 minutes ago
      Oh man. Some low effort supply chain attack that turns every GPU into a cryptominer. It's funny because it's plausible.
    • N_Lens 26 minutes ago
      Ah yes ye olde bold-headers: ^3.13.31;
  • CSMastermind 32 minutes ago
    I assume it cascaded from one provider to the other as people who lost claude access for instance moved to openai who moved to grok when it went down, etc.
  • morkalork 5 minutes ago
    Didn't SpaceX overbuilt infra and leases it out Anthropic? I f their dc goes down it probably takes a chunk out of Claude's capacity before even considering the flood of users switching over
  • kocial 38 minutes ago
    Maybe the stack behind it is down, like AWS or something
  • aslkalska 8 minutes ago
    they all rent compute from each other
  • satvikpendem 15 minutes ago
    They're all using Cloudflare.
  • fidla 8 minutes ago
    ChatGPT is up
  • convivialdingo 49 minutes ago
    The Thundering Herd has thundered, apparently.
  • wejick 31 minutes ago
    Probably same public cloud or CDN in front of them.
  • dgellow 27 minutes ago
    Too early to know, let’s wait and see
  • fidla 8 minutes ago
    chatgpt is back
  • misano 9 minutes ago
    The IRGC has cut the fiber-optic cables in the Strait of Hormuz. LOL
  • Razengan 20 minutes ago
    SkyNet is arming..
  • ratelimitsteve 14 minutes ago
    everything in this thread is raw speculation, obv, but if i had to put money on anything i'd say this is a left-pad incident. some piece of something or other that all of these services happen to depend on went down. Second most likely seems to be some random failure of one leading to an unexpected traffic spike in others, though it seems like we've been talking about automated scalability in web apps for so long that there should at least be a response to, if not a solution for, this sort of problem.
  • pwyq 27 minutes ago
    [dead]