37 comments

  • niobe 23 minutes ago

    Well no one said it yet so I will, "international actors" is at least a possibility. And I don't mean any specific country because pretty much anyone is a potential these days, which makes it a perfect cover for different anyones. Demonstrating vulnerability in the US's AI boom can move the markets. That's a financial incentive and a strong geopolitical one.

    More likely just cascading overload though: "Never attribute to malice what can be explained by incompetence", or in this case, "growing as fast as possible"

  • docheinestages 20 minutes ago

    My gut feeling tells me it has something to do with Cloudflare. Along with AWS, they're two of the main suspects in such incidents.

  • Insanity 33 minutes ago

    Think of it like one big distributed system. OpenAI is down, so people migrate to Claude, now this one gets overloaded and goes down, etc.

    So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess.

      erdos_2 19 minutes ago

      It'd be funny if this is true because that'd prolly mean nobody is touching Gemini even as a fallback.

      fny 6 minutes ago

      I find it hard to believe that enough people would flock to from Claude and Chat to Grok to cause an outage. I feel like Gemini is the dominant release valve in this case especially for enterprise.

      paxys 7 minutes ago

      Especially considering memory/gpu/compute are scarce so these services are likely running with very little buffer.

      toomuchtodo 29 minutes ago

      https://en.wikipedia.org/wiki/Domino_effect

      Edit: Updated per valleyer's suggestion.

  • JackFr 3 minutes ago

    Obvious answer is it's the AI singularity. Been nice run for humanity. So long everyone.

  • dmillar 2 minutes ago

    Seeing 503s on Gemini via API as well

  • Linello 14 minutes ago

    What about a hard-takeoff scenario of an unleashed OpenAI Astra taking other models down for computational resources control?

  • codazoda an hour ago

    I kinda assume it's because one went down and a large amount of work shifted to another.

    I'm also aware that they have overlap in some areas on data centers.

  • faitswulff 11 minutes ago

    Heard on the grape vine that the OpenAI blip was a cloudflare issue

  • Avicebron 8 minutes ago

    I suspect Azure is having issues, Microsoft has had outages the paat two days, especially with email.

  • elar_verole 19 minutes ago

    Pretty sure it's a US thing since it's available here in France. What exactly is down, idk

  • chasd00 11 minutes ago

    claide.ai is working for me, so is chatgpt.com. grok still has a status message about issues, i can't try it without signing up.

  • jedbrooke 15 minutes ago

    according to https://downdetector.com/ Gemini is down too (and copilot, but that just uses ChatGPT right?)

  • maxbaines 41 minutes ago

    They all rent compute from SpaceXAI

      lavezzi 24 minutes ago

      I don't believe OpenAI does

      halcdev 35 minutes ago

      Surely it's a bit more distributed than that, right?

  • elorant 32 minutes ago

    Some npm library that makes headers bold would be broken.

      ibejoeb 10 minutes ago

      Oh man. Some low effort supply chain attack that turns every GPU into a cryptominer. It's funny because it's plausible.

      N_Lens 29 minutes ago

      Ah yes ye olde bold-headers: ^3.13.31;

  • CSMastermind 35 minutes ago

    I assume it cascaded from one provider to the other as people who lost claude access for instance moved to openai who moved to grok when it went down, etc.

  • kocial 41 minutes ago

    Maybe the stack behind it is down, like AWS or something

  • morkalork 8 minutes ago

    Didn't SpaceX overbuilt infra and leases it out Anthropic? I f their dc goes down it probably takes a chunk out of Claude's capacity before even considering the flood of users switching over

  • satvikpendem 18 minutes ago

    They're all using Cloudflare.

  • fidla 11 minutes ago

    chatgpt is back

  • aslkalska 11 minutes ago

    they all rent compute from each other

  • convivialdingo an hour ago

    The Thundering Herd has thundered, apparently.

  • fidla 11 minutes ago

    ChatGPT is up

  • wejick 34 minutes ago

    Probably same public cloud or CDN in front of them.

  • dgellow 30 minutes ago

    Too early to know, let’s wait and see

  • misano 12 minutes ago

    The IRGC has cut the fiber-optic cables in the Strait of Hormuz. LOL

  • Razengan 23 minutes ago

    SkyNet is arming..

  • ratelimitsteve 17 minutes ago

    everything in this thread is raw speculation, obv, but if i had to put money on anything i'd say this is a left-pad incident. some piece of something or other that all of these services happen to depend on went down. Second most likely seems to be some random failure of one leading to an unexpected traffic spike in others, though it seems like we've been talking about automated scalability in web apps for so long that there should at least be a response to, if not a solution for, this sort of problem.