Comment by atraac

4 days ago

Great that there's a new model but they could fix their existing infra. We're considering dropping our Claude Team sub cause it's unusable recently. Constant bugs, dropped sessions, issues switching models, http errors. It's becoming ridiculous

Is it some Claude Team/Enterprise only problematic? I'm using two 20x Max accounts almost non-stop (Fable/Opus) for 1.5 years at this point, zero issues with both client and infra sides (from US and in travels). When I'm reading such messages it feels like either I'm lucky or it's a part of some campaign.

  • I’m on the biggest max plan. It is riddled with annoying bugs for me, only been using it for a little over a month. Settings screen flashes randomly. But most annoyingly: sometimes when forking chats or sometimes for no reason, the UI just straight up eats my previous messages. The model is still aware of them and can recount them if I ask but the visible history is gone. And that’s not even all of them. Fable 5 is just too good that I put up with it but it seriously raises concerns for me that even with infinite compute these companies can’t even deliver a functional chat UI.

    • I've been on one to two of these plans for eight months or so. There used to be a lot of issues with CC's terminal but at least in iterm2 they have largely been sorted.

      What terminal tooling are you using?

      1 reply →

  • I commend you for burning 6 figure losses into Anthropic thanks to their subsidies singlehandedly but I am curious to learn what you've built with this so far, not to be snarky, I just wonder what people actually produce while having these run nonstop.

  • No idea but we have few Team Premium seats and everyone is encountering issues daily for past two weeks. From straight up outages to vscode extension/CLI refusing to process messages. It's been unusable for most of our work hours past two days.

  • I’m using both side by side, my employer pays by the token and I have a Max 20x plan.

    They start hitting timeouts or API errors at the same time on two different computers. As far as I can tell it’s the exact same infrastructure.

  • Over the last 1.5 years, they had a few failures with their auth system; two or three times auth failed for a few hours so I could not work. But otherwise they have been just fine. Some issues, but not significant.

  • I think you’re just lucky. Look at the Claude status page to see just how often they have outages (it’s almost daily). Even most of the green days have issues if you hover over them, they just don’t count them as outages.

Funny that a company selling an AI software developer can't use it to fix their infra.

Fixing those issues still requires humans.

  • How many companies at the size of Anthropic can serve the amount of traffic and manage the amount of compute they have?

    • It doesn't matter. Front end code should not misbehave if servers can't keep up. At worst it should fail gracefully.

    • How many compagnies can manage a mostly stateless workload at "whatever-the-scale-because-it-does-not-matter-because-stateless" ? Lots of people can do that. Massive amount of people can do that.

      2 replies →

  • Let's be honest - they're also still hiring software devs. AI still requires skilled humans in the loop and that's not going away.

    • It is going away for non tech companies though.

      Imagine you are a company that sells concrete. You have a web dev contractor you use to build and maintain your website. It has tools on it to get delivery quotes and a few internal tools to track orders.

      Except now you can just have your sales team also maintain the website with a $20/month Claude subscription.

      3 replies →

    • It's funny how they are at a disadvantage because they feel obligated to AI-max. Would Claude Code, as an interface, be as mediocre if they had software engineers writing its code directly? I doubt. On the other hand - how embarrassing would it be if they sold you a tool to write code but they were careful not to use it too much on their own products?

      1 reply →

  • This is just like any extreme engineering domain. I am okay with occasional delays in Flights, as long as it takes me from X to Y in 10hrs vs months.

  • i know the dream for capitalists is to be able to point an llm at something and say "do and/or fix it" but we still can't even get them to not go quite literally insane if allowed to run for an extended period of time

    and you can only kill weyoun, awaken the next vorta clone and have him 'catch up' on all that its missed so many times before they just end up with a complete mess, so. uh. yeah.

    doubt they can just "fix" their problems like that.

    • That one DS9 episode where there were two Weyouns at the same time is a good analogy for two agents working on a codebase at the same time (as in they don’t work well together).

> coding is largely solved

- Boris

  • The code it outputs, yes! It's fantastic. It's just so frustrating that the product and UX before the code output is so bad. Greatness is so close within their reach, if only they invested in product and QA people.

  • Fixing such issues requires software and site reliability engineering, of which coding is just a part.

    The first page of the score card mentions that this model is not capable to replace engineers.

> Constant bugs, dropped sessions, issues switching models, http errors. It's becoming ridiculous

And memory leaks.

I applied to their reliability team but never heard back. I would love to help them solve this problem!

I found the same thing funny with Computer Use from OpenAI. It struggled to open and close Spotify.

So dangerous! I can't believe they let the public use this technology! /s