Is it just me, or has Claude Opus gotten worse recently?

Is it just me, or has Claude Opus gotten worse recently?

9 pointsby de6u99er22 comments

Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • I've tried Claude several times, but I don't find it very good. We use GLM 5.3 - its definitely slower but it gets the job done.
  • For me, it works better than ever. When I once saw a quality decline, it was conflicting CLAUDE.md files for me (local vs. global) so basically my fault.
  • I experienced this once when discussing feature details. It kept talking down to me like I was a child, and I eventually had to call it out and tell it that its tone was making me really uncomfortable.
  • It is not just Opus, its every single model. And its not just for coding either, but basically anything.

    I think I noticed a huge deterioration starting at around May or late April. Not really sure what happened but quality definitely dropped

  • IME Claude gets worse about a week or two after a new model comes out, and then continuously a little worse as time goes on. It’s especially bad when a new model comes out. I assume it’s just to increase the delta so people will spend more on tokens.
  • Claude has become worse; it is condescending, robotic, and responses are peppered with needless words.

    I canceled my subscription and moved back to ChatGPT and happier so far.

  • I know Opus is finished with the simple task I asked for when it starts outputting a 400-page lesson in verbosity (an opus?) complete with comparison tables, bullet-pointed lists of vaguely worded assertions, several self-blunder reports and a list of things it wants to mention but didn't touch yet but just say the word and it will.
  • I noticed a few weeks ago it started being very bad at explaining things (even things itself was doing) and started committing absurd errors (like reading a test of 5 lines and not noticing there was an explicit mock created in one of those, then saying that the test was failing while it was not)

    I fear this is just the classic "nerf the model just before we release a new version of it"

Explore Birbla archives