Is it just me, or has Claude Opus gotten worse recently?
Things that used to work cleanly in a single prompt now fail completely. In coding workflows, it constantly ignores mandatory CLAUDE.md project rules, makes unsolicited edits to unrelated files, and breaks working code. It starts arguing based on stale comments, fails to update documentation when code changes, and edits based on blind guesswork instead of actually verifying the codebase first.
I even caught myself using Fable for tasks that I always used Opus for in the past, just to get decent results.
It feels like Anthropic is shifting their technical problems onto paying users. Either they are using strict safety filters that silently fall back to cheaper models without telling us, or they are downgrading the compute under heavy load to save money.
The result is the same kind of shrinkflation: we pay the same subscription price, but we either burn way more Opus tokens on endless retries, or we are forced to spend money on more expensive Fable tokens to get the quality we used to have.
Has anyone here successfully moved away from Claude for complex engineering workflows? What are you replacing it with, and how are you handling the transition?
Disclaimer: As a native German speaker, I used Gemini to clean up and polish the English text for this post.
25 comments
[ 0.26 ms ] story [ 39.0 ms ] threadI canceled my subscription and moved back to ChatGPT and happier so far.
I use gemini for rewriting code docs, because the frontier models are so verbose when it comes to writing text
I am now wondering
Lets say it has gotten worse? What are you going to do about it? Jump to Codex? Then what happen if you perceive that to be getting worse? Jump back to Claude? One of the many problems with these tools.
It made half the changes, committed and pushed, then it made the other half of the changes on the same two files as before and... just stopped and reported back with a cheerful "all food, all done".
I fear this is just the classic "nerf the model just before we release a new version of it"
For editing files, it again seems to prefer writing small python scripts, rather than read the file and rewrite it (even small files, like notes). This again causes it to constantly miss duplications, inconsistencies, etc.
And from 4.8 to 5, it doubled down on this. At this point not moderately-long-horizon task is getting done consistently, due to its pigeon-hole view of the workdir.
Anybody else notice this?
I think I noticed a huge deterioration starting at around May or late April. Not really sure what happened but quality definitely dropped