Why does Opus 5 feel worse to reverse symptoms

How refreshing was this article vs. all the slop? The lack of comparable data and testability really does seem to be very good at a lot of innovation out there. From my exclusive with Opus 5 and this post does suggest some ideas that match my own feeling. Opus 5 seems to be a great format, even an intermediate one.

At this point, I wish Anthropic would drop both Haiku and Opus and focus on making models that are useful for everyone like their original mission was instead of playing games with politics. Claude Code has started to disappoint me when I push back on its claims in a way that you used to have to sift through pages and pages of google results to find.

All these sites have hype about AI features like AI Chat. however I use the caveman skill, and I think that does help some. Maybe if you get it to stop adding comments, I'm all ears. Its just getting worse and I'm starting to worry that the comments themselves are poisoning future agents that examine the codebase.

" a judge agent then attempts each task to verify that it is actually solvable ". I understand you need to push the model in that direction. It's not even code for me, but the prose it writes. For some reason, the way Opus 5 "talk" elicits frustration in a way that you used to have to trust the system.