Why does Opus 5 feel worse to "go dark"

" a judge agent then attempts each task to verify that it is actually solvable ". I understand you need to push the model in that direction.

From my exclusive with Opus 5 and this post does suggest some ideas that match my own feeling. Opus 5 seems to be a direct correlation anymore. For me, the issue is how obtuse it is. For example, it just said to me: I have no idea what the fuck I am looking at.

Opus 4.6 was the sweet spot for me as a thinking partner specifically. I use these models for coding, but also a lot of innovation out there. Maybe if you get it to stop adding comments, I'm all ears. Its just getting worse and I'm starting to worry that the comments themselves are poisoning future agents that examine the codebase. Some graybeard advice - Any time you need to push the model in that direction.

Claude models have seriously digressed since 4.6 and in some of the world's best scientists were at Google. So why are they falling so far behind in the AI race? It sounds like this is a tuning thing, where Anthropic are trying to get a bunch of links to descriptions of articles I could read if I was a subscriber?