Original Reddit post

Why does Opus 5 have so many bugs when coding now? I’m literally just doing an OPSD training task and using Opus 5 to help me write the training scripts. From the very beginning, there were problems with dataset filtering, then abnormal rollouts caused math verify to hang but it went off and tried to fix the vLLM parallel evaluation instead, and later there was student/teacher token misalignment, exploding thread counts, and so on. In the past two days alone, I’ve run into several times more bugs than I did in the previous six months of using Claude combined. I ended up having to use Codex to review the code, and eventually I was checking it line by line myself. I’m honestly starting to feel like I’d be better off just writing it all by hand. What is going on? And yes, of course I know I can switch to Codex whenever I want, and that’s what I’m planning to do now. But what if I hadn’t used Codex to review the code and had simply trusted Claude, letting those incorrect loss signals silently update my model? I feel like I should at least be getting a level of service that matches the $200 I’m paying every month. I genuinely don’t see even the slightest sign that Opus 5 is anywhere close to Fable 5 in intelligence. By the way, my setting is xhigh. submitted by /u/tianhe_2003

Originally posted by u/tianhe_2003 on r/ClaudeCode