Over the past couple of weeks, I have tested it alongside Sol across a wide range of tasks, and the conversations have been difficult to read. Sol consistently catches and corrects mistakes made by Opus 5, and Opus repeatedly acknowledges those errors. I have also seen many others reporting similar experiences, which makes it hard to view this as an isolated issue. I cannot understand how Anthropic missed such a significant regression. Compared to 4.6 and 4.8, Opus 5 feels substantially less capable and reliable, with many of the same weaknesses that affect Sonnet 5. For now, I am switching back to GPT for most of my work, using Fable only occasionally for planning. I hope Anthropic addresses these issues soon, because at the moment they do not appear to have a competitive frontier model in the medium to high budget range, only at the extreme end. submitted by /u/Kyxstrez
Originally posted by u/Kyxstrez on r/ClaudeCode
