Original Reddit post

I gave Opus 5 and Opus 5.5 the same five coding tasks in Claude Code and counted all the slop people complain about: Opus 5 vs Opus 5.5, over 20 replies each: Em dashes: 98 -> 0 Denying ramdom things (“this is X, it is Y”): 19 -> 5 Sentences with no verb: 24 -> 0 Metaphor words like “load-bearing”: 4 -> 0 Sentences you have to read twice to understand: 63 -> 6 Blind LLM judge picks for the best readable reply: Opus 5.5 in every message It is honestly refreshing to see this improvement when all the model providers are all trying to win the benchmark games. What is your experience with Opus 5.5? Made a video of it if you wanna look in details: https://www.youtube.com/watch?v=5iTeQSVNjlU submitted by /u/WingsJw

Originally posted by u/WingsJw on r/ClaudeCode