Original Reddit post

I’ve been playing with Opus 5.5 recently, and the quality of code that it spits out is pretty phenomenal. It doesn’t hallucinate as much as the previous models after 125k tokens in context window. What I’ve seen is that, and this is my observation (and it comes from my anecdotal experience), the amount of tokens that these newer models are consuming is much lower than the previous generation of models. Does anyone know what architectural changes allow them to achieve this sort of efficiency, or is it because they have very fierce competition and they just want to give tokens for free? I would love to hear other people’s thoughts if they are having the same kind of experience, but it seems like the newer model is faster and more clear in its chain of thought. Overall, it feels like a better bump than Opus 5. submitted by /u/Godcreatebugs

Originally posted by u/Godcreatebugs on r/ClaudeCode