Original Reddit post

Fable can one shot a task from a good prompt while Opus 5 raises multiple questions and phrase them in such a dogshit way, and then completely butchers the output. How can we trust the benchmark anymore when Opus5 is shown to be close to Fable when it really aint even close submitted by /u/tnh34

Originally posted by u/tnh34 on r/ClaudeCode