Fable can one shot a task from a good prompt while Opus 5 raises multiple questions and phrase them in such a dogshit way, and then completely butchers the output. How can we trust the benchmark anymore when Opus5 is shown to be close to Fable when it really aint even close submitted by /u/tnh34
Originally posted by u/tnh34 on r/ClaudeCode
You must log in or # to comment.
