I was in the middle of a pretty heavy development day, with some long-running, well-defined tasks queued up, when Opus dropped earlier. The way most tasks run in my workflow: Fable orchestrates the design with me and we segment the work. Fable then writes the briefs and dispatches Opus or Sol builders, who submit the PR locally for an adversarial review by the opposite model (OpenAI/Anthropic) — then CI, fixes, etc. I didn’t want to upset the applecart too much, but there was a natural pause when I updated the client, so I did what I usually do and made the main chat context aware that there was a new model in the workflow and to watch out for unexpected behaviour or any deviations from the established workflow. (The change from 4.7 to 4.8 totally borked a bunch of my setup.) I didn’t interact with the model much personally today, but as far as the output goes, it was noticeably more thorough on complex tasks and caught and rolled back more of its own mistakes than the previous model did. It caught maybe 70% of the issues that would normally surface during the first adversarial review. It also followed the agent instructions to escalate rather than improvise around issues much better. It’s a solid ass model. Here’s Fable’s review: submitted by /u/OofDaMae
Originally posted by u/OofDaMae on r/ClaudeCode
