Original Reddit post

One would assume this would be easy since you could rerun the same evals from release day. Has anyone ever produced any proof that Anthropic or ANY frontier AI provider has “nerfed” a model after release? Everyone says it and I’m trying to steelman it first before I attribute it to psychological phenomena. submitted by /u/jimmc414

Originally posted by u/jimmc414 on r/ClaudeCode