Original Reddit post

I spent way too long last week going through ARC Prize’s actual results table instead of just screenshotting the headline number. Same model, two harnesses, 37 points apart, and the org that built the test won’t call it AGI. Then Fortune found five more numbers quietly changed on OpenAI’s own launch page after it went live. Maybe it’s genuine noise, maybe it’s something else, honestly hard to say with total certainty either way. Wrote up the whole harness breakdown plus the Llama 4 precedent. Here’s what I found. https://srutiosocial.com/gpt-6-astra-arc-agi-3-score-explained/ submitted by /u/mixtapedmonk

Originally posted by u/mixtapedmonk on r/ArtificialInteligence