OpenAIResearchMateriality 2

OpenAI says GPT-5.6 Sol tops Opus 5 on ARC-AGI-3 with API settings

OpenAI claims GPT-5.6 Sol scores 38.3 percent on ARC-AGI-3 using its own API features, versus 7.8 percent in the official test setup. ARC Prize says its environment is provider-neutral but may have used an outdated API, skewing the comparison with Opus 5.

First seen 30 Jul, 09:03 UTC on The Decoder2 sourcesLast update 52d ago

What the sources say

Linked, never rewritten. Official means the lab itself.

How it unfolded

Every source in the order it appeared. Times in UTC.