Technology
OpenAI Claims GPT-5.6 Sol Outperforms Opus 5 on ARC-AGI-3

OpenAI Claims GPT-5.6 Sol Outperforms Opus 5 on ARC-AGI-3

the-decoder1h
THE BRIEF
1

OpenAI reports that its GPT-5.6 Sol model achieved a 38.3% score on the ARC-AGI-3 benchmark, surpassing the 30.2% score previously set by Anthropic's Claude Opus 5.

2

The results were achieved using OpenAI's proprietary 'Retained Reasoning' and 'Compaction' API settings, which differ from the standardized official test environment.

3

ARC Prize co-founder François Chollet noted that while provider-specific settings create parity issues, they are acceptable if costs and configurations are clearly reported.