OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settings
OpenAI counters Anthropic's ARC-AGI-3 record: GPT-5.6 Sol scores 38.3 percent, but only with its own API features instead of the official test setup, where the model landed at 7.8 percent. ARC Prize claims its test environment is provider-neutral, but may have used an outdated API that skewed the comparison with Opus 5. The article OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settings appeared first on The Decoder .
Record details
Published: 30 July 2026
Source: The Decoder
Category: News
Topics: Environment
Retrieved: 31 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems
The Decoder · 31 July 2026
Investor pressure forces Nvidia to shrink its OpenAI bet just as Anthropic's numbers defy bubble warnings
The Decoder · 15 August 2026
OpenAI open-sources Codex Security CLI to help developers find and fix vulnerabilities from the command line
The Decoder · 29 July 2026
US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns
The Decoder · 26 July 2026
Alibaba's new Qwen model is also taking your job, but this time it's great
The Decoder · 3 August 2026
Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion
The Decoder · 22 July 2026
How to cite this record
ethics.ai (30 July 2026), “OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settings,” evidence record 15161, https://ethics.ai/record/15161 (originally published by The Decoder).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.