FrontierThe story, in brief

How enabling two settings tripled our scores on the ARC-AGI-3 benchmark

Two API toggles tripled GPT-5.6's score on ARC-AGI-3. Here's what changed.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI disclosed a significant capability unlock on a major reasoning benchmark by toggling reasoning retention and compaction settings — practical tuning that practitioners may need to adopt to match published performance claims.

The key facts

6 to know
  1. GPT-5.6 scores tripled on ARC-AGI-3 with two API settings

  2. Settings: reasoning retention and compaction enabled

  3. Benchmark: ARC-AGI-3 (abstract reasoning task)

  4. Source: OpenAI official blog post

  5. Date: July 29, 2026

  6. Implication: published benchmarks may depend on non-default configurations

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier