FrontierThe story, in brief

OpenAI says its internal model solved over 100 long-standing math problems after just a month of training

100+ unsolved math problems. One month of training. OpenAI's internal model just reset what we thought was possible in automated theorem-solving.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI demonstrated a significant capability leap in mathematical reasoning—a frontier benchmark for reasoning models—but the company is managing external scrutiny by funding (but limiting) independent review at the Institute for Advanced Study.

The key facts

6 to know
  1. OpenAI internal model solved 100+ long-standing open math problems

  2. Training duration: approximately one month

  3. Company facing criticism from mathematicians on claims/methodology

  4. OpenAI backing independent advisory group at Institute for Advanced Study

  5. Company has excluded its research pace from advisory group's scope

  6. Suggests reasoning capability breakthrough, but governance/transparency questions remain

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: OpenAI says a new internal model solved more than 100 open math problems after just a month of training. Facing criticism from mathematicians, the company is backing an independent advisory group at the Institute for Advanced Study. But OpenAI has excluded its research pace from the group's…
Read original report
Back to today's editionMore frontier news

The wider picture

View all
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier01

OpenAI launches GPT-6 Sol and Luna, boasting lower cost and fewer mistakes

Two-model strategy signals OpenAI's bet on specialization over one-size-fits-all frontier capability. Practitioners choosing between cost and quality now have official paths; enthusiasts watch if this reshapes the lab-race playbook.

TechCrunch AI
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier02

Claude Opus 5.5 matches Fable 5.1 performance at lower cost and promises less "Claudish" writing

A new generation of Claude models arrives with meaningful cost reduction and claimed capability parity to Anthropic's previous flagship, while positioning against OpenAI's latest. This matters for practitioners choosing between models and for understanding the efficiency frontier in the lab race.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier03

Anthropic releases Opus 5.5 with lower prices and Fable-level performance

A new flagship model from a frontier lab claims best-in-class performance while undercutting rivals on price—a capability + economics shift that reshapes the competitive landscape and forces practitioners to re-evaluate their model strategies.

TechCrunch AI