FrontierThe story, in brief

Addendum to GPT-5 system card: GPT-5-Codex

OpenAI just released GPT-5-Codex. Here's what dynamic thinking effort means for your engineering workflow.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI is releasing a specialized variant of GPT-5 optimized for agentic coding tasks, featuring adaptive compute allocation based on task complexity. This signals a shift toward task-specific model variants and dynamic inference as a competitive capability.

The key facts

5 to know
  1. GPT-5-Codex is a specialized variant of GPT-5 optimized for agentic coding

  2. Dynamic thinking effort allocation based on task complexity

  3. Responds quickly to simple queries/small tasks, allocates longer thinking for complex tasks

  4. Released as addendum to GPT-5 system card

  5. Published September 15, 2025

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: This addendum to the GPT-5 system card shares a new model: GPT-5-Codex, a version of GPT-5 further optimized for agentic coding in Codex. GPT-5-Codex adjusts its thinking effort more dynamically based on task complexity, responding quickly to simple conversational queries or small tasks, while…
Read original report
Back to today's editionMore frontier news

The wider picture

View all
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier01

Simulated students that make realistic mistakes help AI tutors learn faster

A novel approach to AI training using realistic synthetic feedback loops is accelerating tutor model development and reducing the cost of evaluation data. This represents a meaningful shift in how frontier labs can iterate on capability without massive labeled datasets.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier02

Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages

Alibaba's new Qwen3.8-LiveTranslate represents a meaningful advance in multimodal capability (speech-to-speech interpretation with low latency, speaker diarization, and long-context understanding), shipped and available now. It's a model release that changes the frontier benchmark for real-time translation and interpretation.

MarkTechPost
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier03

Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks

A new multimodal model from Alibaba's Qwen achieves near-parity with Google's flagship on audio-video tasks while undercutting its pricing, intensifying competition in the agent-native model race and raising questions about Google's cost positioning.

The Decoder