FrontierThe story, in brief

GPT-5 System Card

GPT-5 isn't one model anymore. OpenAI just split it into three: main, thinking, and nano. Here's what that means for your inference costs.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI has architected GPT-5 as a multi-tier routing system, enabling developers to optimize latency and cost by selecting appropriate model variants for different task complexities—a strategic shift toward efficient model deployment at scale.

The key facts

5 to know
  1. GPT-5 split into three variants: gpt-5-main, gpt-5-thinking, gpt-5-thinking-nano

  2. Unified model routing system for task-specific optimization

  3. Lightweight versions available for cost-sensitive workloads

  4. Published August 7, 2025 on OpenAI official channel

  5. Focus on latency and performance differentiation across model tiers

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: This GPT-5 system card explains how a unified model routing system powers fast and smart responses using gpt-5-main, gpt-5-thinking, and lightweight versions like gpt-5-thinking-nano, optimized for different tasks and developer use.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier