FrontierThe story, in brief

Thinking Machines bets on efficiency over size with its second model, Inkling Small

Mira Murati's Thinking Machines just shipped a smaller model that outperforms its predecessor on reasoning and code — a direct bet against the scaling narrative.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Open-weight reasoning models are becoming more efficient; Inkling Small's sub-linear scaling relative to its predecessor signals a shift in how frontier labs optimize for cost and capability tradeoffs. Practitioners evaluating reasoning models now have a new efficiency benchmark to consider.

The key facts

6 to know
  1. Thinking Machines released Inkling Small

  2. Inkling Small is less than one-third the size of the original Inkling model

  3. Inkling Small beats original Inkling on multiple coding and reasoning benchmarks

  4. Open-weights reasoning model

  5. Lab leadership: Mira Murati (former OpenAI CTO)

  6. Published 31 July 2026

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: Thinking Machines, the AI lab from former OpenAI CTO Mira Murati, has released Inkling Small. The open-weights reasoning model is less than a third the size of Inkling but beats it on several coding and reasoning benchmarks.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier