Thinking Machines Rolls Out Broad but Efficient Model
OpenAI's former CTO just shipped a model designed to do more with less. Inkling trades raw scale for token efficiency—and that matters when inference costs are eating margins.

Why it matters
Thinking Machines' Inkling represents a counter-narrative to the scaling-at-all-costs trend: a general-purpose model optimized for token efficiency. For founders and investors, this signals renewed competition on operational cost rather than capability alone—a shift that could reshape unit economics across AI applications.
The key facts
4 to knowThinking Machines founded by OpenAI's former CTO
Model name: Inkling
Positioning: general-purpose with token efficiency focus
Release date: July 16, 2026
Go to the source
AI Businessaibusiness.com
Publisher excerpt: The AI startup, founded by OpenAI's former CTO, released Inkling, a general-purpose model that keeps token use in mind.