Thinking Machines wants to build an AI that actually listens while it talks
Thinking Machines just broke the turn-taking bottleneck. AI that listens while it talks isn't sci-fi anymore.

Why it matters
A fundamental shift in model architecture from sequential turn-taking to simultaneous input-output processing could reshape latency expectations and user experience across all conversational AI products. This represents a capability advancement that competitors will need to match.
The key facts
4 to knowThinking Machines building model with simultaneous input processing and response generation
Current AI models use sequential architecture: user input → model processes → model responds → user listens
Proposed approach mimics real-time phone conversation vs. text-based exchange
Addresses latency and interactivity as core capability differentiator
Go to the source
TechCrunch AItechcrunch.com
Publisher excerpt: Right now, every AI model you've ever used works the same way. You talk, it listens. It responds, you listen. Thinking Machines is trying to change that by building a model that processes your input and generates a response at the same time, so it's more like a phone call than a text chain.