FrontierThe story, in brief

Tencent's Gander aims to keep talking while it works in the background

Tencent's Gander splits the brain: one part keeps you talking, another handles the work. Interruptions drop to 8%—but accuracy still lags.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

A novel architecture for multimodal agents that separates conversational continuity from task execution. Demonstrates a real capability tradeoff: smoother UX vs. task reliability. Relevant to how frontier labs are rethinking agent design.

The key facts

5 to know
  1. Tencent's Gander uses 'cerebellum' for speech/conversation and swappable 'brain' for background tasks

  2. Handles speech, images, and text multimodally

  3. 8% interrupt rate in benchmarks, better than rivals

  4. Task accuracy trails competitors

  5. Users can interrupt or redirect mid-conversation

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: Tencent's Gander processes speech, images, and text while handling tasks in the background. A "cerebellum" keeps the conversation going, while a swappable "brain" searches files, writes code, or tackles other complex work. Users can interrupt or change the task mid-conversation. In benchmarks,…
Read original report
Back to today's editionMore frontier news

The wider picture

View all
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier01

Alibaba's open-weight Qwen-Image-2.1 claims to beat closed models in image generation with just 7 billion parameters

A capable open-weight image model at 7B parameters challenges the closed-model dominance in generation and editing, expanding practitioner options for on-device and cost-efficient image workflows.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier02

Simulated students that make realistic mistakes help AI tutors learn faster

A novel approach to AI training using realistic synthetic feedback loops is accelerating tutor model development and reducing the cost of evaluation data. This represents a meaningful shift in how frontier labs can iterate on capability without massive labeled datasets.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier03

Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages

Alibaba's new Qwen3.8-LiveTranslate represents a meaningful advance in multimodal capability (speech-to-speech interpretation with low latency, speaker diarization, and long-context understanding), shipped and available now. It's a model release that changes the frontier benchmark for real-time translation and interpretation.

MarkTechPost