FrontierThe story, in brief

State-of-the-art video and image generation with Veo 2 and Imagen 3

Google just shipped Veo 2. Here's what it means for the video generation wars.

Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
New tools for building and creating with AI.AI illustration by KeyNews
The KeyNews take

Why it matters

Google DeepMind's Veo 2 and Imagen 3 updates represent a significant capability jump in generative video and image models, directly challenging OpenAI and other competitors in the multimodal generation space. For founders building video/creative tools, this raises the bar on what users will expect.

The key facts

5 to know
  1. Veo 2: new state-of-the-art video generation model

  2. Imagen 3: updates to image generation capability

  3. Whisk: new experimental feature/tool introduced

  4. Published Dec 16 2024 — recent release

  5. Google DeepMind source — credible, official announcement

Go to the source

Google DeepMind Blogdeepmind.google

Publisher excerpt: We’re rolling out a new, state-of-the-art video model, Veo 2, and updates to Imagen 3. Plus, check out our new experiment, Whisk.
Read original report
Back to today's editionMore frontier news

The wider picture

View all
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier01

Alibaba's open-weight Qwen-Image-2.1 claims to beat closed models in image generation with just 7 billion parameters

A capable open-weight image model at 7B parameters challenges the closed-model dominance in generation and editing, expanding practitioner options for on-device and cost-efficient image workflows.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier02

Tencent's Gander aims to keep talking while it works in the background

A novel architecture for multimodal agents that separates conversational continuity from task execution. Demonstrates a real capability tradeoff: smoother UX vs. task reliability. Relevant to how frontier labs are rethinking agent design.

The Decoder
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier03

Simulated students that make realistic mistakes help AI tutors learn faster

A novel approach to AI training using realistic synthetic feedback loops is accelerating tutor model development and reducing the cost of evaluation data. This represents a meaningful shift in how frontier labs can iterate on capability without massive labeled datasets.

The Decoder