FrontierThe story, in brief

ByteDance's Seedance 2.5 generates 30-second video clips with built-in audio

ByteDance's Seedance 2.5 generates 30-second video with native audio — 3x longer than Gemini Omni Flash.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

A multimodal model capability leap: native audio-video generation and 3x context length vs. the frontier competition signals where the lab race is moving. Practitioners building video tools need to reset their baseline.

The key facts

5 to know
  1. Seedance 2.5 produces up to 30-second video clips with integrated audio generation

  2. 3x longer output than Google's Gemini Omni Flash

  3. Accepts dozens of images, videos, and audio files as reference inputs

  4. Native audio-video generation (not post-processed)

  5. ByteDance product — direct competition with Google, OpenAI multimodal capabilities

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: ByteDance just shipped Seedance 2.5, an AI video model that produces video and audio together in one go. Each clip runs up to 30 seconds, three times what Google's Gemini Omni Flash puts out. Users can feed in dozens of images, videos, and audio files as reference. For ad teams, this could kill the…
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier