FrontierThe story, in brief

Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One

Reka's Rho-1: one 19B model reads text, generates video, outputs robot actions—all in a shared KV cache. Distilled variant hits 5.3-second clips in ~1 second.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

A unified omni-modal architecture consolidating text, image, video, and robot-action reasoning in a single network is a noteworthy capability shift. No open weights yet, and it's research-preview status, so deployment impact remains unproven—but the model design (shared KV cache across modalities) is substantive enough for frontier tracking.

The key facts

6 to know
  1. Model size: 19B parameters

  2. Modalities: reads/generates text, images, video; outputs robot actions

  3. Architecture: single shared KV cache across all modalities

  4. Distilled variant: 5.3-second video clip generated in ~1 second

  5. Status: research preview, no public weights yet

  6. Unified reasoning: one network, not modular pipelines

The story so far

Earlier coverage of this storyline

  1. Reka AI's omni-model Rho-1 handles text, images, video, and robot control in a single modelThe Decoder
  2. This story

Go to the source

MarkTechPostmarktechpost.com

Publisher excerpt: Reka has released Rho-1, a 19B omni-reasoning model trained from scratch. One network reads and generates text, images, video and robot actions over a shared KV cache. A distilled variant returns a 5.3-second clip in about a second. It is a research preview with no public weights yet.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier