WorkThe story, in brief

Researchers define what counts as a world model and text-to-video generators do not

NOBODY TALKING: While everyone obsesses over Sora, researchers just redefined what actually counts as a world model—and text-to-video doesn't make the cut.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

An international research team is establishing formal definitions and taxonomies for world models through OpenWorldLib, challenging industry assumptions about what capabilities constitute genuine world modeling. This matters because definitional clarity in AI research shapes which capabilities get funded, benchmarked, and pursued—and the exclusion of text-to-video generators signals a fundamental disagreement with the market narrative around Sora.

The key facts

9 to know
  1. OpenWorldLib framework establishes formal definition of world models

  2. Text-to-video generators (Sora) explicitly excluded from definition

  3. International research collaboration to standardize fragmented landscape

  4. Definitional disagreement suggests text-to-video and true world modeling are distinct capabilities

  5. Published April 12, 2026 (very recent)

  6. International research team released OpenWorldLib framework

  7. Text-to-video generators like Sora excluded from world model definition

  8. Addresses fragmented world model research landscape

  9. Published April 2026

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: An international research team wants to bring order to the fragmented world model research landscape with OpenWorldLib. Text-to-video models like Sora are explicitly left out of their definition.
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work