WorkThe story, in brief

The Future Of AI Training Data Is Human. The Question Is How

NOBODY TALKING: Everyone chases synthetic data. A new partnership just proved human behavioral data from virtual worlds could be the moat.

Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
People, judgement and the changing nature of work.AI illustration by KeyNews
The KeyNews take

Why it matters

As synthetic data hits diminishing returns, a novel approach to sourcing training data from human behavior in virtual environments raises critical questions about data quality, consent, and competitive advantage in model development.

The key facts

8 to know
  1. Partnership between VLGE (metaverse startup) and Protege (data firm)

  2. Strategy: leverage natural human behavioral data from virtual environments for training sets

  3. Published June 2026 - emerging data sourcing trend

  4. Shifts training data paradigm from synthetic to human-behavioral sourcing

  5. Partnership: VLGE (metaverse startup) + Protege (data firm)

  6. Training data sourced from natural human behavior in virtual environments

  7. Signals industry move away from traditional web-scraped/synthetic data

  8. Raises governance questions: consent, privacy, data rights in human-derived training sets

Go to the source

Forbes Innovationforbes.com

Publisher excerpt: A new partnership between metaverse startup VLGE and data firm Protege leverages natural human behavioral data from virtual environments to build training sets.
Read original report
Back to today's editionMore work news

Keep reading

Related stories

More from Work