WorkThe story, in brief

Gemma Scope 2: helping the AI safety community deepen understanding of complex language model behavior

Google just open-sourced interpretability tools for an entire model family. Here's why the AI safety community is paying attention.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Google's Gemma Scope 2 release democratizes model interpretability across the Gemma 3 family, giving the safety and research community open-source tools to audit and understand LLM behavior—a critical capability as models become more complex and deployed at scale.

The key facts

5 to know
  1. Gemma Scope 2 released with open interpretability tools

  2. Tools available across entire Gemma 3 model family

  3. Targets AI safety community and model behavior research

  4. Open-source distribution model

  5. Released Dec 16, 2025

Go to the source

Google DeepMind Blogdeepmind.google

Publisher excerpt: Open interpretability tools for language models are now available across the entire Gemma 3 family with the release of Gemma Scope 2.
Read original report
Back to today's editionMore work news

The wider picture

View all
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work01

Trump now says he wants to form an ‘AI Force’

A major political signal on AI governance: the administration is positioning itself to accelerate rather than constrain AI development, with formal institutional backing (czar + task force). Practitioners and policy-watchers need to know the regulatory stance is shifting toward facilitation.

The Verge AI
Illustration of two anonymous hands arranging task cards around an amber tool on a shared desk.
AI illustration by KeyNews
Work02

The 'robot relations' department may become reality in workplace of the future

As corporations deploy autonomous systems across operations, workers face real changes to pay, autonomy, and job structure. The organizational and policy implications of managing human-AI work dynamics are becoming immediate workplace issues.

CNBC Technology
Paper-cut illustration of an amber microchip with circuit paths extending into a row of data-center cabinets.
AI illustration by KeyNews
Work03

AI, data center alarms dominate Congressional Black Caucus week in Washington

AI regulation and data-center expansion are now front-and-center in a major political forum, signaling emerging consensus-building around policy that will affect enterprise AI deployment and the communities hosting compute infrastructure.

CNBC Technology