These AI Experts Want to Do High-Stakes Research Out in the Open
Trillium Labs is publishing research on self-improving AI models—work most frontier labs keep behind closed doors. Here's why transparency matters (and what it risks).

Why it matters
A new lab is challenging the frontier-lab norm of secrecy around high-risk research by committing to open publication of self-improvement and model-behavior work. This reflects a genuine tension in AI governance: how to balance safety-through-transparency against safety-through-containment.
The key facts
10 to knowTrillium Labs position: conducting self-improvement and model-behavior research in the open
Contrasts with frontier-lab practice of restricting high-stakes research to internal/private channels
Focus areas: self-improvement (agents learning from deployment) and model behavior (interpretability, alignment)
No funding amount, founding team size, or publication schedule disclosed
Raises governance question: whether open research on self-improving systems increases or decreases safety risk
Trillium Labs focuses on self-improvement and model behavior research
Strategy contrasts with industry norm of keeping high-risk research private
Open publication affects how frontier capability research is validated
No specific model release, benchmark, or technical capability data provided in article description
Positioning is strategic/reputational rather than announcing a capability milestone
The story so far
Earlier coverage of this storyline
- Elon Musk, SpaceXAI subpoenaed by NYC in AI safety investigationCNBC Technology
- Pope Leo criticises Nvidia’s Jensen Huang over AI safetyFinancial Times Technology
- Trump’s Crazy AI Rebrand Was a Loyalty Test for Tech Execs—and It WorkedWired AI
- This story
Go to the source
Wired AIwired.com
Publisher excerpt: Many frontier labs keep their risky research locked away. Trillium Labs wants to show off its work when it comes to self-improvement and model behavior.