Thursday, May 22, 2025

Start of archive·May 16

Top story

One Useful Thing (Ethan Mollick)

Making AI Work: Leadership, Lab, and Crowd

Companies are failing at AI adoption not because of technology limitations, but because they lack the proper organizational framework combining top-down leadership, dedicated experimentation labs, and grassroots employee engagement.

Three-pillar framework: Leadership, Lab, and Crowd

The briefs

OpenAI's latest models (o3, o4-mini, GPT-4.1) are moving beyond benchmarks into production developer tools. CodeRabbit's deployment shows real-world ROI for code review automation—a concrete use case that investors and CTOs need to track.

Models deployed: o3, o4-mini, GPT-4.1

OpenAI is upgrading its Operator agent from GPT-4o to o3, signaling confidence in o3's reasoning and task-execution capabilities. This is a real-world deployment signal that o3 outperforms GPT-4o on agentic workloads—a critical benchmark for AI labs competing on agent feasibility.

OpenAI Operator now runs on o3 (upgraded from GPT-4o)