Full KeyRank table

Every qualifying story, ranked. How KeyRank works

#StoryPillarKeyRankSource
1Anthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the benchmark designed to measure real intelligenceFrontier85The Decoder
2DeepsecBench: evaluating model performance in finding cybersecurity vulnerabilitiesFrontier82Vercel Blog
3Anthropic says its Mythos model found vulnerabilities in cryptographic algorithms that secure the internetFrontier78The Decoder
4Microsoft launches MAI-Cyber-1-Flash cybersecurity AI model - qz.comFrontier78Reuters Technology
5Moonshot AI releases Kimi K3 open weights and infrastructure after shaking up the frontier model raceFrontier78The Decoder
6Exclusive: Cogent Security debuts VR-1, a frontier model built to prove attack pathsFrontier78SiliconAngle
7Why China is giving away its best AI modelsFrontier78The Verge AI
8Black Forest Labs Releases FLUX 3: A Multimodal Flow Model for Image, Video, Audio and Robot Action PredictionFrontier78MarkTechPost
9Anthropic's Claude Opus 5 costs well below Fable 5 while matching or beating it across most benchmarksFrontier78The Decoder
10How GPT-5.6 fuses frontier intelligence with frontier efficiencyFrontier75OpenAI Blog
11Thinking Machines bets on efficiency over size with its second model, Inkling SmallFrontier72The Decoder
12Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoidsFrontier72The Decoder
13DeepSeek Upgrades DeepSeek-V4-Flash-0731 with Major Agentic and Coding GainsFrontier72MarkTechPost
14PolyAI launches new real-time voice conversation model to make AI-driven calls more humanFrontier72SiliconAngle
15Google DeepMind debuts Gemini Robotics 2 model series for humanoid robotsFrontier72SiliconAngle
16Microsoft AI bets on cheap specialist models instead of chasing the frontierFrontier72The Decoder
17Ex-OpenAI researcher bets $100 billion will flow into training data because scaling alone won't cut itFrontier72The Decoder
18Tencent Open-Sources AngelSpec: A Unified Training Framework for MTP and Block-Parallel Speculative Decoding on Hy3 ModelsFrontier72MarkTechPost
19Google DeepMind Ships Three Physical AI Models For Whole Body Control, Dexterity And Multi Robot CollaborationFrontier72MarkTechPost
20PolyAI Releases Dialog-RSN-1: An Audio-Native Dialog Model That Fuses Turn-Taking, Speech Recognition, Function Calling, And ResponseFrontier72MarkTechPost
21China’s open-weight model lead exposes America’s AI blind spotFrontier72CNBC Technology
22Gemini Robotics 2 Brings Google's AI Into the Physical WorldFrontier72Wired AI
23Google DeepMind’s new AI model can control a robot’s entire bodyFrontier72The Verge AI
24Gemini Robotics 2 brings whole body intelligence to robotsFrontier72Google DeepMind Blog
25Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaborationFrontier72Google DeepMind Blog
26Microsoft is openly competing with OpenAI, Anthropic more than everFrontier72TechCrunch AI
27It’s Frighteningly Easy to Jailbreak Some Frontier AI ModelsFrontier72Wired AI
28ByteDance’s plan to dominate AIFrontier72Financial Times Technology
29How enabling two settings tripled our scores on the ARC-AGI-3 benchmarkFrontier72OpenAI Blog
30Amazon reportedly scales back its Nova AI models and bets on a new Frontier research teamFrontier72The Decoder
31moonshotai/Kimi-K3Frontier72Simon Willison
32Microsoft launches its own cybersecurity model MAI-Cyber-1-Flash but still depends on OpenAI for the toughest tasksFrontier72The Decoder
33Kimi AI and kvcache-ai Open Sources ‘AgentENV’: A Distributed System that Powers Agentic Reinforcement Learning (RL) Training for Kimi K3Frontier72MarkTechPost
34Microsoft AI Releases MAI-Cyber-1-Flash: A 5B-Active-Parameter Cyber Model That Pushes MDASH to 95.95% on CyberGymFrontier72MarkTechPost
35Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI’s accidental AI hackerFrontier72Import AI (Blog)
36The AI giants’ new problem: open AI - The VergeFrontier72Reuters Technology
37Induction Labs Photon-1 Simulates Desktops, Plays Checkers, and Models Billiard Physics From One Pretraining RunFrontier72MarkTechPost
38KwaiKAT Team Releases KAT-Coder-V2.5: An Agentic Coding Model Trained on 100,000+ Verifiable Repository EnvironmentsFrontier72MarkTechPost
39Opus 5 may have solved browser-based prompt injection, the biggest security flaw haunting AI agentsFrontier72The Decoder
40Meet Open Dreamer: A JAX/Flax Reproduction of the Dreamer 4 World Model Pipeline, With the Full Training Recipe PublishedFrontier72MarkTechPost
41Sakana AI Releases Fugu-Cyber: An Orchestration Model Reporting 86.9% on CyberGym and 72.1% on CTI-REALMFrontier72MarkTechPost
42FAIRChem v2 UMA for Multidomain Atomistic Simulation across Molecules, Catalysts, Materials, Vibrations, and Molecular DynamicsFrontier72MarkTechPost
43Introducing Claude Opus 5Frontier72Simon Willison
44Anthropic launches Claude Opus 5 with efficiency, safety improvementsFrontier72SiliconAngle
45Datalab Marker v2 vs MinerU, Docling, and Liteparse: Benchmark BreakdownFrontier72MarkTechPost
46Open weights vs. closed: An AI civil war's afoot, and the stakes are existentialFrontier68ZDNet AI
47Language models can't spark scientific revolutions, but world models mightFrontier62The Decoder
48GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error ratesFrontier62The Decoder
49Liquid AI Releases LFM2.5-Encoder-230M and LFM2.5-Encoder-350M: Bidirectional Encoders That Stay Fast at 8K Context on CPUFrontier62MarkTechPost
50Moonshot AI Open-Sources MoonEP: A Perfectly Balanced Expert Parallelism Library for MoE TrainingFrontier62MarkTechPost
51LFM2.5-Encoders for Fast Long-Context Inference on CPUFrontier62Hugging Face Blog
52EvoLib: Turning experience into evolving knowledgeFrontier55Microsoft Research
53smevals - a small eval suite for evaluating models, prompts, and harnessesFrontier52Simon Willison
54Nvidia’s Open Source Alliance Is Missing Some Key Names: OpenAI and AnthropicFrontier52Wired AI
55MoMo: Dial Motion Mode in Robot Manipulation with Spatiotemporal Action TokenizationFrontier45Apple Machine Learning
56Everyone Is Freaking Out About OpenAI and Anthropic’s Race for DominanceFrontier45Wired AI
57OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settingsFrontier45The Decoder
58Latest AI Uses Tabular Foundation Models To Turn Columnar Data Into Vital InsightsFrontier45Forbes Innovation
59How controllers from industrial machinery can coordinate multitask machine learningFrontier42Amazon Science