| 1 | Claude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and a new price war | Frontier | 88 | Simon Willison |
| 2 | Introducing GPT-6 Sol and Luna | Frontier | 88 | OpenAI Blog |
| 3 | Anthropic’s biolab made a discovery it’s comparing to Crispr | Frontier | 78 | The Verge AI |
| 4 | Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Than Opus 5 | Frontier | 78 | MarkTechPost |
| 5 | Anthropic and OpenAI roll out cheaper models in first release since call for slowdown | Frontier | 78 | CNBC Technology |
| 6 | Xiaomi's affordable flagship AI leads the open models, and Anthropic says Claude helped get it there | Frontier | 78 | The Decoder |
| 7 | GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmarkRising | Frontier | 75 | The Decoder |
| 8 | SpaceX launches Grok 4.7 with long-horizon processing, safety upgrades | Frontier | 75 | SiliconAngle |
| 9 | Deepmind was built to chase AGI, but its new chief just wants Gemini 4 out the door | Frontier | 72 | The Decoder |
| 10 | Anthropic says Claude discovered a new enzyme system, but CRISPR researchers call it routine genome mining | Frontier | 72 | The Decoder |
| 11 | Black Forest Labs launches FLUX 3 Action, an open robotics AI model | Frontier | 72 | The Decoder |
| 12 | Google launches two benchmark-topping speech generation models | Frontier | 72 | SiliconAngle |
| 13 | NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time | Frontier | 72 | MarkTechPost |
| 14 | Introducing MentalHealthBench | Frontier | 72 | OpenAI Blog |
| 15 | Anthropic unveils Opus 5.5: powerful performance, still premium price | Frontier | 72 | AI Business |
| 16 | Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent | Frontier | 72 | The Decoder |
| 17 | Alibaba Unveils Zhenwu V900 — and Plans Qwen Models With Up to 10 Trillion Parameters | Frontier | 72 | TechRepublic |
| 18 | Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning | Frontier | 72 | MarkTechPost |
| 19 | Meta admits Muse’s likeness to OpenClaw isn’t a coincidence | Frontier | 72 | TechCrunch AI |
| 20 | Google’s robotics unit Intrinsic open-sources its foundational infrastructure for intelligent robots | Frontier | 72 | SiliconAngle |
| 21 | OpenAI says its internal model solved over 100 long-standing math problems after just a month of training | Frontier | 72 | The Decoder |
| 22 | Anthropic Builds Biology Lab to Test What Claude Can Do in the Real World | Frontier | 72 | TechRepublic |
| 23 | Jev introduces a new shape of LLM - System One, aka Decision Models | Frontier | 72 | Simon Willison |
| 24 | OpenAI forms math advisory group as its AI resolves more than 100 open problems | Frontier | 72 | TechCrunch AI |
| 25 | Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing | Frontier | 72 | MarkTechPost |
| 26 | StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work | Frontier | 72 | MarkTechPost |
| 27 | Tether addresses AI underinvestment in Africa with open-source machine translation models | Frontier | 72 | CIO |
| 28 | Anthropic is operating a lab that conducts biology experiments | Frontier | 72 | TechCrunch AI |
| 29 | PrismML Releases Ternary Bonsai 2 27B: A 5.9 GB Apache 2.0 Model Retaining 98.2% of Qwen3.8 27B Performance | Frontier | 72 | MarkTechPost |
| 30 | Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs | Frontier | 72 | MarkTechPost |
| 31 | PrismML launches Bonsai 2 27B, a high-intelligence AI model so small it fits on consumer hardware | Frontier | 72 | SiliconAngle |
| 32 | The U.S. says China's AI progress is down to 'distillation.' But is it that clear cut? | Frontier | 72 | CNBC Technology |
| 33 | Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use | Frontier | 72 | MarkTechPost |
| 34 | Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarksRising | Frontier | 72 | The Decoder |
| 35 | Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 LanguagesRising | Frontier | 68 | MarkTechPost |
| 36 | Top AI experts badly underestimated how fast the field is moving, study finds | Frontier | 68 | The Decoder |
| 37 | Simulated students that make realistic mistakes help AI tutors learn fasterRising | Frontier | 62 | The Decoder |
| 38 | Vals, backed by Andreessen Horowitz, is looking to become the gold standard for AI benchmarkingRising | Frontier | 65 | TechCrunch AI |
| 39 | OpenAI Expands Outside Safety Reviews Into Model Training | Frontier | 62 | TechRepublic |
| 40 | Anthropic engineer explains why Claude's writing got worse although the model got smarter | Frontier | 62 | The Decoder |
| 41 | How UK AISI and EvalEval Are Making Benchmark Results Reproducible | Frontier | 62 | Hugging Face Blog |
| 42 | AI Models Built From Rat Brains Just Got Closer to Reality | Frontier | 62 | Wired AI |
| 43 | Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model | Frontier | 62 | MarkTechPost |
| 44 | Anthropic brings in Accenture for AI safety testing | Frontier | 62 | Financial Times Technology |
| 45 | World model companies are keeping a lot of secrets | Frontier | 62 | TechCrunch AI |
| 46 | Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think | Frontier | 62 | The Decoder |
| 47 | Visible chains of thought are a safety advantage for AI, but that transparency is slipping away | Frontier | 62 | The Decoder |
| 48 | Anthropic and OpenAI need truly independent safety evaluators, experts say in public letter | Frontier | 62 | CNBC Technology |
| 49 | OpenAI Introduces Triage Framework and Case Studies to Report Model Misalignment | Frontier | 62 | InfoQ AI/ML |
| 50 | Tencent's Gander aims to keep talking while it works in the background | Frontier | 57 | The Decoder |
| 51 | BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens at a 0.86pp Accuracy Cost | Frontier | 52 | MarkTechPost |
| 52 | Compressing Streaming Neural Audio Encoders via Latent-Space Distillation | Frontier | 52 | Apple Machine Learning |
| 53 | Priorities and principles for effective third party assessments | Frontier | 52 | OpenAI Blog |
| 54 | Advancing AI for biology: Teaching models to design and characterize antibodies | Frontier | 52 | Amazon Science |
| 55 | tokenizers v1: encode, decode and scaling, measured | Frontier | 52 | Hugging Face Blog |
| 56 | Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem | Frontier | 52 | Hugging Face Blog |
| 57 | How to Guide Your Language Flow | Frontier | 48 | Apple Machine Learning |
| 58 | Accelerating vision-language models with LFM2.5-VL-DSpark | Frontier | 45 | Hugging Face Blog |