FrontierThe story, in brief

xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to Claude and GPT-6

46 vs 53: Grok 4.7 lands mid-pack on Artificial Analysis, but the price-to-capability math might change the frontier calculus.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

xAI's new model release underperforms Claude and GPT-6 on published benchmarks, but aggressive pricing could reshape how practitioners evaluate the capability-cost tradeoff in the lab race. A clear signal of competitive positioning and the emergence of a two-tier frontier.

The key facts

12 to know
  1. Grok 4.7 scores 46 on Artificial Analysis Intelligence Index

  2. Claude Fable 5.1 scores 53

  3. GPT-6 scores 53

  4. Grok 4.7 significantly underperforms on agentic coding benchmarks

  5. Positioned as a low-cost alternative despite capability gap

  6. Published Mon Sep 21 2026

  7. Grok 4.7 released by xAI

  8. Artificial Analysis Intelligence Index: Grok 4.7 scores 46 points

  9. Claude Fable 5.1 scores 53 points

  10. GPT-6 scores 53 points

  11. Wider gap in agentic coding benchmarks

  12. Positioned as lower-cost alternative

Go to the source

The Decoderthe-decoder.com

Publisher excerpt: xAI has released Grok 4.7, its most capable model yet. But on the Artificial Analysis Intelligence Index, it scores just 46 points, landing mid-pack and well behind Claude Fable 5.1 and GPT-6 at 53 each. The gap grows even wider in agentic coding. The upside is it's cheap.
Read original report
Back to today's editionMore frontier news

The wider picture

View all
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier01

Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing

A capable open-weight diffusion model with multi-reference editing and transparency support raises the bar for accessible image generation; practitioners can now evaluate a credible alternative to closed models, and the architecture (prefix KV cache for fast edits) offers a technical playbook.

MarkTechPost
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier02

Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem

A novel pruning technique using physics-inspired optimization could make it cheaper and faster to compress frontier models into deployable sizes — practical for practitioners deploying at scale, and a methodological innovation worth tracking in the lab-race toolkit.

Hugging Face Blog
Illustration of a transparent lens revealing connected networks across layers of paper.
AI illustration by KeyNews
Frontier03

StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work

StepFun enters the sparse MoE race with a large-context model positioned for agent and knowledge-work use cases. The pricing and multimodal capabilities signal competitive pressure in the frontier lab space, and the October open-weight timeline affects the open-model landscape.

MarkTechPost