Fast mode for Opus 4.7 available on AI Gateway
2.5x faster. 6x the cost. Anthropic's new fast mode for Opus 4.7 is now live—here's the speed-vs-spend tradeoff.

Why it matters
Anthropic is shipping a production-ready speed tier for Opus 4.7, signaling confidence in the model's quality while introducing a clear monetization lever for latency-sensitive workloads. This is a tactical move to compete with inference optimization plays from OpenAI and others.
The key facts
6 to knowOpus 4.7 fast mode delivers ~2.5x faster output token generation
Fast mode priced at 6x standard Opus rates ($150/1M output tokens vs. $25/1M)
Available on Anthropic AI Gateway in research preview
Integrates with Claude Code via environment variables
Standard pricing multipliers (e.g., prompt caching) stack on top of fast mode rates
Feature is experimental and early-stage
Go to the source
Vercel Blogvercel.com
Publisher excerpt: Fast mode for Claude Opus 4.7 is now available on in research preview.AI Gateway Fast mode delivers ~2.5x faster output token generation with full Opus 4.7 intelligence. This is an early, experimental feature. To enable fast mode, pass in the provider options with .speed:…