Start building with Gemini 2.0 Flash and Flash-Lite
Gemini 2.0 Flash-Lite is now in production. Google's betting on lightweight inference to win the build layer.

Why it matters
Google is shipping a production-ready lightweight model variant designed for cost-sensitive and latency-critical deployments. This signals a strategic push to capture developers building at scale who can't afford flagship model pricing.
The key facts
5 to knowGemini 2.0 Flash-Lite now generally available in Gemini API
Available in Google AI Studio and Vertex AI for enterprise
Production-ready deployment
Positioned for cost and latency optimization vs. Flash flagship
Released February 25, 2025
Go to the source
Google DeepMind Blogdeepmind.google
Publisher excerpt: Gemini 2.0 Flash-Lite is now generally available in the Gemini API for production use in Google AI Studio and for enterprise customers on Vertex AI