Gemini 2.0 Flash-Lite
Google DeepMind · 2025-02-25
Gemini 2.0 Flash-Lite is Google's Gemini 2.0 model optimized for speed, scale, and cost efficiency; the gemini-2.0-flash-lite API model became generally available on February 25, 2025. It uses the series' sparse Mixture-of-Experts Transformer architecture, accepts text, image, audio, and video inputs with a 1,048,576-token context window, and returns up to 8,192 text tokens.