Gemini 2.0 Flash Lite
by Google
Ultra-cost-effective variant of Gemini 2.0 Flash optimized for maximum efficiency and low latency deployment. Delivers Gemini 2.0 capabilities with the most competitive pricing for high-volume applications requiring speed over maximum intelligence.
- Ultra-cost-effective pricing at $0.075 input, $0.30 output per 1M tokens for budget-conscious applications
- Optimized for cost efficiency and low latency while maintaining Gemini 2.0 core capabilities
- High-volume deployment ready for applications requiring speed and cost optimization
- Multimodal support across text, audio, images, video, and documents with 1M context window
- Native tool integration for practical agentic workflows with function calling support
JSON Mode
Output responses in valid JSON format.
Image Generator
Generate images from text prompts.
Model Information
Supported Formats
Google's most intelligent stable model for coding tasks, released July 21, 2026 as the successor to Gemini 3.5 Flash. Combines frontier-class intelligence with Flash-tier speed, with full support for thinking, function calling, structured outputs, code execution, URL context, and search/Maps grounding.
The fastest and cheapest model in Google's 3.5 line, released July 21, 2026 as the successor to Gemini 3.1 Flash Lite. Built for high-volume, low-latency jobs like agentic search, document processing, and translation.
Google's most advanced reasoning model (Gemini 3.1 Pro) with configurable thinking levels (minimal/low/medium/high) for optimal cost-performance balance. Successor to Gemini 3 Pro with enhanced reasoning and tool capabilities including Google Maps grounding.