Gemini 2.5 Flash
by Google
Best price-performance model with hybrid reasoning capabilities for large-scale processing and agentic use cases. Offers well-rounded capabilities with thinking features optimized for low-latency, high-volume tasks requiring both speed and intelligence.
- Optimal price-performance ratio delivering well-rounded capabilities with hybrid reasoning
- Best for large-scale processing, low-latency applications, and high-volume agentic workflows
- Comprehensive multimodal support with native audio capabilities and Live API integration
- Enhanced thinking capabilities with configurable reasoning budgets for complex problem solving
- 1M token context window with superior speed and native tool use for real-time applications
Web Search
Search the web for current info.
Code Execution
Run Python code in a sandbox.
JSON Mode
Output responses in valid JSON format.
URL Context
Fetch and process content from URLs.
Google Maps
Ground responses in real-time Google Maps data for location-aware queries.
Image Generator
Generate images from text prompts.
Model Information
Supported Formats
Google's most intelligent stable model for coding tasks, released July 21, 2026 as the successor to Gemini 3.5 Flash. Combines frontier-class intelligence with Flash-tier speed, with full support for thinking, function calling, structured outputs, code execution, URL context, and search/Maps grounding.
The fastest and cheapest model in Google's 3.5 line, released July 21, 2026 as the successor to Gemini 3.1 Flash Lite. Built for high-volume, low-latency jobs like agentic search, document processing, and translation.
Google's most advanced reasoning model (Gemini 3.1 Pro) with configurable thinking levels (minimal/low/medium/high) for optimal cost-performance balance. Successor to Gemini 3 Pro with enhanced reasoning and tool capabilities including Google Maps grounding.