The fastest and cheapest model in Google's 3.5 line, released July 21, 2026 as the successor to Gemini 3.1 Flash Lite. Built for high-volume, low-latency jobs like agentic search, document processing, and translation.
- Fastest and cheapest model in the Gemini 3.5 line, succeeding Gemini 3.1 Flash Lite.
- 1M token context window with full multimodal support for text, audio, images, video, and documents.
- Optimized for high-volume, low-latency applications requiring speed and cost efficiency.
- Configurable thinking capabilities for flexible reasoning depth based on task complexity.
- Pricing is $0.30 per 1M input tokens and $2.50 per 1M output tokens.
Web Search
Search the web for current info.
Code Execution
Run Python code in a sandbox.
JSON Mode
Output responses in valid JSON format.
URL Context
Fetch and process content from URLs.
Google Maps
Ground responses in real-time Google Maps data for location-aware queries.
Image Generator
Generate images from text prompts.
Model Information
Supported Formats
Google's most intelligent stable model for coding tasks, released July 21, 2026 as the successor to Gemini 3.5 Flash. Combines frontier-class intelligence with Flash-tier speed, with full support for thinking, function calling, structured outputs, code execution, URL context, and search/Maps grounding.
Google's most advanced reasoning model (Gemini 3.1 Pro) with configurable thinking levels (minimal/low/medium/high) for optimal cost-performance balance. Successor to Gemini 3 Pro with enhanced reasoning and tool capabilities including Google Maps grounding.
Google's most intelligent stable model for coding tasks in the Gemini 3 family. Gemini 3.5 Flash combines frontier-class intelligence with Flash-tier speed, making it ideal for complex coding workflows and agentic tasks at a cost-effective price.