Gemini 3.5 Flash Lite logo

Gemini 3.5 Flash Lite

Chat

by Google

Overview

The fastest and cheapest model in Google's 3.5 line, released July 21, 2026 as the successor to Gemini 3.1 Flash Lite. Built for high-volume, low-latency jobs like agentic search, document processing, and translation.

Key Features
  • Fastest and cheapest model in the Gemini 3.5 line, succeeding Gemini 3.1 Flash Lite.
  • 1M token context window with full multimodal support for text, audio, images, video, and documents.
  • Optimized for high-volume, low-latency applications requiring speed and cost efficiency.
  • Configurable thinking capabilities for flexible reasoning depth based on task complexity.
  • Pricing is $0.30 per 1M input tokens and $2.50 per 1M output tokens.
Input Capabilities
Text
Audio
Image
Video
Document
Available Tools

Web Search

Search the web for current info.

Code Execution

Run Python code in a sandbox.

JSON Mode

Output responses in valid JSON format.

URL Context

Fetch and process content from URLs.

Google Maps

Ground responses in real-time Google Maps data for location-aware queries.

Image Generator

Generate images from text prompts.

10k FREE Credits50+ AI Models

Start Building with AI Today

Join thousands of developers using our unified platform to access 50+ premium AI models without multiple subscriptions.

OpenAI
Anthropic
Gemini
Grok
Meta
Runway
DeepMind
DeepSeek
Ideogram
ElevenLabs
Stability
Perplexity
Recraft