gemini-3.5-flash-lite

Common Name: Gemini 3.5 Flash-Lite

Google
-10%On SaleReleased on Jul 21 12:00 AMKnowledge Cutoff Mar 1 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Google's fastest, most cost-effective Gemini 3.5 model, delivering 350 output tokens/sec for high-throughput tasks like agentic search, document processing, and translation.

Specifications

Context
1000K
Maximum Output
64K
Inputtext, image, audio, video
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Standard
Batch
Input/MTokens
$0.27
$0.135
Cached Input/MTokens
$0.027
$0.018
Output/MTokens
$2.25
$1.125
Thinking Output/MTokens
$2.25
$1.125

Availability Trend (24h)

Performance Metrics (24h)

Similar Models

$0.27/$2.25/M-10%
ctx1.0Mmax66Kavailtps

Google's most efficient workhorse model designed for speed and low-cost. Improved across key benchmarks for reasoning, multimodality, code and long context while being 20-30% more efficient.

$0.45/$2.70/M-10%
ctx1.0Mmax64Kavailtps

Preview of Google's next-generation Gemini 3 Flash model, optimized for speed with frontier intelligence combined with superior search and grounding capabilities.

$0.225/$1.35/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Google's most cost-efficient Gemini 3 series model, optimized for high-volume agentic tasks, translation, and simple data processing with 2.5X faster time to first token than 2.5 Flash.

$0.45/$2.70/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Gemini 3.1 Flash Image generation model designed for speed and efficiency, effective for quick interactive image responses and high throughput.