Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is the fastest, most cost-effective 3.5 model for high-throughput, low-latency tasks like agentic search and document processing.

gemini-3.5-flash-lite
STABLEGet StartedView uptime
1,048,576 context
Starting at $0.30/M input tokens
Starting at $2.50/M output tokens
Streaming
Vision
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

All Providers for Gemini 3.5 Flash Lite

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Google AI Studio
Context: 1.0M
Input
$0.3
/M tokens
Cache Read
$0.03
/M tokens
Output
$2.5
/M tokens
Cache Write 5m
$0.0833
/M tokens
Cache Write 1h
$0.0833
/M tokens
+ $0.014 per search
Get Started
Google Vertex AI
Context: 1.0M
Input
$0.3
/M tokens
Cache Read
$0.03
/M tokens
Output
$2.5
/M tokens
Cache Write 5m
$0.0833
/M tokens
Cache Write 1h
$0.0833
/M tokens
+ $0.014 per search
Get Started