Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is the fastest, most cost-effective 3.5 model for high-throughput, low-latency tasks like agentic search and document processing.

gemini-3.5-flash-lite
STABLEGet StartedView uptime
1,048,576 context
Starting at $0.30/M input tokens
Starting at $2.50/M output tokens
Streaming
Vision
Tools
Reasoning
JSON Output

All Providers for Gemini 3.5 Flash Lite

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Google AI Studio
Context: 1.0M
Input
$0.3
/M tokens
Cached
$0.03
/M tokens
Output
$2.5
/M tokens
+ $0.014 per search
Get Started
Google Vertex AI
Context: 1.0M
Input
$0.3
/M tokens
Cached
$0.03
/M tokens
Output
$2.5
/M tokens
+ $0.014 per search
Get Started