Gemini 3.5 Flash-Lite
Low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution.
Gemini 3.5 Flash-Lite is a specific Gemini model release: a low-latency, cost-effective, natively multimodal model optimized for high-throughput, low-cost agentic workflows, subagent tasks, document parsing, translation, classification, and data extraction. It supports text, image, video, audio, and PDF inputs with text output.
Pricing
Model profile · Current snapshot
Input / 1M
$0.30
Output / 1M
$2.50
Blended / 1M
$0.85
Output TPS
322
TTFT (s)
7.75
Model Intelligence
Context window
1,048,576 tokens
Arena ranking
37
Benchmarkable
Yes
Model level
release
Intelligence Index
37.4
Coding Index
49.3
GPQA
0.84
HLE
0.19
SciCode
0.41
LCR
0.75
Recent stories
1 linked story