Skip to content
AI Primer

Gemini 3.5 Flash-Lite

Low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution.

Gemini 3.5 Flash-Lite is a specific Gemini model release: a low-latency, cost-effective, natively multimodal model optimized for high-throughput, low-cost agentic workflows, subagent tasks, document parsing, translation, classification, and data extraction. It supports text, image, video, audio, and PDF inputs with text output.

Pricing

Model profile · Current snapshot
Input / 1M
$0.30
Output / 1M
$2.50
Blended / 1M
$0.85
Output TPS
322
TTFT (s)
7.75

Model Intelligence

Context window
1,048,576 tokens
Arena ranking
37
Benchmarkable
Yes
Model level
release
Intelligence Index
37.4
Coding Index
49.3
GPQA
0.84
HLE
0.19
SciCode
0.41
LCR
0.75

Recent stories

1 linked story
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.