Skip to content
AI Primer

Gemini 3.5 Flash-Lite

Low-latency, cost-effective multimodal model for high-throughput agentic workflows.

Gemini 3.5 Flash-Lite is a low-latency, cost-effective Gemini 3.5-class model release optimized for high-throughput, low-cost subagent tasks, document parsing, data extraction, coding, and other agentic workflows. It accepts text, image, video, audio, and PDF inputs and produces text output.

Pricing

Model profile · Current snapshot
Input / 1M
$0.30
Output / 1M
$2.50
Blended / 1M
$0.85
Output TPS
395
TTFT (s)
8.26

Model Intelligence

Context window
1,048,576 tokens
Arena ranking
37
Benchmarkable
Yes
Model level
release
Intelligence Index
36.5
Coding Index
49.3
GPQA
0.84
HLE
0.18
SciCode
0.41
LCR
0.62

Recent stories

1 linked story
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.