Gemini 3.6 Flash
Our workhorse model that delivers better coding, knowledge work, and multimodal performance.
A Google Gemini 3-series natively multimodal reasoning model optimized for efficient agentic workflows, coding, knowledge work, and multimodal tasks.
Pricing
Model profile · Current snapshot
Input / 1M
$0.75
Output / 1M
$3.75
Blended / 1M
$1.50
Output TPS
185
TTFT (s)
11.58
Model Intelligence
Context window
1,048,576 tokens
Arena ranking
34
Benchmarkable
Yes
Model level
release
Intelligence Index
34.3
Coding Index
69.2
GPQA
0.93
HLE
0.41
SciCode
0.53
LCR
0.8
Recent stories
2 linked stories
releaseSECONDARY2026-07-28
Gemini API adds token budget caps for Managed Agents
Google added token budget caps and other controls for Managed Agents in the Gemini API. The release also adds sandbox hooks, cron triggers, model configuration, free-tier support, and Gemini 3.6 Flash defaults.
releasePRIMARY2026-07-21
Google ships Gemini 3.6 Flash and 3.5 Flash-Lite to serving platforms
Gemini 3.6 Flash and 3.5 Flash-Lite went live on OpenRouter, Venice, Hyperbrowser, and Google surfaces. Early benchmarks show lower token use and cost, but mixed document and coding results.