Gemini 3.6 Flash
Google's workhorse Flash model for better coding, knowledge work, and multimodal performance with improved token efficiency.
Gemini 3.6 Flash is a stable Gemini model release optimized for agentic workflows, coding, knowledge work, spatial reasoning, and multimodal understanding. It accepts text, image, video, audio, and PDF inputs and outputs text, with a 1,048,576-token input context window and 65,536-token output limit.
Pricing
Model Intelligence
Recent stories
Google added token budget caps and other controls for Managed Agents in the Gemini API. The release also adds sandbox hooks, cron triggers, model configuration, free-tier support, and Gemini 3.6 Flash defaults.
Gemini 3.6 Flash and 3.5 Flash-Lite went live on OpenRouter, Venice, Hyperbrowser, and Google surfaces. Early benchmarks show lower token use and cost, but mixed document and coding results.