Gemini 3.6 Flash
About this model
Gemini 3.6 Flash is Google's July 2026 update to the efficiency-focused Flash line, positioned between low-cost throughput models and the Pro tier. Google positions it for real-world agentic tasks—code generation, multi-step orchestration, and spatial reasoning. It carries a one-million-token context window and accepts text, image, audio, and video inputs, outputting up to 64K tokens. Google's documentation confirms it is the successor to Gemini 3.5 Flash.
Against that predecessor, Google emphasizes token efficiency: it reports 3.6 Flash uses fewer tokens and completes multi-step workflows in fewer turns, while lowering compile-failure and revision rates in coding environments. DeepMind states it executes code migrations with lower latency and higher quality than 3.5 Flash. A built-in computer use tool and configurable thinking levels (from minimal to high) let developers trade latency against reasoning depth.
It launched alongside Gemini 3.5 Flash-Lite, a higher-throughput sibling, and follows earlier Flash checkpoints including Gemini 3 Flash Preview. Independent evaluator Artificial Analysis characterizes 3.6 Flash as fast and fairly concise, and confirms its reasoning behavior and one-million-token context. For deeper reasoning, Google also offers Gemini 3.1 Pro Preview in the Pro tier.
This About section is AI-generated from public sources (Claude Opus 4.8), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.
Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 1d ago