GoogleGoogle·💬 Text Generation·New

Gemini 3.6 Flash

ReasoningVisionFunction CallingWeb SearchAudioanonymized
🧠 Try in Intelligence →Try on Venice.ai ↗
Quick reference
Gemini 3.6 Flash — TLDR
  • 🆕 Google's latest Flash-tier model, based on Gemini 3.5 Flash.
  • 📏 One million token context, up to 64K output tokens.
  • 👁️ Multimodal inputs: text, image, audio, and video.
  • 🧠 Reasoning model with configurable thinking levels for cost control.
  • 🔧 Built-in computer use tool for agentic execution.
  • ⚡ Google reports fewer output tokens and fewer turns than predecessor.
  • 🎯 Optimized for coding, agentic loops, and multi-step orchestration.
  • 🏢 Available via Gemini API, Google Cloud, and Gemini app.
💰 Pricing
$1.88 / $9.38
per 1M · input / output
📏 Context
1M tokens
📅 On Venice since
Jul 9, 2026
13 days ago
Provider

Google is an American multinational technology corporation and one of the world's most valuable brands. A subsidiary of parent company Alphabet Inc., Google operates across search, cloud computing, consumer electronics, and artificial intelligence. Its…

Read full profile →
32 models on Venice
12 text · 11 video · 3 image · 3 inpaint · 1 music · 1 embedding · 1 tts
Since Oct 15, 2024

About this model

Gemini 3.6 Flash is Google's July 2026 update to the efficiency-focused Flash line, positioned between low-cost throughput models and the Pro tier. Google positions it for real-world agentic tasks—code generation, multi-step orchestration, and spatial reasoning. It carries a one-million-token context window and accepts text, image, audio, and video inputs, outputting up to 64K tokens. Google's documentation confirms it is the successor to Gemini 3.5 Flash.

Against that predecessor, Google emphasizes token efficiency: it reports 3.6 Flash uses fewer tokens and completes multi-step workflows in fewer turns, while lowering compile-failure and revision rates in coding environments. DeepMind states it executes code migrations with lower latency and higher quality than 3.5 Flash. A built-in computer use tool and configurable thinking levels (from minimal to high) let developers trade latency against reasoning depth.

It launched alongside Gemini 3.5 Flash-Lite, a higher-throughput sibling, and follows earlier Flash checkpoints including Gemini 3 Flash Preview. Independent evaluator Artificial Analysis characterizes 3.6 Flash as fast and fairly concise, and confirms its reasoning behavior and one-million-token context. For deeper reasoning, Google also offers Gemini 3.1 Pro Preview in the Pro tier.

This About section is AI-generated from public sources (Claude Opus 4.8), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 1d ago