SPIRITT logoSPIRITTGoogle DeepMindGemini 3.6 Flash

Gemini 3.6 Flash is faster Flash, not a new IQ tier

Google DeepMind's July 21, 2026 Flash update keeps the same independent AA Intelligence band as 3.5 Flash (~50–52 depending on snapshot) while roughly halving time per task, using fewer tokens, and lowering output price to $7.50/M. Multimodal inputs, 1M context, and stronger vendor coding/computer-use numbers—run it as a real agent on SPIRITT.

Across X on Gemini 3.6 Flash

Spec reviews, multi-agent setups, token-efficiency takes, and real product teams shipping on Flash.

Official launch card

Google's efficiency release for agent workloads: same Flash intelligence band, less time and fewer tokens per task.

What Google built it for

A multimodal Flash workhorse (text, image, video, audio, PDF in; text out) with 1M context, thinking mode, function calling, and Google Search/Maps grounding.

Vendor-reported gains on coding agents, ML research (MLE-Bench), and computer use, plus lower output price ($7.50/M) and ~17% fewer AA output tokens versus 3.5 Flash.

Multimodal1M contextAgent efficiencyComputer useThinking modeFlash tier

AA Intelligence Index

Claude Fable 5
60
Kimi K3
57
Grok 4.5
54
Gemini 3.6 Flash
50
Gemini 3.5 Flash
50
Gemini 3.1 Pro Preview
46

AA: Gemini 3.6 Flash ~50–52 depending on methodology snapshot—effectively tied with 3.5 Flash on composite IQ, not a new frontier tier.

Where Gemini Flash is strong vs other models

Same score means nothing alone. These are head-to-head charts so you can see where it leads, where it is close, and what that means for real agent work.

Speed
Winner

Output tokens / sec (AA)

Gemini 3.5 Flash-Lite
350
Gemini 3.6 Flash
304
DeepSeek V4 Flash 0731
122

AA places Gemini 3.6 Flash among the fastest measured (~232–304 tok/s depending on run), second only to Flash-Lite class peaks.

Task time
Winner

Avg AA task time (minutes)

Gemini 3.6 Flash
1.3
Gemini 3.5 Flash
2.7

AA: ~2.7 → ~1.3 minutes average task time vs 3.5 Flash.

Token efficiency
Winner

Relative AA output tokens

Gemini 3.6 Flash
83
Gemini 3.5 Flash
100

Google/AA: ~17% fewer output tokens than 3.5 Flash on the Index suite.

Coding agents (vendor)
Winner

DeepSWE % (Google table)

Gemini 3.6 Flash
49
Gemini 3.5 Flash
37

Vendor-reported: DeepSWE 37% → 49%. Independent composite IQ did not move the same way.

Computer use (vendor)
Winner

OSWorld-Verified % (Google table)

Gemini 3.6 Flash
83
Gemini 3.5 Flash
78.4

Vendor-reported computer use: 78.4% → 83.0%.

ML research (vendor)
Winner

MLE-Bench % (Google table)

Gemini 3.6 Flash
63.9
Gemini 3.5 Flash
49.7

Vendor-reported MLE-Bench: 49.7% → 63.9%.

Price
Near frontier

Output price ($ / M tokens)

GPT-5.6 Luna (ref)
$1.20
Gemini 3.6 Flash
$7.50
Gemini 3.5 Flash
$9.00
Claude Sonnet 5 (ref)
$15.00

Output $9 → $7.50/M; input stays $1.50/M. Still pricier than GPT-5.6 Luna list rates.

Task cost
Winner

AA avg cost per task

Gemini 3.6 Flash
$0.50
Gemini 3.5 Flash
$0.59

Coverage citing AA: ~$0.59 → ~$0.50 average evaluated task cost vs 3.5 Flash.

Sources: Google DeepMind Gemini 3.6 Flash launch; Artificial Analysis Gemini 3.6 Flash model page and halving-time article; third-party summaries of AA task time/cost. Composite IQ is flat vs 3.5 Flash; efficiency and vendor coding/computer-use gains are the real story. On SPIRITT, use Flash where multimodal speed matters more than max reasoning.

How It Works

From zero to Gemini 3.6 Flash running real work in a cloud agent environment

01

Open a workspace

Open a workspace and land in a fully equipped cloud computer: browser, files, terminal, integrations, and memory. No local setup. No thin chat box pretending to be an agent.

Open a SPIRITT workspace
02

Pick Gemini 3.6 Flash

Open the model picker and choose Gemini 3.6 Flash. Same model shipped for agentic work, now inside a workspace that already has tools, browser control, and durable context.

Select Gemini 3.6 Flash in the model picker
03

Build or automate

Tell it what to ship or what to run. Gemini Flash can code, call tools, drive the browser, coordinate multi-step work, and keep going while you step away. The point is not another chat window. It is an agentic environment where Gemini Flash actually does the job.

Build and automate with Gemini 3.6 Flash

Try Gemini 3.6 Flash in an agentic environment

Multimodal Flash speed inside a full cloud agent workspace.

Questions

Frequently asked questions

01What is Gemini 3.6 Flash?+
Google DeepMind's July 2026 Flash-tier multimodal model for fast, cheaper agent work. 1M context, thinking mode, tool use, and strong speed—without claiming a new overall intelligence tier vs 3.5 Flash.
02How is 3.6 Flash different from 3.5 Flash?+
Similar AA Intelligence, much faster task completion, fewer output tokens, lower output price, and Google-reported gains on coding agents, computer use, and MLE-Bench.
03What is Gemini 3.6 Flash best for?+
High-volume multimodal agents, UI/computer-use loops, coding assistants that need speed, and Google-ecosystem grounded workflows where latency and token burn dominate cost.
04What are some of the main limitations of this model?+
It does not leapfrog frontier IQ on AA; TTFT can be high on some AA runs; list price still loses to cheaper Luna-class options; vendor benches need independent confirmation on your tasks.
05What is a good example of using Gemini 3.6 Flash well?+
SPIRITT workspace agent that reads screenshots/PDFs, calls tools, and iterates quickly—optimizing for wall-clock and tokens rather than max closed-model reasoning.
06Where is Gemini 3.6 Flash best to test for full agentic capabilities?+
SPIRITT Workspaces. Pick Gemini 3.6 Flash in the model picker and run real work in a fully equipped cloud environment: tools, browser, files, terminal, memory, and durable sessions. That is where the model can act as an agent, not just chat.
Buy from builders who use what they sellBuilt usingSPIRITT