Gemini 3.6 Flash
Released Jul 21, 2026 · knowledge cutoff 2026-03
- Status
- Frontier
- Location
- United States / United Kingdom
- Modality
- Multimodal
- Context window
- 1M
- Max output
- 66K
- Speed (high effort)
- 214 tok/s · 16s to first answer token
- Price ($/MTok in / out)
- $1.50 / $7.50
- Cost per task (high effort)
- $0.50
- Open weights
- No
Benchmarks4 of 8 reported
GPQA Diamond
Terminal-Bench 2.1
Humanity's Last Exam
LMArena Elo
MMLU-Pro
SWE-bench Verified
AIME (latest)
ARC-AGI-2
Strengths & weaknesses
Strengths
- Sticks to the change you asked for instead of touching adjacent files during diagnostic work
- Reaches the same result with fewer output tokens than the Flash before it
- Available across Google's whole developer and enterprise surface on day one
Weaknesses
- Flash-tier depth ceiling — it executes well but doesn't crack genuinely hard reasoning
- Arrived with no matching Pro-tier model in its generation, so there was nothing to escalate to
Model card reports SWE-Bench Pro 58.7%, DeepSWE v1.1 49%, GDPval-AA v2 Elo 1421, and GDM-MRCR v2 91.8% at 128k; Google published none of the site's tracked benchmarks at launch. Artificial Analysis Intelligence Index 50 (third-party). GPQA Diamond 92.8% and HLE 38.3% are Artificial Analysis-measured, not Google-reported. LMArena Elo 1485 is the 'gemini-3.6-flash' text-arena listing (rank 12, arena.ai); it also sits #12 in the Frontend Code Arena at 1537. Announced 2026-07-21 alongside Gemini 3.5 Flash-Lite and the limited-access Gemini 3.5 Flash Cyber.
For developers
API model strings
- google
gemini-3.6-flash
Licence
Proprietary — weights not released
Retirement
No retirement announced
Lineage
No recorded predecessor