Gemini 3.6 Flash
A versatile model, balanced across most use cases.
Overview
Gemini 3.6 Flash, released July 21 2026, becomes the default model in the Gemini family, succeeding 3.5 Flash. 1,048,576-token context, 65,536 max output, multimodal input (text, image, video, audio, PDF). March 2026 knowledge cutoff versus January 2025 for its predecessor — a 14-month jump. Unusually, output pricing dropped ($7.50 from $9) while the model uses roughly 17% fewer output tokens, compounding the savings. It leads on long context (91.8% on GDM-MRCR v2, best in its comparison set) and computer use (83% on OSWorld-Verified), but trails GPT-5.6 Luna and Grok 4.5 on coding. Around 280 tokens per second.
Skill profile
Not disclosed
Strengths
- World-class on Arena
- Very long context
Limitations
- No native GDPR guarantee
Who is it for
- you handle sensitive EU data
Access & availability
Key specifications
Estimate your monthly cost
Per-token API pricingIndicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.
Privacy
Advanced data · for expertsArchitecture, modalities, detailed cost, full benchmarks▾
| Arena Elo | 1483 |
| MMLU | — |
| GPQA | — |
| HumanEval | — |
| SWE-Bench | — |
| MATH | — |