3.6 Flash
Best for token efficiency in coding, knowledge work, and multimodal tasks
Frontier intelligence with action
Best for token efficiency in coding, knowledge work, and multimodal tasks
Best for high-volume tasks that need efficiency and intelligence
Best for complex tasks and bringing creative concepts to life
Best for modern challenges across science, research and engineering
Introducing our latest series of models combining frontier intelligence with action. Build more capable, intelligent agents efficiently.
Completing everyday tasks, or solving your most challenging problems. Discover the right model for what you need
Best for token efficiency in coding, knowledge work, and multimodal tasks
Best for high-volume tasks that need efficiency and intelligence
Best for complex tasks and bringing creative concepts to life
Tackle complex, development tasks with advanced reasoning at speed.
Transform text, images, video and audio into rich interactive user interfaces.
Execute sophisticated workflows over extended timeframes.
Leverage advanced tools to solve demanding, real-world problems.
| Benchmark | Notes | Gemini 3.6 Flash | Gemini 3.5 Flash | Gemini 3.1 Pro | GPT-5.6 Luna | Grok 4.5 | Claude Sonnet 5 |
|---|---|---|---|---|---|---|---|
| Input price $/1M tokens, no caching | $1.50 | $1.50 | $2.00 | $1.00 | $2.00 | $3.00 $2.00 (temp discount) | |
| Output price $/1M tokens | $7.50 | $9.00 | $12.00 | $6.00 | $6.00 | $15.00 $10.00 (temp discount) | |
| SWE-Bench Pro (Public) Diverse agentic coding tasks | 58.7% | 55.1% | 54.2% | 62.7% | 64.7% | 63.2% | |
| DeepSWE v1.1 Long-horizon software engineering | 49% | 37% | 12% | 67% | 54% | 54% | |
| Terminal-bench 2.1 Agentic terminal coding | Terminus-2 harness | 78.0% | 76.2% | 73.8% | 84.7% | 83.3% | 80.4% |
| MLE-Bench Machine Learning Engineering | 63.9% | 49.7% | 42.6% | 47.6% | 43.2% | 66.9% | |
| OSWorld-Verified Agentic computer use | 83.0% | 78.4% | 76.2% | 72.6% | — | 81.2% | |
| GDPVal-AA v2 Knowledge work | Elo | 1421 | 1349 | 965 | 1584 | 1535 | 1607 |
| CharXiv Reasoning Information synthesis from complex charts | No tools | 85.2% | 84.2% | 83.3% | 82.7% | 81.6% | 77.0% |
| With tools | 89.4% | 84.9% | 83.2% | — | — | 88.3% | |
| GDM-MRCR v2 (8-needle) Long context performance | 128k (average) | 91.8% | 77.3% | 84.9% | 74.8% | 81.4% | 71.6% |
| 1M (pointwise) | 54.0% | 26.6% | 26.3% | — | — | — |
For details on our evaluation methodology please see deepmind.google/models/evals-methodology/gemini-3-6-flash
Explore what you can do with Gemini 3.6 Flash and 3.5 Flash-Lite
3.6 Flash shows better token efficiency and reduced verbosity than 3.5 Flash in an OSWorld verified task.
3.6 Flash, using Managed Agents on AIS, can help parse through and analyze financial data and transcripts more efficiently and accurately than 3.5 Flash.
3.6 Flash executes code migrations, using multi-agent orchestration on AGY, with lower latency and higher quality than 3.5 Flash.
3.6 Flash helps develop a photographic texture extractor for 3D workflows, using canvas.
3.5 Flash-Lite executes high volume tasks at a lower latency than 3.5 Flash.
Working alongside 3.6 Flash as the master agent, 3.5 Flash-Lite instantly generates 25 unique, ready-to-explore web design concepts.
3.5 Flash-Lite can scale receipt translation and summarization with its multimodal understanding.
3.5 Flash-Lite builds a game by instantly generating and iterating through multiple options.
Shopify is running subagents in parallel to analyze complex data over a long horizon for more accurate merchant growth forecasts at a global scale.
Macquarie Bank is piloting how 3.5 Flash can accelerate customer onboarding by reasoning over complex 100+ page documents, retrieving relevant information and making reliable recommendations with low latency.
Salesforce is integrating 3.5 Flash into Agentforce to reliably automate complicated enterprise tasks by deploying multiple subagents that retain context and execute complex, multi-turn tool calling.
3.5 Flash is helping Ramp enable smarter, more reliable OCR through multimodal understanding of complex invoices combined with reasoning over historical patterns.
Xero is deploying agents to autonomously manage complex, multi-week workflows, such as identifying suppliers and gathering information for 1099 tax forms, enabling small businesses to automate tedious admin tasks.
Databricks is using agentic workflows to monitor and retrieve real-time information, reason across massive datasets to diagnose issues, identify fixes and propose solutions for data scientists.
Build with Gemini 3.5
Our AI-first development platform that allows anyone to be a builder
Leap from prompt to production
Get started building with cutting-edge AI models
Build, scale, and govern agents
Building with responsibility at the core
As we develop these new technologies, we recognize the responsibility it entails, and aim to prioritize safety and security in all our efforts.
Finding and fixing vulnerabilities quickly and efficiently
Best for modern challenges across science, research and engineering
Create anything from anything, starting with video
State-of-the-art image generation and editing models, built on Gemini
Advanced real-time audio models, built on Gemini
Our most advanced vision-language-action model
State-of-the-art multimodal embedding model
Supercharge your creativity and productivity
Ask whatever's on your mind to get an AI powered response
Your research and thinking partner
The fastest path from prompt to production
Our AI-first development platform that allows anyone to be a builder
Get started building with cutting-edge AI models
Build, scale, and govern agents