Our workhorse model that reduces output token usage by 17% compared to 3.5 Flash, according to Artificial Analysis Index.



Performance

3.6 Flash is more token efficient than 3.5 Flash and a step-up in coding and knowledge work.

BenchmarkNotesGemini 3.6 FlashGemini 3.5 FlashGemini 3.1 ProGPT-5.6 LunaGrok 4.5Claude Sonnet 5
Input price $/1M tokens, no caching$1.50$1.50$2.00$1.00$2.00$3.00 $2.00 (temp discount)
Output price $/1M tokens$7.50$9.00$12.00$6.00$6.00$15.00 $10.00 (temp discount)
SWE-Bench Pro (Public) Diverse agentic coding tasks58.7%55.1%54.2%62.7%64.7%63.2%
DeepSWE v1.1 Long-horizon software engineering49%37%12%67%54%54%
Terminal-bench 2.1 Agentic terminal codingTerminus-2 harness78.0%76.2%73.8%84.7%83.3%80.4%
MLE-Bench Machine Learning Engineering63.9%49.7%42.6%47.6%43.2%66.9%
OSWorld-Verified Agentic computer use83.0%78.4%76.2%72.6%81.2%
GDPVal-AA v2 Knowledge workElo14211349965158415351607
CharXiv Reasoning Information synthesis from complex chartsNo tools85.2%84.2%83.3%82.7%81.6%77.0%
With tools89.4%84.9%83.2%88.3%
GDM-MRCR v2 (8-needle) Long context performance128k (average)91.8%77.3%84.9%74.8%81.4%71.6%
1M (pointwise)54.0%26.6%26.3%

For details on our evaluation methodology please see deepmind.google/models/evals-methodology/gemini-3-6-flash

Model information

Name
3.6 Flash
Status
General availability
Input
  • Text
  • Image
  • Video
  • Audio
  • PDF
Output
  • Text
Input tokens
1M
Output tokens
64k
Tool use
  • Function calling
  • Search as a tool
  • Computer use
Best for
  • Everyday tasks
  • Agentic coding
  • Advanced reasoning
  • Multimodal understanding
  • Long context understanding
Availability
  • Gemini App
  • Gemini Enterprise App
  • Gemini Enterprise Agent Platform
  • Google AI Studio
  • Gemini API
  • Google Antigravity
Documentation
View developer docs
Model card
View model card