3.6 Flash
Best for token efficiency in coding, knowledge work, and multimodal tasks
Best for token efficiency in coding, knowledge work, and multimodal tasks
Best for high-volume tasks that need efficiency and intelligence
Best for complex tasks and bringing creative concepts to life
Slide 1 of 4
Tackle complex, development tasks with advanced reasoning at speed.
Transform text, images, video and audio into rich interactive user interfaces.
Execute sophisticated workflows over extended timeframes.
Leverage advanced tools to solve demanding, real-world problems.
| Benchmark | Notes | Gemini 3.6 Flash | Gemini 3.5 Flash | Gemini 3.1 Pro | GPT-5.6 Luna | Grok 4.5 | Claude Sonnet 5 |
|---|---|---|---|---|---|---|---|
| Input price $/1M tokens, no caching | $1.50 | $1.50 | $2.00 | $1.00 | $2.00 | $3.00 $2.00 (temp discount) | |
| Output price $/1M tokens | $7.50 | $9.00 | $12.00 | $6.00 | $6.00 | $15.00 $10.00 (temp discount) | |
| SWE-Bench Pro (Public) Diverse agentic coding tasks | 58.7% | 55.1% | 54.2% | 62.7% | 64.7% | 63.2% | |
| DeepSWE v1.1 Long-horizon software engineering | 49% | 37% | 12% | 67% | 54% | 54% | |
| Terminal-bench 2.1 Agentic terminal coding | Terminus-2 harness | 78.0% | 76.2% | 73.8% | 84.7% | 83.3% | 80.4% |
| MLE-Bench Machine Learning Engineering | 63.9% | 49.7% | 42.6% | 47.6% | 43.2% | 66.9% | |
| OSWorld-Verified Agentic computer use | 83.0% | 78.4% | 76.2% | 72.6% | — | 81.2% | |
| GDPVal-AA v2 Knowledge work | Elo | 1421 | 1349 | 965 | 1584 | 1535 | 1607 |
| CharXiv Reasoning Information synthesis from complex charts | No tools | 85.2% | 84.2% | 83.3% | 82.7% | 81.6% | 77.0% |
| With tools | 89.4% | 84.9% | 83.2% | — | — | 88.3% | |
| GDM-MRCR v2 (8-needle) Long context performance | 128k (average) | 91.8% | 77.3% | 84.9% | 74.8% | 81.4% | 71.6% |
| 1M (pointwise) | 54.0% | 26.6% | 26.3% | — | — | — |
Slide 1 of 8
3.6 Flash shows better token efficiency and reduced verbosity than 3.5 Flash in an OSWorld verified task.
3.6 Flash, using Managed Agents on AIS, can help parse through and analyze financial data and transcripts more efficiently and accurately than 3.5 Flash.
3.6 Flash executes code migrations, using multi-agent orchestration on AGY, with lower latency and higher quality than 3.5 Flash.
3.6 Flash helps develop a photographic texture extractor for 3D workflows, using canvas.
3.5 Flash-Lite executes high volume tasks at a lower latency than 3.5 Flash.
Working alongside 3.6 Flash as the master agent, 3.5 Flash-Lite instantly generates 25 unique, ready-to-explore web design concepts.
3.5 Flash-Lite can scale receipt translation and summarization with its multimodal understanding.
3.5 Flash-Lite builds a game by instantly generating and iterating through multiple options.
Slide 1 of 4
Our AI-first development platform that allows anyone to be a builder
Leap from prompt to production
Get started building with cutting-edge AI models
Build, scale, and govern agents
Slide 1 of 7
Finding and fixing vulnerabilities quickly and efficiently
Best for modern challenges across science, research and engineering
Create anything from anything, starting with video
State-of-the-art image generation and editing models, built on Gemini
Advanced real-time audio models, built on Gemini
Our most advanced vision-language-action model
State-of-the-art multimodal embedding model
Finding and fixing vulnerabilities quickly and efficiently
Best for modern challenges across science, research and engineering
Create anything from anything, starting with video
State-of-the-art image generation and editing models, built on Gemini
Advanced real-time audio models, built on Gemini
Our most advanced vision-language-action model
State-of-the-art multimodal embedding model
Supercharge your creativity and productivity
Ask whatever's on your mind to get an AI powered response
Your research and thinking partner
The fastest path from prompt to production
Our AI-first development platform that allows anyone to be a builder
Get started building with cutting-edge AI models
Build, scale, and govern agents