Google DeepMind

Best for tackling complex agentic tasks at scale

Our most intelligent workhorse model yet for coding and agents.

Slide 1 of 4

Versatility across agentic tasks

Optimized for high performance across software engineering, web development and knowledge work tasks.

Smart and rigorous

Rigorous reasoning efforts for better quality output.

Reliable in agentic execution

Navigates roadblocks and resolves coding and real world issues with accuracy.

Truly multimodal

Multimodal understanding across text, audio, images, code, and video.


Slide 1 of 4

3D game built in Google Antigravity

From a simple text prompt to a fully playable 3D game. We used Gemini 3.7 Flash combined with Nano Banana to dynamically generate characters, items, and textures in real-time.

Interactive parallax landing pages

Stunning, interactive landing pages generated in a single shot. We used Gemini 3.7 Flash to orchestrate sub-agents, using Gemini Omni to create smooth, interactive parallax components.

Train a robotics model with multimodal graph looping

Watch a robotics model learn faster. We combined Gemini 3.7 Flash's multimodal understanding with a 3-agent graph loop to speed up the training loop.

Interactive annual report

From a static PDF to an interactive data story. Watch how complex annual reports are transformed into engaging web experiences complete with live charts and aggregated insights.


BenchmarkNotesGemini 3.7 FlashGemini 3.6 FlashClaude Sonnet 5GPT-5.6 TerraMuse Spark 1.2
Input price $/1M tokens$0.75*$0.75*$2.00$2.00$1.25
Output price $/1M tokens$3.75*$3.75*$10.00$12.00$4.25
Artificial Analysis Intelligence Index Composite model intelligence5652555757
FrontierCode 1.1 Main Production code qualityScore43.6%34.4%42.7%41.3%
DeepSWE v1.1 Long-horizon software engineering65.3%48.6%53.8%69.6%54.9%
Code Arena Web developmentElo15881538154115231535
Terminal-bench 2.1 Agentic terminal coding85.8%78.0%80.4%87.4%82.9%
Terminal-bench 3.0 General agent capabilities14.9%5.4%14.6%20.8%
AutomationBench Enterprise workflow automationPrivate set30.4%17.0%10.7%23.6%
GDPVal-AA v2 Knowledge workElo15251422159815781628
Harvey LAB-AA Complex legal workflows90.7%85.1%90.1%85.2%
GDP.pdf Expert PDF document comprehension34.0%22.0%28.0%24.7%16.0%
CharXiv Reasoning Information synthesis from complex chartsNo tools84.5%85.2%77.0%85.9%
With tools88.7%89.4%88.3%
LVBench Long video understanding85.4%84.2%68.5%78.9%
GDM-MRCR v2 (8-needle) Long context performance128k (average)97.0%91.8%81.5%93.5%
OSWorld-2.0 Agentic computer use47.9%33.8%50.2%
Agent's Last Exam Multimodal desktop and OS agent tasksPass rate26.3%24.2%33.3%28.0%
HLE-Verified Multidisciplinary expert reasoning53.6%51.2%31.0%51.1%
BioMysteryBench Bioinformatics research reasoningHuman solvable87.1%80.6%87.5%83.8%
Human difficult43.5%41.2%34.1%49.4%
LABBench2 Biology real-world research tasks82.1%76.1%80.1%81.2%
Name
3.7 Flash
Status
General availability
Input
  • Text
  • Image
  • Video
  • Audio
  • PDF
Output
  • Text
Input tokens
1M
Output tokens
64k
Tool use
  • Function calling
  • Search as a tool
  • Computer use
Best for
  • Everyday tasks
  • Agentic coding
  • Advanced reasoning
  • Multimodal understanding
  • Knowledge work
Availability
  • Gemini App
  • Gemini Enterprise App
  • Gemini Enterprise Agent Platform
  • Google AI Studio
  • Gemini API
  • Google Antigravity
Documentation
View developer docs
Model card
View model card

Read the original on deepmind.google ↗