💰 Intro $0.75/$3.75 (thru 2026/12/31) → $1.50/$7.50 from 2027/1/1

Gemini 3.7 Flash

"Our most capable Flash model for agentic workflows and multimodal reasoning." Shipped three weeks after 3.6 Flash at half its launch price — $0.75/$3.75 through the end of 2026.

$0.75
Input / 1M tokens (intro)
$3.75
Output / 1M tokens (intro)
1M
Context window
64K
Max output tokens
Last updated: | Sources: Google Announcement · Google AI Pricing · Model Card | All Pricing · Calculator · Grok 4.6 · 日本語版

Pricing (intro vs. standard)

Verified against Google's official pricing page (ai.google.dev). All prices in USD per 1M tokens. Output pricing includes thinking tokens.

TierInputOutputCached InputNotes
Standard (intro, thru 12/31)$0.75$3.75$0.075Through Dec 31, 2026
Standard (from 1/1)$1.50$7.50$0.15From Jan 1, 2027
Batch / Flex (intro, thru 12/31)$0.375$1.875$0.0375Half of Standard
Priority (intro, thru 12/31)$1.35$6.75$0.135High priority. 1.8× Standard
⚠️ The intro price expires (critical): $0.75/$3.75 is valid only through December 31, 2026. From January 1, 2027 the standard price doubles to $1.50/$7.50. Google published both tiers at launch, so this is a scheduled — not speculative — 100% increase. Budget on the standard rate. The same intro price applies to Gemini 3.6 Flash.
💡 Grounding with Google Search: 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. Google Maps grounding has a similar free allowance plus metered pricing.

Primary sources: ai.google.dev/gemini-api/docs/pricing · Google Blog: Introducing Gemini 3.7 Flash (verified 2026-08-17)

Model Specifications

GA Release
Aug 14, 2026
Knowledge Cutoff
Mar 2026 (some Jan 2025)
Context Window
1M (1,048,576) tokens
Max Output
64K (65,536) tokens
Input Modalities
Text + Image + Audio + Video
Output Modalities
Text
Tool Use
Function Calling ✓
Base Model
Gemini 3.6 Flash
Model ID
gemini-3.7-flash

Sources: Google DeepMind Model Card · Google Blog (model card published Aug 13, 2026; blog announcement Aug 14, 2026)

Performance — Big Coding & Agent Gains Over 3.6 Flash

An algorithmic improvement built on Gemini 3.6 Flash, with gains across coding, agents, and knowledge work.

  • FrontierCode 1.1 Main: 43.6% — production code quality, up from 3.6 Flash (34.4%). Above Claude Sonnet 5 (42.7%) and GPT-5.6 Terra (41.3%).
  • DeepSWE v1.1: 65.3% — long-horizon software engineering, +16.7 pts over 3.6 Flash (48.6%).
  • Code Arena (WebDev) Elo: 1588 — above 3.6 Flash (1538).
  • Terminal-bench 2.1: 85.8% — agentic terminal coding (3.6 Flash: 78.0%).
  • GDP.pdf: 34.0% — expert PDF comprehension, far above 3.6 Flash (22.0%).
  • AutomationBench: 30.4% — enterprise workflow automation (3.6 Flash: 17.0%).
  • Harvey LAB-AA: 90.7% — complex legal workflows, above Claude Sonnet 5 (90.1%).
  • AA Intelligence Index: 56 — up from 3.6 Flash (52).

Google positions 3.7 Flash as "our most capable Flash model for agentic workflows and multimodal reasoning," with customizable thinking configurations to control quality, cost, and latency.

Sources: Model Card (evals) · Google Blog

Competitor Price Comparison

The intro price of $0.75/$3.75 is among the cheapest in the Flash tier. All prices in USD per 1M tokens (Standard, text input).

ModelProviderInputOutputvs 3.7 FlashNotes
Gemini 3.7 FlashGoogle$0.75→$1.50$3.75→$7.50Intro/standard. Multimodal
Gemini 3.6 FlashGoogle$0.75→$1.50$3.75→$7.50SameSame intro price applies
Gemini 3.5 FlashGoogle$1.50$9.003.7 cheaper50% cheaper input, 58% cheaper output (intro)
GPT-5.6 LunaOpenAI$0.20$1.20Luna cheaperText + image only. Cheapest in absolute terms
Grok 4.6xAI$2.00$6.003.7 cheaper62% cheaper input, 37% cheaper output (intro)
GPT-5.6 TerraOpenAI$2.00$12.003.7 cheaper62% cheaper input, 68% cheaper output (intro)
Claude Sonnet 5Anthropic$2.00$10.003.7 cheaperIntro thru 8/31. 62% cheaper output (intro)

* Gemini 3.7 Flash's intro price ($0.75/$3.75) runs through Dec 31, 2026; standard $1.50/$7.50 follows. GPT-5.6 Luna ($0.20/$1.20) and Gemini 3.1 Flash-Lite ($0.25/$1.50) are cheaper in absolute terms but serve high-throughput niches; 3.7 Flash targets high-difficulty coding and agentic work.

Sources: official pricing pages from each provider (verified 2026-08-17). See full pricing comparison. Grok 4.6 details →

Ideal Use Cases

Coding, agents, and knowledge work.

💻

Coding & Code Generation

Strong FrontierCode/DeepSWE. Better first-pass accuracy and production quality.

🤖

Agentic Workflows

Multi-step planning and tool use with more persistent reasoning. Less manual oversight.

📄

Document Processing

High GDP.pdf score. Understand and transform complex PDFs like annual reports.

🏦

Finance, Law & Bio

Strong Harvey LAB-AA / BioMysteryBench. Expert-domain reasoning and accuracy.

🌐

Web Dev & UI Generation

Practical layouts and features from fewer prompts. Code Arena Elo 1588.

🔍

Long-Context Processing

97.0% on GDM-MRCR v2. 1M context for large-document Q&A.

Gemini 3.x Family Selection Guide

ModelInputOutputWhen to Use
Gemini 3.1 Pro Preview$2.00$12.00Hardest reasoning, agents, complex multi-step problems
Gemini 3.7 Flash ⬅$0.75→$1.50$3.75→$7.50Coding, agents, knowledge work (half price thru 2026)
Gemini 3.6 Flash$0.75→$1.50$3.75→$7.50Speed + Grounding (same intro price as 3.7)
Gemini 3.5 Flash$1.50$9.00Standard Flash tier. Slightly higher output than 3.7
Gemini 3.1 Flash-Lite$0.25$1.50High-throughput, cost-sensitive, multimodal bulk processing

💡 Strategy: use Flash-Lite for routine high-frequency tasks, 3.7 Flash for production coding/agents, and Pro for the hardest problems. 3.7/3.6 are available at $0.75/$3.75 through the end of 2026.

Estimate Your Actual Costs

Compare token costs for Gemini 3.7 Flash and all major models. Just enter your token counts.

🧮 Open Token Calculator →

📊 Full Pricing Comparison · 🚀 Compare with Grok 4.6 · 🗓️ Release Timeline · 🔗 Google Official Pricing

⚠️ Disclaimer

  • Pricing verified against Google's official AI page (ai.google.dev) on 2026-08-17. The intro price runs through Dec 31, 2026, then $1.50/$7.50. Always check official pages before contracting.
  • This site is informational only. We are not affiliated with or endorsed by any provider.
  • Performance claims are based on Google's published data. Your mileage may vary.