Pricing (intro vs. standard)
Verified against Google's official pricing page (ai.google.dev). All prices in USD per 1M tokens. Output pricing includes thinking tokens.
| Tier | Input | Output | Cached Input | Notes |
|---|---|---|---|---|
| Standard (intro, thru 12/31) | $0.75 | $3.75 | $0.075 | Through Dec 31, 2026 |
| Standard (from 1/1) | $1.50 | $7.50 | $0.15 | From Jan 1, 2027 |
| Batch / Flex (intro, thru 12/31) | $0.375 | $1.875 | $0.0375 | Half of Standard |
| Priority (intro, thru 12/31) | $1.35 | $6.75 | $0.135 | High priority. 1.8× Standard |
Primary sources: ai.google.dev/gemini-api/docs/pricing · Google Blog: Introducing Gemini 3.7 Flash (verified 2026-08-17)
Model Specifications
Sources: Google DeepMind Model Card · Google Blog (model card published Aug 13, 2026; blog announcement Aug 14, 2026)
Performance — Big Coding & Agent Gains Over 3.6 Flash
An algorithmic improvement built on Gemini 3.6 Flash, with gains across coding, agents, and knowledge work.
- FrontierCode 1.1 Main: 43.6% — production code quality, up from 3.6 Flash (34.4%). Above Claude Sonnet 5 (42.7%) and GPT-5.6 Terra (41.3%).
- DeepSWE v1.1: 65.3% — long-horizon software engineering, +16.7 pts over 3.6 Flash (48.6%).
- Code Arena (WebDev) Elo: 1588 — above 3.6 Flash (1538).
- Terminal-bench 2.1: 85.8% — agentic terminal coding (3.6 Flash: 78.0%).
- GDP.pdf: 34.0% — expert PDF comprehension, far above 3.6 Flash (22.0%).
- AutomationBench: 30.4% — enterprise workflow automation (3.6 Flash: 17.0%).
- Harvey LAB-AA: 90.7% — complex legal workflows, above Claude Sonnet 5 (90.1%).
- AA Intelligence Index: 56 — up from 3.6 Flash (52).
Google positions 3.7 Flash as "our most capable Flash model for agentic workflows and multimodal reasoning," with customizable thinking configurations to control quality, cost, and latency.
Sources: Model Card (evals) · Google Blog
Competitor Price Comparison
The intro price of $0.75/$3.75 is among the cheapest in the Flash tier. All prices in USD per 1M tokens (Standard, text input).
| Model | Provider | Input | Output | vs 3.7 Flash | Notes |
|---|---|---|---|---|---|
| Gemini 3.7 Flash | $0.75→$1.50 | $3.75→$7.50 | — | Intro/standard. Multimodal | |
| Gemini 3.6 Flash | $0.75→$1.50 | $3.75→$7.50 | Same | Same intro price applies | |
| Gemini 3.5 Flash | $1.50 | $9.00 | 3.7 cheaper | 50% cheaper input, 58% cheaper output (intro) | |
| GPT-5.6 Luna | OpenAI | $0.20 | $1.20 | Luna cheaper | Text + image only. Cheapest in absolute terms |
| Grok 4.6 | xAI | $2.00 | $6.00 | 3.7 cheaper | 62% cheaper input, 37% cheaper output (intro) |
| GPT-5.6 Terra | OpenAI | $2.00 | $12.00 | 3.7 cheaper | 62% cheaper input, 68% cheaper output (intro) |
| Claude Sonnet 5 | Anthropic | $2.00 | $10.00 | 3.7 cheaper | Intro thru 8/31. 62% cheaper output (intro) |
* Gemini 3.7 Flash's intro price ($0.75/$3.75) runs through Dec 31, 2026; standard $1.50/$7.50 follows. GPT-5.6 Luna ($0.20/$1.20) and Gemini 3.1 Flash-Lite ($0.25/$1.50) are cheaper in absolute terms but serve high-throughput niches; 3.7 Flash targets high-difficulty coding and agentic work.
Sources: official pricing pages from each provider (verified 2026-08-17). See full pricing comparison. Grok 4.6 details →
Ideal Use Cases
Coding, agents, and knowledge work.
Coding & Code Generation
Strong FrontierCode/DeepSWE. Better first-pass accuracy and production quality.
Agentic Workflows
Multi-step planning and tool use with more persistent reasoning. Less manual oversight.
Document Processing
High GDP.pdf score. Understand and transform complex PDFs like annual reports.
Finance, Law & Bio
Strong Harvey LAB-AA / BioMysteryBench. Expert-domain reasoning and accuracy.
Web Dev & UI Generation
Practical layouts and features from fewer prompts. Code Arena Elo 1588.
Long-Context Processing
97.0% on GDM-MRCR v2. 1M context for large-document Q&A.
Gemini 3.x Family Selection Guide
| Model | Input | Output | When to Use |
|---|---|---|---|
| Gemini 3.1 Pro Preview | $2.00 | $12.00 | Hardest reasoning, agents, complex multi-step problems |
| Gemini 3.7 Flash ⬅ | $0.75→$1.50 | $3.75→$7.50 | Coding, agents, knowledge work (half price thru 2026) |
| Gemini 3.6 Flash | $0.75→$1.50 | $3.75→$7.50 | Speed + Grounding (same intro price as 3.7) |
| Gemini 3.5 Flash | $1.50 | $9.00 | Standard Flash tier. Slightly higher output than 3.7 |
| Gemini 3.1 Flash-Lite | $0.25 | $1.50 | High-throughput, cost-sensitive, multimodal bulk processing |
💡 Strategy: use Flash-Lite for routine high-frequency tasks, 3.7 Flash for production coding/agents, and Pro for the hardest problems. 3.7/3.6 are available at $0.75/$3.75 through the end of 2026.
Estimate Your Actual Costs
Compare token costs for Gemini 3.7 Flash and all major models. Just enter your token counts.
🧮 Open Token Calculator →📊 Full Pricing Comparison · 🚀 Compare with Grok 4.6 · 🗓️ Release Timeline · 🔗 Google Official Pricing
⚠️ Disclaimer
- Pricing verified against Google's official AI page (ai.google.dev) on 2026-08-17. The intro price runs through Dec 31, 2026, then $1.50/$7.50. Always check official pages before contracting.
- This site is informational only. We are not affiliated with or endorsed by any provider.
- Performance claims are based on Google's published data. Your mileage may vary.