Major LLM API Pricing Comparison [August 2026]

We reference each provider's official pricing page and list only re-verified figures. All prices in USD / 1M tokens. Use English calculator to estimate costs. See GPT-5.6 Luna deep dive →. 日本語版

Page updated: (pricing verified 2026-08-14 JST for DeepSeek/Gemini, 2026-08-12 for Grok, others 2026-08-04 JST) | Anthropic · Google · OpenAI · DeepSeek · xAI

OpenAI

📋 Standard, short-context rate. Long-context: see official page

✅ Verified via Firecrawl (2026-08-04 JST). GPT-5.6 Terra/Luna price cuts (July 30, 2026) reflected.
ModelInputCached InputOutputNotes
GPT-5.6 Sol$5.00$0.50$30.00Current flagship. Long: $10/$1/$45
GPT-5.6 Terra $2.00$0.20$12.00Cut 2026-07-30 (was $2.50/$15). Long: $4/$0.40/$18
GPT-5.6 Luna $0.20$0.02$1.20Cut 2026-07-30 (was $1/$6). Long: $0.40/$0.04/$1.80 →Deep dive
GPT-5.5$5.00$0.50$30.00Previous flagship
GPT-5.4$2.50$0.25$15.00Balanced (superseded by Terra)
GPT-5.4 mini$0.75$0.075$4.50Lightweight tier
GPT-5.4 nano$0.20$0.02$1.25Cheapest tier

Primary source: developers.openai.com/api/docs/pricing (2026-08-04 JST). Priority processing renamed to Fast mode. 📄 GPT-5.6 Luna: full pricing, benchmarks & competitor comparison →

Anthropic (Claude)

📋 Standard API / text input / no cache

✅ Verified via Firecrawl (2026-08-04 JST). Claude Opus 5 (released July 24, 2026) reflected.
ModelInputCache Write 5mCache HitOutputNotes
Claude Fable 5$10.00$12.50$1.00$50.00Top-tier flagship. New tokenizer.
Claude Opus 5 New$5.00$6.25$0.50$25.00Released 2026-07-24. Same price as Opus 4.8. Fast mode $10/$50
Claude Opus 4.8$5.00$6.25$0.50$25.00Previous Opus. Fallback for Claude Max
Claude Sonnet 5$2.00 until 8/31
$3.00(9/1~)
$2.50
$3.75(9/1~)
$0.20
$0.30(9/1~)
$10.00 until 8/31
$15.00(9/1~)
Introductory pricing period
Claude Sonnet 4.6$3.00$3.75$0.30$15.00Older balanced model
Claude Haiku 4.5$1.00$1.25$0.10$5.00High-volume, classification

Primary source: platform.claude.com (2026-08-04 JST). Batch 50% off. Cache hit 90% off. Flat rate full 1M context.

Google (Gemini Developer API)

📋 Paid tier / Standard / text, image, video input / standard API

✅ Verified via Firecrawl (2026-08-14 JST). Gemini 3.7 Flash (Aug 14) and the Gemini 3.6 Flash introductory price cut reflected.
ModelInputInput>200KOutputOutput>200KNotes
Gemini 3.7 Flash New$0.75 until 12/31
$1.50(1/1~)
$3.75 until 12/31
$7.50(1/1~)
Released 2026-08-14. Most capable Flash for coding/agents →Deep dive
Gemini 3.6 Flash $0.75 until 12/31
$1.50(1/1~)
$3.75 until 12/31
$7.50(1/1~)
Cut to introductory price (was $1.50/$7.50)
Gemini 3.5 Flash$1.50$9.00GA May 2026
Gemini 3.5 Flash-Lite$0.30$2.50GA July 2026. Cheapest Gemini 3.x
Gemini 3.1 Pro Preview Preview$2.00$4.00$12.00$18.00Preview pricing
Gemini 3.1 Flash-Lite$0.25$1.50Audio $0.50
Gemini 2.5 Pro$1.25$2.50$10.00$15.00
Gemini 2.5 Flash$0.30$2.50Audio $1.00
Gemini 2.5 Flash-Lite$0.10$0.40Audio $0.30

Gemini 3.7/3.6 Flash introductory price ($0.75/$3.75) valid through Dec 31, 2026; standard price $1.50/$7.50 from Jan 1, 2027. Primary source: ai.google.dev (2026-08-14 JST)

DeepSeek

📋 Peak/off-peak pricing (effective 2026-08-16 16:00 UTC). Cache miss/hit differ

✅ Verified via Firecrawl (2026-08-14). Moved to official V4-Pro (-0813) / V4-Flash (-0731) with peak/off-peak pricing. Peak: 01:00-04:00 and 06:00-10:00 UTC; off-peak is half the peak rate.
ModelInput (miss)Input (hit)OutputNotes
DeepSeek V4 Pro New
(-0813)
$0.66 (off-peak)
$1.32 (peak)
$0.022 (off-peak)
$0.044 (peak)
$1.98 (off-peak)
$3.96 (peak)
1M ctx, 384K out. GA
DeepSeek V4 Flash New
(-0731)
$0.22 (off-peak)
$0.44 (peak)
$0.007 (off-peak)
$0.014 (peak)
$0.66 (off-peak)
$1.32 (peak)
Lightweight. Auto disk cache

Peak hours: 01:00-04:00 and 06:00-10:00 UTC (JST 10:00-13:00 / 15:00-19:00; all other hours off-peak at half the peak rate). Effective 2026-08-16 16:00 UTC (JST 8/17 01:00). Old promo rates (V4 Pro $0.435/$0.87, V4 Flash $0.14/$0.28) retired. deepseek-chat/reasoner aliases retired 2026-07-24. Source: api-docs.deepseek.com (2026-08-14 JST) · V4-Pro GA release notes

xAI (Grok)

📋 Standard API rate. Fast variant is 2× (see official docs for cache/other modes)

✅ From x.ai official announcement (2026-08-12). Grok 4.6 released. Input $2 / output $6 (fast variant 2×) as stated in the announcement.
ModelInputOutputNotes
Grok 4.6 New$2.00$6.00Released 2026-08-12. Fast variant 2× ($4/$12) →Deep dive

Primary source: x.ai/news/grok-4-6 ("Pricing starts at $2 per million input tokens and $6 per million output tokens") · x.ai/pricing (consumer/enterprise plans; see docs.x.ai for full API details)

🔍 Spotlight

DeepSeek V4 Pro — now peak/off-peak (effective 8/16)

Old promo ($0.435/$0.87) retired. New: $0.66/$1.98 off-peak, $1.32/$3.96 peak (cache-miss input / output). Shift heavy jobs to off-peak to contain costs.

⚠️ Disclaimer

  • Pricing verified 2026-08-14 JST (DeepSeek/Gemini), 2026-08-12 (Grok), others 2026-08-04 JST.
  • Pricing may change without notice.
  • DeepSeek peak/off-peak pricing takes effect 2026-08-16 16:00 UTC. Always verify on the official page around the cutover.
  • This site is informational only.