Major LLM API Pricing Comparison [August 2026]
We reference each provider's official pricing page and list only re-verified figures. All prices in USD / 1M tokens. Use English calculator to estimate costs. See GPT-5.6 Luna deep dive →. 日本語版
OpenAI
📋 Standard, short-context rate. Long-context: see official page
| Model | Input | Cached Input | Output | Notes |
|---|---|---|---|---|
| GPT-5.6 Sol | $5.00 | $0.50 | $30.00 | Current flagship. Long: $10/$1/$45 |
| GPT-5.6 Terra ↓ | $2.00 | $0.20 | $12.00 | Cut 2026-07-30 (was $2.50/$15). Long: $4/$0.40/$18 |
| GPT-5.6 Luna ↓ | $0.20 | $0.02 | $1.20 | Cut 2026-07-30 (was $1/$6). Long: $0.40/$0.04/$1.80 →Deep dive |
| GPT-5.5 | $5.00 | $0.50 | $30.00 | Previous flagship |
| GPT-5.4 | $2.50 | $0.25 | $15.00 | Balanced (superseded by Terra) |
| GPT-5.4 mini | $0.75 | $0.075 | $4.50 | Lightweight tier |
| GPT-5.4 nano | $0.20 | $0.02 | $1.25 | Cheapest tier |
Primary source: developers.openai.com/api/docs/pricing (2026-08-04 JST). Priority processing renamed to Fast mode. 📄 GPT-5.6 Luna: full pricing, benchmarks & competitor comparison →
Anthropic (Claude)
📋 Standard API / text input / no cache
| Model | Input | Cache Write 5m | Cache Hit | Output | Notes |
|---|---|---|---|---|---|
| Claude Fable 5 | $10.00 | $12.50 | $1.00 | $50.00 | Top-tier flagship. New tokenizer. |
| Claude Opus 5 New | $5.00 | $6.25 | $0.50 | $25.00 | Released 2026-07-24. Same price as Opus 4.8. Fast mode $10/$50 |
| Claude Opus 4.8 | $5.00 | $6.25 | $0.50 | $25.00 | Previous Opus. Fallback for Claude Max |
| Claude Sonnet 5 | $2.00 until 8/31 $3.00(9/1~) | $2.50 $3.75(9/1~) | $0.20 $0.30(9/1~) | $10.00 until 8/31 $15.00(9/1~) | Introductory pricing period |
| Claude Sonnet 4.6 | $3.00 | $3.75 | $0.30 | $15.00 | Older balanced model |
| Claude Haiku 4.5 | $1.00 | $1.25 | $0.10 | $5.00 | High-volume, classification |
Primary source: platform.claude.com (2026-08-04 JST). Batch 50% off. Cache hit 90% off. Flat rate full 1M context.
Google (Gemini Developer API)
📋 Paid tier / Standard / text, image, video input / standard API
| Model | Input | Input>200K | Output | Output>200K | Notes |
|---|---|---|---|---|---|
| Gemini 3.7 Flash New | $0.75 until 12/31 $1.50(1/1~) | — | $3.75 until 12/31 $7.50(1/1~) | — | Released 2026-08-14. Most capable Flash for coding/agents →Deep dive |
| Gemini 3.6 Flash ↓ | $0.75 until 12/31 $1.50(1/1~) | — | $3.75 until 12/31 $7.50(1/1~) | — | Cut to introductory price (was $1.50/$7.50) |
| Gemini 3.5 Flash | $1.50 | — | $9.00 | — | GA May 2026 |
| Gemini 3.5 Flash-Lite | $0.30 | — | $2.50 | — | GA July 2026. Cheapest Gemini 3.x |
| Gemini 3.1 Pro Preview Preview | $2.00 | $4.00 | $12.00 | $18.00 | Preview pricing |
| Gemini 3.1 Flash-Lite | $0.25 | — | $1.50 | — | Audio $0.50 |
| Gemini 2.5 Pro | $1.25 | $2.50 | $10.00 | $15.00 | |
| Gemini 2.5 Flash | $0.30 | — | $2.50 | — | Audio $1.00 |
| Gemini 2.5 Flash-Lite | $0.10 | — | $0.40 | — | Audio $0.30 |
Gemini 3.7/3.6 Flash introductory price ($0.75/$3.75) valid through Dec 31, 2026; standard price $1.50/$7.50 from Jan 1, 2027. Primary source: ai.google.dev (2026-08-14 JST)
DeepSeek
📋 Peak/off-peak pricing (effective 2026-08-16 16:00 UTC). Cache miss/hit differ
| Model | Input (miss) | Input (hit) | Output | Notes |
|---|---|---|---|---|
| DeepSeek V4 Pro New (-0813) | $0.66 (off-peak) $1.32 (peak) | $0.022 (off-peak) $0.044 (peak) | $1.98 (off-peak) $3.96 (peak) | 1M ctx, 384K out. GA |
| DeepSeek V4 Flash New (-0731) | $0.22 (off-peak) $0.44 (peak) | $0.007 (off-peak) $0.014 (peak) | $0.66 (off-peak) $1.32 (peak) | Lightweight. Auto disk cache |
⏰ Peak hours: 01:00-04:00 and 06:00-10:00 UTC (JST 10:00-13:00 / 15:00-19:00; all other hours off-peak at half the peak rate). Effective 2026-08-16 16:00 UTC (JST 8/17 01:00). Old promo rates (V4 Pro $0.435/$0.87, V4 Flash $0.14/$0.28) retired. deepseek-chat/reasoner aliases retired 2026-07-24. Source: api-docs.deepseek.com (2026-08-14 JST) · V4-Pro GA release notes
xAI (Grok)
📋 Standard API rate. Fast variant is 2× (see official docs for cache/other modes)
| Model | Input | Output | Notes |
|---|---|---|---|
| Grok 4.6 New | $2.00 | $6.00 | Released 2026-08-12. Fast variant 2× ($4/$12) →Deep dive |
Primary source: x.ai/news/grok-4-6 ("Pricing starts at $2 per million input tokens and $6 per million output tokens") · x.ai/pricing (consumer/enterprise plans; see docs.x.ai for full API details)
DeepSeek V4 Pro — now peak/off-peak (effective 8/16)
Old promo ($0.435/$0.87) retired. New: $0.66/$1.98 off-peak, $1.32/$3.96 peak (cache-miss input / output). Shift heavy jobs to off-peak to contain costs.
⚠️ Disclaimer
- Pricing verified 2026-08-14 JST (DeepSeek/Gemini), 2026-08-12 (Grok), others 2026-08-04 JST.
- Pricing may change without notice.
- DeepSeek peak/off-peak pricing takes effect 2026-08-16 16:00 UTC. Always verify on the official page around the cutover.
- This site is informational only.