✅ Verified against primary sources · Updated continuously · No ads

Find LLM API pricing & model data
quickly and accurately

Token pricing verified against official OpenAI, Anthropic, Google Gemini and other provider pages.
A practical hub for cost decisions by engineers and product teams.

See pricing comparison →🧮 Estimate with the calculator →

日本語: 日本語トップページ →

Last updated: GLM 5.3 Flash · GPT-5.6 Sol promo · China AI overview · Grok 4.6 · Gemini 3.7 Flash

Content

Engineer-focused LLM cost data, verified against official primary sources

📊

Major LLM API Pricing Comparison

OpenAI (GPT-5.6 Sol/Terra/Luna, GPT-5.5/5.4), Anthropic Claude (Opus 5/4.8, Fable 5, Sonnet 5, Haiku 4.5), Google Gemini (3.7/3.6/3.5 Flash, 3.1 Flash-Lite), DeepSeek V4 Pro/Flash and Z.ai GLM — in USD per 1M tokens, verified against each provider's official page.

See pricing comparison →→ 日本語版
🧮

Token Cost Calculator

Enter input and output token counts to instantly compare costs across major LLM APIs in USD (plus JPY reference). No network calls — everything runs in your browser.

Open the calculator →→ 日本語版
🗓️

Major LLM Release Timeline

From DeepSeek R1 (Jan 2025) to GPT-5.6, Claude Opus 5, Gemini 3.7 Flash and GLM 5.3 Flash — major model releases verified against primary sources and organized chronologically, each with an official link.

See the timeline →→ 日本語版

GLM 5.3 Flash — frontier intelligence at Flash cost

Z.ai (Zhipu)'s first natively multimodal GLM-5 model. 320B MoE (18B active) beating GLM-5.2 at ~1/10 the price. Input $0.075 / output $0.25 (promo thru 9/9). Specs, benchmarks and comparisons.

GLM 5.3 Flash details →→ 日本語版
🌙

GPT-5.6 Luna — 80% price cut, best cost-per-performance lightweight model

Dropped to $0.20/$1.20. Cheaper than Gemini Flash-Lite with GPT-5.5-beating performance. Competitor comparisons, use cases and a quick selection guide.

Luna details →→ 日本語版
☀️

GPT-5.6 Sol — OpenAI's flagship (promo price)

Cut to $4/$20 (at least through 2026-11-21). 1.05M context, 128K output, top-tier reasoning. Comparison vs Claude Opus 5 and the Cloudflare AI Gateway 50% off.

Sol details →→ 日本語版

Gemini 3.1 Flash-Lite — Google's cheapest, fastest high-throughput model

$0.25/$1.50, multimodal. Built on the Gemini 3 Pro architecture with 2.5 Flash-level quality. Deep comparison vs GPT-5.6 Luna and thinking-level control.

Flash-Lite details →→ 日本語版
🚀

Grok 4.6 — xAI's new flagship for code & agents

$2/$6 standard, fast variant 2×, 500K context. AA Intelligence Index 61, tying GPT-5.6 Sol. Pricing, performance and comparisons from primary sources.

Grok 4.6 details →→ 日本語版
💻

Gemini 3.7 Flash — the best Flash for coding & agents (half-price launch)

$0.75/$3.75 launch (thru 12/31) → standard $1.50/$7.50. 1M context, 64K output. FrontierCode 43.6% · DeepSWE 65.3%.

3.7 Flash details →→ 日本語版
🐋

DeepSeek V4 Pro × GitHub Copilot practical guide

Setup, official pricing, and a real-world comparison vs Claude Sonnet, with measured cost examples from the operator's own use.

Practical guide →
🇨🇳

Chinese AI Models Overview (permanent page)

DeepSeek, Qwen (Alibaba), GLM (Zhipu), Kimi (Moonshot) and MiniMax at a glance. Indicative pricing, open-weight status and an update log. Only what can be verified against primary sources.

China AI overview →→ 日本語版

About this site

Data quality policy

📌 Listing criteria

  • Primary sources only — We reference each provider's official pricing page and list only verified figures
  • Sources cited — Every table links its source URL and retrieval date
  • "Verified only" principle — Unverified figures are marked "see official page"
  • Continuously updated — Last-updated dates are always shown
  • No ads — No bias toward any single provider