🚀 Released September 21, 2026 — $2/$6 unchanged · Fast costs 2x

Grok 4.7

xAI’s (x.ai) newest flagship. It uses a larger base model than Grok 4.6 while keeping the same $2/$6 pricing. Grok 4.7 beats Grok 4.6 on all seven official benchmarks and leads the four-model field on EEBench and the Harvey Legal Agent Benchmark.

$2.00
Input / 1M tokens
$6.00
Output / 1M tokens
500K
Context window
Text+Image
Input modality
●Last updated: | Sources: x.ai announcement · docs.x.ai model card · docs.x.ai pricing | Pricing · Calculator · Grok 4.6

Pricing (as of the September 21, 2026 launch)

Verified on the docs.x.ai API pricing page. All amounts in USD per 1M tokens. Standard pricing is unchanged from Grok 4.6. "Cached input" is the cheaper rate applied when prompt content is served from the cache (prompt caching explained).

TierInputCached inputOutputNotes
Standard (short <200K)$2.00$0.50$6.00Default rate, same as Grok 4.6
Long context (over 200K)$4.00$1.00$12.00Applies to the whole request

Grok 4.7 Fast (Cursor and Grok Build only — the same Grok 4.7 served on faster infrastructure at twice the standard token rates)

Prompt lengthInputCached inputOutput
Below 200K$4.00$1.00$12.00
Above 200K$6.00$1.50$18.00

Thresholds follow the official wording (below 200k / above 200k). For the exact treatment of a prompt of exactly 200K tokens, see the docs.x.ai pricing table.

⚠️ The 200K long-context boundary: once a prompt reaches 200K tokens, the long-context rates ($4/$12) are billed for every token in that request (stated explicitly on docs.x.ai). Agent loops that accumulate tool output can cross this line and see the bill jump sharply.
💡 Three things that decide your real cost: (1) Cached input is $0.50 — a 75% discount off standard input — which matters a lot when you resend the same system prompt or repository context. (2) The 20% Batch API discount covers only grok-4.3 and the grok-4.20 family; grok-4.7 is not eligible (verified in the docs.x.ai Batch discount table). (3) The US regional endpoint (us.api.x.ai) bills all token rates at 1.1x: for grok-4.7 that is $2.20/$0.55/$6.60 below 200K and $4.40/$1.10/$13.20 above.
🧩 Server-side tools are billed separately (docs.x.ai pricing): web_search $5 / 1k calls, x_search $5 / 1k posts and $10 / 1k profiles, code_execution $5 / 1k calls, attachment_search $10 / 1k calls, collections_search (file_search) $2.50 / 1k calls. File storage $0.025/GiB/day, collection storage $0.10/GiB/day, downloads $0.20/GiB.

Primary sources: docs.x.ai/developers/pricing (Last updated: September 21, 2026) · docs.x.ai Grok 4.7 · x.ai/news/grok-4-7 (checked 2026-09-23)

Worked cost examples (standard tier, USD)

Input rate x input tokens + output rate x output tokens. Once a request passes the 200K long-context threshold, the higher rates apply to the entire request.

CaseInput tokensOutput tokensEstimated cost (USD)
Short exchange100,00010,000$0.26
Same input, 80% cached100,000 (80,000 cache hits)10,000$0.14
Exactly 200K (boundary)200,00020,000$0.52 → $1.04 if long-context applies
Long context (above 200K)250,00020,000$1.24

Excludes cache-write charges, server-side tools and the 1.1x US regional multiplier. For your own token counts, use the token calculator.

Model specifications

GA release
2026-09-21
Knowledge cutoff
May 2026
Context window
500,000 tokens
Long-context threshold
200,000 tokens
Input modalities
Text + Image
Output modalities
Text
Reasoning effort
low / medium / high / xhigh (xhigh reasons deepest)
Default effort
high
Function calling
Supported ✓
Structured outputs
Supported
Region
us-east-1 (+ US regional)
Model ID
grok-4.7

Sources: docs.x.ai Grok 4.7 · xAI announcement. xAI says the model “uses a new, larger base model compared to Grok 4.6” and is better at verifying its own work and managing longer context. Note: the x.ai and docs.x.ai sites are branded “SpaceXAI” as of 2026-09-23; because a corporate rename has not been confirmed first-hand, this page refers to the provider as xAI.

Performance — ahead of Grok 4.6 on all seven official benchmarks

From xAI’s published benchmark table, comparing Grok 4.7, Grok 4.6, GPT-5.6 Sol and Fable 5.1 (Anthropic). These are vendor-reported figures and test conditions are not necessarily identical.

Benchmark (domain)Grok 4.7Grok 4.6GPT-5.6 SolFable 5.1Note
CursorBench 4.0 (SWE)46.3%40.4%41.7%51.8%+5.9pt vs. previous gen
DeepSWE v1.1 (long-horizon SWE)71.0%*65.2%72.7%70.0%2nd of four (* high effort)
EEBench (electrical engineering)64.0%53.0%39.4%56.4%Best of four
AA Briefcase v1.1 (multi-hour office work, Elo)1,6571,5461,4871,678Fable 5.1 slightly ahead
Terminal-Bench 4.0 (terminal work)37.6%20.3%37.3%57.9%~1.9x vs. previous gen
Harvey Legal Agent Benchmark (legal)19.6%15.8%2.5%6.7%Best of four (~7.8x GPT-5.6 Sol)
HealthBench Professional (clinical reasoning)56.7%48.5%60.5%62.1%3rd of four

“Winner” marks the top score in each benchmark. DeepSWE v1.1 is Datacurve’s long-horizon software engineering benchmark (113 tasks across 5 languages).

  • GDPval (Elo): Grok 4.7 (xhigh) 1,695 — vs. Fable 5.1 (max) 1,735, Grok 4.6 (high) 1,605 and GPT-6 Astra (max) 1,542. That is +90 over the previous generation, with xAI highlighting better documents and presentations.
  • Pricing unchanged: $2 input / $6 output (xHigh), the same as Grok 4.6 (High) in the vendor’s own chart. The same chart lists GPT-5.6 Sol Max at $4/$20 and Fable 5.1 Max at $10/$50.
  • Bottom line: the vendor claim is “improved across the board at the same price.” But in the four-model field Fable 5.1 still leads three of the seven benchmarks and GPT-5.6 Sol one, so Grok 4.7 is not the outright leader.

Source: xAI: Introducing Grok 4.7 (benchmark table and GDPval chart)

Third-party view — affordable, but not at the frontier

Independent measurements are worth reading alongside the vendor’s numbers.

  • Mid-pack on the composite intelligence index: Grok 4.7 scores 46 on Artificial Analysis’ Intelligence Index, while Claude Fable 5.1 and GPT-6 each score 53, according to reporting by the-decoder (2026-09-21). The reported gap widens further on agentic coding.
  • Token consumption can outweigh the low unit price: VentureBeat notes (2026-09-21) that a model with cheap per-token pricing can still cost more per finished workload if it burns substantially more reasoning tokens. Compare measured token usage, not just list prices.
  • Wide distribution: Grok 4.7 began rolling out in GitHub Copilot on 2026-09-21 (VS Code, JetBrains, Xcode and more). It is also available via the Grok API, Cursor, Grok Build, third-party coding harnesses, and model routers and cloud platforms.

Sources: the-decoder · VentureBeat · GitHub Changelog (all checked 2026-09-23). Third-party figures depend on each tester’s methodology.

How it compares

Based on xAI’s own comparison chart. Each model is shown at a different reasoning-effort setting, so this is not a like-for-like comparison.

ModelSettingInputOutputVs. Grok 4.7
Grok 4.7xHigh$2.00$6.00— (baseline)
Grok 4.6High$2.00$6.00Same
GPT-5.6 SolMax$4.00$20.00Grok cheaper (half input, 70% less output)
Claude Fable 5.1Max$10.00$50.00Grok cheaper (1/5 input, ~1/8 output)

The table above reflects xAI’s published chart, i.e. top reasoning-effort tiers; it may differ from the standard rates we verify independently on each vendor’s pricing page. See our pricing comparison for standard rates across all models.

Source: xAI: Introducing Grok 4.7 · standard vendor rates via our pricing page (as of 2026-09-23)

Where the benchmarks show strength

Coding, agentic work and specialist domains.

💻

Coding & code review

46.3% on CursorBench 4.0 (+5.9pt over the previous generation). Available day one in Cursor and Grok Build.

🖥️

Long terminal sessions

Terminal-Bench 4.0 jumps from 20.3% to 37.6% (~1.9x). Good for multi-step shell automation.

⚖️

Legal & specialist domains

19.6% on the Harvey Legal Agent Benchmark — best of the four models compared.

🔌

Electrical engineering

64.0% on EEBench — best of the four models compared.

📊

Documents & presentations

1,657 on AA Briefcase and 1,695 Elo on GDPval; xAI highlights improvement on professional deliverables.

🧠

Long-context workloads

500K context window — but rates double above 200K, so manage the length/cost trade-off.

When Grok 4.7 is a good fit — and when it isn't

  • A strong price-per-token pick at the frontier tier: $2 input / $6 output is clearly cheaper than GPT-5.6 Sol or Fable 5.1 at their top effort settings. xAI claims “twice as fast, at half the price of comparable models.”
  • Judge total cost by measured tokens: high reasoning effort (default high up to xhigh) inflates output tokens and narrows the price gap. Lower reasoning_effort per task difficulty.
  • Mind the 200K wall: long-context calls bill at $4/$12 ($6/$1.50/$18 with Fast). Using cached input ($0.50) is the main lever for real savings.
  • No Batch discount: the 20% Batch API discount covers only grok-4.3 and the grok-4.20 family, so grok-4.7 is not eligible.
  • If you need the absolute best: on the Artificial Analysis Intelligence Index cited here, Claude Fable 5.1 and GPT-6 (53 each) score ahead of Grok 4.7 (46); consider routing the hardest work to those.

Estimate your actual cost

Calculate token costs for Grok 4.7 and every other model at once — just enter your input and output token counts.

🧮 Open the token calculator →

📊 All-model pricing · 🤖 Compare with Grok 4.6 · 🗓️ Release timeline · 🔗 xAI official pricing

⚠️ Disclaimer

  • Pricing was verified on the official docs.x.ai pages on 2026-09-23 and may change without notice. Always confirm on the official pages before you commit.
  • This site is for information only and is not a recommendation or agent for any provider.
  • Benchmark figures are vendor-reported and do not guarantee real-world results; third-party figures depend on each tester’s methodology.