Last verified
NEWEST GROK500K CONTEXTTEXT + IMAGE INPUTREASONING TO XHIGH2x ABOVE 200K PROMPTUS REGIONS ONLY

Grok 4.6 API Pricing

Grok 4.6 is SpaceXAI's newest frontier model for coding, agentic tasks, and knowledge work, released August 12, 2026. The pricing page lists $2.00/M input, $0.50/M cached input, and $6.00/M output for prompts under 200K tokens. At or above 200K every rate doubles — $4.00 / $1.00 / $12.00 — and the higher rate applies to the whole request, not just the tokens past the line.

Input - per 1M tokens
$2.00/M
Under 200K prompt live
Output - per 1M tokens
$6.00/M
Reasoning billed as output live
Cached input
$0.50/M
Prompt cache SpaceXAI row -75%
Effective - agentic blend
$1.19/M
92/8 split - 82% cache
§ 01 / TERMINAL

Run the numbers.

Live calculator pre-loaded with Grok 4.6's sub-200K rates from SpaceXAI's pricing page. Reasoning effort is configurable across low, medium, high (default) and xhigh, and reasoning tokens bill at the output rate. Note the threshold: send a prompt of 200K tokens or more and every rate in this calculator doubles.

$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
Words equivalent (English)
Effective rate
Open full calculator (all models · share URL · CSV) →
§ 02 / SCENARIOS

Real-world presets.

§ 03 / TOKENIZER

Paste text. See tokens. See cost.

Estimate · grok-tokenizer-estimate · ≈3.85 chars/token Auto-counts as you type

This is a chars-per-token approximation, not a real tokenizer. Actual tokens vary by language, code density, and tool-call overhead — counts are typically ±10–20% off for English prose, more for code or non-Latin scripts. For exact billing, use the vendor's official tokenizer.

Characters 524
Words 85
Tokens (estimated) 136 tokens
Cost as input · uncached $0.00027 USD
Cost as output · uncached $0.00082 USD
Cost as cached input $0.00007 USD
§ 04 / SHELF

Up against the shelf.

All models →
Model Input /M Output /M Effective blended Context Best for
Grok 4.6 Current $2.00 cache $0.50 $6.00 $1.19 newest Grok 500K Coding, agents, and knowledge work
Grok 4.5 $2.00 cache $0.30 $6.00 $1.04 same rates, cheaper cache 500K Cache-heavy loops on the prior Grok
Grok 4.3 $1.25 cache $0.20 $2.50 $0.558 cheaper, 1M context 1M Lower-cost Grok chat and tools
Grok Build 0.1 $1.00 cache $0.20 $2.00 $0.476 coding early access 256K Dedicated coding workflows
Grok 4.20 Multi-Agent $1.25 cache $0.20 $2.50 $0.558 multi-agent scaffolding 1M Pinned long-horizon agent loops
Claude Sonnet 5 $2.00 cache $0.20 $10.00 $1.36 Anthropic coding 1M Claude-class coding agents
GPT-5.6 Sol $4.00 cache $0.40 $20.00 $2.73 OpenAI flagship 1.05M OpenAI frontier workloads

Frequently asked.

Grok 4.6 pricing questions, with SpaceXAI's token rates separated from the long-context threshold and cache assumptions.

Q · 01 What is Grok 4.6's API price? +
SpaceXAI lists grok-4.6 at $2.00/M input, $0.50/M cached input, and $6.00/M output for prompts below 200K tokens. These are USD prices per 1M tokens.
Q · 02 Why does Grok 4.6 have two sets of prices? +
SpaceXAI charges long-context requests at double the rate: $4.00/M input, $1.00/M cached input, and $12.00/M output. The trigger is the prompt reaching 200K tokens, and the higher rate then applies to every token in that request — not only the tokens beyond 200K. A 199K-token prompt and a 201K-token prompt therefore differ in cost by roughly 2x, not by 1%.
Q · 03 Is Grok 4.6 more expensive than Grok 4.5? +
On paper they are identical: both list $2/M input and $6/M output. The difference is cache. Grok 4.5's cached input was cut to $0.30/M on the day Grok 4.6 launched, while Grok 4.6 reads cache at $0.50/M. Under our standard blend (92% input, 82% cache hits) that puts Grok 4.6 at $1.19/M effective against Grok 4.5's $1.04/M — so on cache-heavy agent loops the older model is currently the cheaper one.
Q · 04 What context window does Grok 4.6 have? +
The SpaceXAI model page lists 500,000 tokens, the same window as Grok 4.5. Grok 4.3 and the Grok 4.20 variants are listed at 1M, so Grok 4.6 is not the longest-context Grok — it is the newest.
Q · 05 Does Grok 4.6 support batch pricing? +
Not that SpaceXAI publishes. The Grok 4.20 model pages carry a Show batch API pricing control; the Grok 4.6 page does not. We do not assume a batch discount that the vendor has not stated, so no batch rate is recorded for this model.
Q · 06 Where can I call Grok 4.6? +
The model page lists regions us-east-1 and us-west-2 only, with rate limits of 150 requests per second and 50M tokens per minute. It supports function calling, structured outputs, and reasoning at low, medium, high (default), and xhigh effort.