Grok 4.6 API Pricing
Grok 4.6 is SpaceXAI's newest frontier model for coding, agentic tasks, and knowledge work, released August 12, 2026. The pricing page lists $2.00/M input, $0.50/M cached input, and $6.00/M output for prompts under 200K tokens. At or above 200K every rate doubles — $4.00 / $1.00 / $12.00 — and the higher rate applies to the whole request, not just the tokens past the line.
Run the numbers.
Live calculator pre-loaded with Grok 4.6's sub-200K rates from SpaceXAI's pricing page. Reasoning effort is configurable across low, medium, high (default) and xhigh, and reasoning tokens bill at the output rate. Note the threshold: send a prompt of 200K tokens or more and every rate in this calculator doubles.
Real-world presets.
Tool-calling agent
Research pack synthesis
Power-user chat
Knowledge base RAG
Paste text. See tokens. See cost.
This is a chars-per-token approximation, not a real tokenizer. Actual tokens vary by language, code density, and tool-call overhead — counts are typically ±10–20% off for English prose, more for code or non-Latin scripts. For exact billing, use the vendor's official tokenizer.
| Model | Input /M | Output /M | Effective blended | Context | Best for |
|---|---|---|---|---|---|
| Grok 4.6 Current | $2.00 cache $0.50 | $6.00 | $1.19 newest Grok | 500K | Coding, agents, and knowledge work |
| Grok 4.5 | $2.00 cache $0.30 | $6.00 | $1.04 same rates, cheaper cache | 500K | Cache-heavy loops on the prior Grok |
| Grok 4.3 | $1.25 cache $0.20 | $2.50 | $0.558 cheaper, 1M context | 1M | Lower-cost Grok chat and tools |
| Grok Build 0.1 | $1.00 cache $0.20 | $2.00 | $0.476 coding early access | 256K | Dedicated coding workflows |
| Grok 4.20 Multi-Agent | $1.25 cache $0.20 | $2.50 | $0.558 multi-agent scaffolding | 1M | Pinned long-horizon agent loops |
| Claude Sonnet 5 | $2.00 cache $0.20 | $10.00 | $1.36 Anthropic coding | 1M | Claude-class coding agents |
| GPT-5.6 Sol | $4.00 cache $0.40 | $20.00 | $2.73 OpenAI flagship | 1.05M | OpenAI frontier workloads |
Frequently asked.
Grok 4.6 pricing questions, with SpaceXAI's token rates separated from the long-context threshold and cache assumptions.
Q · 01 What is Grok 4.6's API price? +
grok-4.6 at $2.00/M input, $0.50/M cached input, and $6.00/M output for prompts below 200K tokens. These are USD prices per 1M tokens.Q · 02 Why does Grok 4.6 have two sets of prices? +
$4.00/M input, $1.00/M cached input, and $12.00/M output. The trigger is the prompt reaching 200K tokens, and the higher rate then applies to every token in that request — not only the tokens beyond 200K. A 199K-token prompt and a 201K-token prompt therefore differ in cost by roughly 2x, not by 1%.Q · 03 Is Grok 4.6 more expensive than Grok 4.5? +
$2/M input and $6/M output. The difference is cache. Grok 4.5's cached input was cut to $0.30/M on the day Grok 4.6 launched, while Grok 4.6 reads cache at $0.50/M. Under our standard blend (92% input, 82% cache hits) that puts Grok 4.6 at $1.19/M effective against Grok 4.5's $1.04/M — so on cache-heavy agent loops the older model is currently the cheaper one.Q · 04 What context window does Grok 4.6 have? +
500,000 tokens, the same window as Grok 4.5. Grok 4.3 and the Grok 4.20 variants are listed at 1M, so Grok 4.6 is not the longest-context Grok — it is the newest.Q · 05 Does Grok 4.6 support batch pricing? +
Q · 06 Where can I call Grok 4.6? +
us-east-1 and us-west-2 only, with rate limits of 150 requests per second and 50M tokens per minute. It supports function calling, structured outputs, and reasoning at low, medium, high (default), and xhigh effort.