Grok 4.6
The latest xAI reasoning model. Successor to Grok 4.5.
Grok 4.6 is xAI’s latest reasoning model, released August 12, 2026 as the successor to Grok 4.5. It carries the same 500k context window and two-tier pricing structure while delivering incremental improvements in reasoning accuracy and instruction following.
Like Grok 4.5, it is positioned as a budget alternative to Claude and GPT for long-document analysis and codebase-scale reasoning. The pricing remains $2/M input and $6/M output for prompts under 200k tokens, doubling to $4/M and $12/M for longer prompts.
We sell it at half xAI’s list price — $1.00 per million input tokens against $2.00, regardless of prompt length. That makes long-context work on Grok cheaper here than buying directly from xAI, just as it is on Grok 4.5.
What it costs you per month
Real-world budget scenarios. Numbers are simple sums — official list price vs llmrelay's 50% off tier.
What it's good at
- +500K context window
- +Improved reasoning from 4.5
- +Reasoning traces
Best for
- — Long-document analysis
- — Whole-repo review
- — Budget reasoning
Pick Grok 4.6 when
- +You are already using Grok 4.5 and want the latest iteration. The version jump suggests measurable improvements, though xAI has not published specific benchmark deltas.
- +Long-document reasoning on a budget. At $1.00/M input Grok 4.6 costs the same as Claude Sonnet 5 and half what GPT-5.6 Terra charges, while still carrying a 500k window.
- +You value reasoning transparency. Grok’s step-by-step trace mode remains available and is useful for understanding how the model reached a conclusion.
- +Whole-repo code review where 500k context is enough. That window fits most repositories, and the per-token savings over Claude or GPT add up fast on large codebases.
Choose something else when
- !You need more than 500k context. Claude Opus, Claude Sonnet, GPT-5.6 Sol and GPT-5.6 Terra all carry 1M windows, which is twice what Grok offers.
- !Tool use and function calling are load-bearing. Grok supports it, but Claude’s and OpenAI’s implementations are more mature and more widely integrated.
- !You are locked into a specific ecosystem. If your workflow is built around Claude Code or OpenAI function calling, switching to Grok for incremental improvements is not worth the friction.
- !Grok 4.5 is working fine for you. The 4.5 → 4.6 jump is incremental, not generational. Unless you hit specific shortcomings in 4.5, there is no urgent reason to migrate.
Questions people ask about Grok 4.6
How much does Grok 4.6 cost?
xAI’s list price has two tiers: under 200k input tokens it is $2.00 per million input and $6.00 per million output; at or above 200k input it is $4.00 and $12.00. On llmrelay we charge the lower tier regardless of prompt length — $1.00 and $3.00 — which makes long-context work on Grok cheaper here than buying directly from xAI.
What changed from Grok 4.5 to 4.6?
xAI has not published detailed benchmarks, but version increments in this family historically mean improvements in reasoning accuracy, instruction following, and edge-case handling. The context window, pricing structure, and API shape remain unchanged.
Should I use Grok 4.6 or stick with Grok 4.5?
If 4.5 is meeting your needs, there is no urgent reason to switch — the improvement is incremental. If you are starting fresh or hitting specific shortcomings in 4.5, 4.6 is the natural pick.
What is the context window on Grok 4.6?
500,000 tokens input, with up to 32,768 output tokens. That is the same as Grok 4.5 and half the window of Claude Opus, Claude Sonnet, or GPT-5.6 models (all 1M).
Try Grok 4.6 at half the price
Free to create an account, no subscription, $10 minimum top-up — enough to run this model against a real task and judge quality yourself.
Get API key →Compare Grok 4.6 against the alternatives
The comparisons this model appears in, the models nearest it on price, and the full rate card.
- GPT-5.6 Luna pricing and specs$0.50/M in, $3.00/M out — half list. The small GPT-5.6. Cheap enough for high-volume routing and classification.
- Claude Haiku 4.5 pricing and specs$0.50/M in, $2.50/M out — half list. The fastest and cheapest Claude. Perfect for classification and routing.
- Full price listEvery model we serve, at 50% of official list. No subscription, no volume gate.
- Cost calculatorPut your own monthly token volume in and see the bill against vendor list price.
- Tool setup guidesBase-URL and key steps for Cursor, Cline, Claude Code, Continue and others.