← All AI Ask BlogTest models →
Model releases · 7 min read

Grok 4.6: API Pricing, 500K Context, and What Changed

Grok 4.6 is xAI’s new 500K-context flagship for coding, agents, and knowledge work, priced at $2/M input and $6/M output with cached input listed separately. Consider it for tool-using workflows, but test the >200K context pricing boundary, regional availability, and agent behavior before switching from 4.5.

By Todd · Published 2026-08-14 · Verified 2026-08-14

Endpointgrok-4.6
Context500,000 tokens
Input$2/M; $0.50/M cached input
Output$6/M
ToolsFunction calling and structured output
Verified2026-08-14

What changed from Grok 4.5

Grok 4.6 keeps the same broad developer shape—text and image input, configurable reasoning, function calling, and structured output—while positioning the endpoint as xAI’s current flagship for coding and agents. The 500K context window is large enough for many repositories and document workflows without claiming a universal long-context advantage.

Price and cached input

The published standard rates are $2 per million input tokens, $0.50 per million cached input tokens, and $6 per million output tokens. The >200K-context pricing treatment is a separate verification point: keep it visible as a caveat until the pricing page clearly defines how that threshold is billed.

Who should use it—and who should wait

Use Grok 4.6 when configurable reasoning and tool use matter and your team can run a provider-specific regression suite. Do not switch blind if your traffic routinely crosses 200K input tokens, depends on a particular region, or assumes 4.5’s exact output style.

Migration checklist and unknowns

Pin grok-4.6, replay tool-call fixtures, compare cached and uncached traffic separately, and record the context size for every cost sample. All AI Ask has not published a controlled first-party quality or latency verdict for this release; provider positioning is labeled as such.

Explore the canonical data

Grok 4.6 model pageGrok 4.6 pricingGrok 4.5 → 4.6 comparisonxAI provider profileGrok 4.6 vs Claude Sonnet 5

Test it yourself

Run the same prompt across grok 4.6, claude sonnet 5, gemini 3.7 flash in one workspace.

Open the comparison playground →

FAQ

How much does Grok 4.6 cost?

The published standard rates are $2/M input, $0.50/M cached input, and $6/M output. Check the pricing page for any context-threshold terms.

Does Grok 4.6 support function calling?

Yes, the model documentation lists function calling and structured output.

Sources