Grok 4.6

xAI · Active · available in 28 of 31 regions

Tool useVisionPrompt cachingStreaming Batch inferenceEmbeddingsFine-tuning
Context window
500K
500,000 tokens
Max output
Input / 1M
$2.00
us-east-1
Output / 1M
$6.00

What Grok 4.6 is good for

editorial

Grok 4.6 offers a 500K-token context window with tool use and prompt caching at $2.00 per 1M input tokens. Callable in 28 of 31 regions.

Suits

  • Long documents. 500K tokens covers long documents and multi-file context without splitting them.
  • Repeated prompts. Supports prompt caching, so a re-sent prefix costs about 4× less than fresh input. On agentic loops that resend a long system prompt every step, this moves the bill more than the headline price does.
  • Tool use. Can call functions, so it can drive retrieval, look things up, and take actions rather than only answering.
  • Images. Accepts image input: screenshots, scanned documents, charts and UI captures.

Think twice if

  • Invocation. Reachable only through a cross-region inference profile — calling the bare model ID returns a validation error. Use an ID such as global.xai.grok-4.6.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Agentic workflows

Multi-step work where the model calls tools, reads the results and decides what to do next.

  • · A support agent that looks up the order, checks the refund policy and drafts the reply before a human approves it.
  • · An on-call assistant that reads logs and metrics, forms a hypothesis, and proposes the one command to run.
  • · A migration bot that opens a pull request per file and reruns the test suite between each change.

Coding

Writing, reviewing and refactoring code, usually across more than one file.

  • · A review pass that flags real defects with a severity and leaves style alone.
  • · A framework upgrade applied across a repository, one module at a time, tests green between each.
  • · Turning a failing bug report into a reproducing test and then a fix.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Cross-region inference profiles (2)

Regions marked Inference profile only require a profile ID rather than the base model ID — invoking the base ID there returns a validation error. Regional (us., eu.) profiles carry a 10% premium over global.

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $2.00 $6.00 $0.500 curated
us-east-2 US East (Ohio) $2.00 $6.00 $0.500 curated
us-west-1 US West (N. California) $2.00 $6.00 $0.500 curated
us-west-2 US West (Oregon) $2.00 $6.00 $0.500 curated
ca-central-1 Canada (Central) $2.00 $6.00 $0.500 curated
ca-west-1 Canada West (Calgary) $2.00 $6.00 $0.500 curated
eu-west-1 EU (Ireland) $2.00 $6.00 $0.500 curated
eu-west-2 EU (London) $2.00 $6.00 $0.500 curated
eu-west-3 EU (Paris) $2.00 $6.00 $0.500 curated
eu-central-1 EU (Frankfurt) $2.00 $6.00 $0.500 curated
eu-central-2 Europe (Zurich) $2.00 $6.00 $0.500 curated
eu-north-1 EU (Stockholm) $2.00 $6.00 $0.500 curated
eu-south-1 EU (Milan) $2.00 $6.00 $0.500 curated
eu-south-2 Europe (Spain) $2.00 $6.00 $0.500 curated
ap-east-2 Asia Pacific (Taipei) $2.00 $6.00 $0.500 curated
ap-northeast-1 Asia Pacific (Tokyo) $2.00 $6.00 $0.500 curated
ap-northeast-2 Asia Pacific (Seoul) $2.00 $6.00 $0.500 curated
ap-northeast-3 Asia Pacific (Osaka) $2.00 $6.00 $0.500 curated
ap-south-1 Asia Pacific (Mumbai) $2.00 $6.00 $0.500 curated
ap-south-2 Asia Pacific (Hyderabad) $2.00 $6.00 $0.500 curated
ap-southeast-1 Asia Pacific (Singapore) $2.00 $6.00 $0.500 curated
ap-southeast-2 Asia Pacific (Sydney) $2.00 $6.00 $0.500 curated
ap-southeast-3 Asia Pacific (Jakarta) $2.00 $6.00 $0.500 curated
ap-southeast-4 Asia Pacific (Melbourne) $2.00 $6.00 $0.500 curated
ap-southeast-5 Asia Pacific (Malaysia) $2.00 $6.00 $0.500 curated
ap-southeast-7 Asia Pacific (Thailand) $2.00 $6.00 $0.500 curated
il-central-1 Israel (Tel Aviv) $2.00 $6.00 $0.500 curated
sa-east-1 South America (Sao Paulo) $2.00 $6.00 $0.500 curated

Curated pricing. The AWS Price List API publishes no SKU for this model, so these figures are hand-maintained from AWS's public pricing page and verified on Aug 27, 2026. They are not machine-derived — confirm in the AWS console before committing spend. Global endpoint. In-Region and US-geo profiles bill 10% more ($2.20 / $6.60).

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Grok 4.6 can be served without the request leaving 1 of the 25 jurisdictions with a Bedrock region.

Grok 4.6 compared

Side-by-side pages against every model people weigh this one against.

Grok 4.6 vs Claude Fable 5Grok 4.6 costs 5.0× less per input token than Claude Fable 5, and Claude Fable 5 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Haiku 4.5Claude Haiku 4.5 costs 2.0× less per input token than Grok 4.6, and Grok 4.6 holds 500K tokens against 200K for Claude Haiku 4.5.Grok 4.6 vs Claude Opus 4.5Grok 4.6 costs 2.5× less per input token than Claude Opus 4.5, and Grok 4.6 holds 500K tokens against 200K for Claude Opus 4.5.Grok 4.6 vs Claude Opus 4.6Grok 4.6 costs 2.5× less per input token than Claude Opus 4.6, and Claude Opus 4.6 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Opus 4.7Grok 4.6 costs 2.5× less per input token than Claude Opus 4.7, and Claude Opus 4.7 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Opus 4.8Grok 4.6 costs 2.5× less per input token than Claude Opus 4.8, and Claude Opus 4.8 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Opus 5Grok 4.6 costs 2.5× less per input token than Claude Opus 5, and Claude Opus 5 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Sonnet 4.5Grok 4.6 costs 1.5× less per input token than Claude Sonnet 4.5, and Grok 4.6 holds 500K tokens against 200K for Claude Sonnet 4.5.Grok 4.6 vs Claude Sonnet 4.6Grok 4.6 costs 1.5× less per input token than Claude Sonnet 4.6, and Claude Sonnet 4.6 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Claude Sonnet 5Grok 4.6 costs 1.5× less per input token than Claude Sonnet 5, and Claude Sonnet 5 holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs GPT-5.6 LunaGPT-5.6 Luna costs 10× less per input token than Grok 4.6, and GPT-5.6 Luna holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs GPT-5.6 SolGrok 4.6 costs 2.0× less per input token than GPT-5.6 Sol, and GPT-5.6 Sol holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs GPT-5.6 TerraSame input price, and GPT-5.6 Terra holds 1M tokens against 500K for Grok 4.6.Grok 4.6 vs Nova 2 LiteNova 2 Lite costs 6.7× less per input token than Grok 4.6, and only Grok 4.6 does tool use (Nova 2 Lite does not).Grok 4.6 vs Nova LiteNova Lite costs 33× less per input token than Grok 4.6, and Grok 4.6 holds 500K tokens against 300K for Nova Lite.

Grok 4.6 appears in

Common questions

How much does Grok 4.6 cost on Amazon Bedrock?
$2.00 per 1M input tokens and $6.00 per 1M output tokens in us-east-1. Cached input reads at $0.500, roughly a tenth of the input rate, which dominates the bill on agentic workloads that re-send a long prefix.
Which AWS regions support Grok 4.6?
28 of 31 regions: us-east-1, us-east-2, us-west-1, us-west-2, ca-central-1, ca-west-1, eu-west-1, eu-west-2, eu-west-3, eu-central-1, eu-central-2, eu-north-1, eu-south-1, eu-south-2, ap-east-2, ap-northeast-1, ap-northeast-2, ap-northeast-3, ap-south-1, ap-south-2, ap-southeast-1, ap-southeast-2, ap-southeast-3, ap-southeast-4, ap-southeast-5, ap-southeast-7, il-central-1, sa-east-1. Regions marked profile-only need a cross-region inference profile ID such as global.xai.grok-4.6 rather than the bare model ID.
What is the context window of Grok 4.6?
500,000 tokens (500K).
What is the model ID for Grok 4.6 on Bedrock?
xai.grok-4.6. In regions where it is only reachable through a cross-region inference profile, use a prefixed ID instead: global.xai.grok-4.6, us.xai.grok-4.6.