GPT-5.6 Terra
OpenAI · Active · available in 28 of 31 regions
- Context window
- 1M
- 1,000,000 tokens
- Max output
- —
- Input / 1M
- $2.00
- us-east-1
- Output / 1M
- $12.00
What GPT-5.6 Terra is good for
editorialGPT-5.6 Terra offers a 1M-token context window with tool use and prompt caching at $2.00 per 1M input tokens. Callable in 28 of 31 regions.
Suits
- Very long inputs. A 1M-token window holds an entire codebase or a day of transcripts in one call, so you can skip chunking.
- Repeated prompts. Supports prompt caching, so a re-sent prefix costs about 10× less than fresh input. On agentic loops that resend a long system prompt every step, this moves the bill more than the headline price does.
- Tool use. Can call functions, so it can drive retrieval, look things up, and take actions rather than only answering.
- Images. Accepts image input: screenshots, scanned documents, charts and UI captures.
Think twice if
- Invocation. Reachable only through a cross-region inference profile — calling the bare model ID returns a validation error. Use an ID such as global.openai.gpt-5.6-terra.
This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.
Where teams typically use it
Coding
Writing, reviewing and refactoring code, usually across more than one file.
- · A review pass that flags real defects with a severity and leaves style alone.
- · A framework upgrade applied across a repository, one module at a time, tests green between each.
- · Turning a failing bug report into a reproducing test and then a fix.
Agentic workflows
Multi-step work where the model calls tools, reads the results and decides what to do next.
- · A support agent that looks up the order, checks the refund policy and drafts the reply before a human approves it.
- · An on-call assistant that reads logs and metrics, forms a hypothesis, and proposes the one command to run.
- · A migration bot that opens a pull request per file and reruns the test suite between each change.
Summarisation and extraction
High-volume, structured output from unstructured input. Usually the cheapest workload to run well.
- · Turning a million support tickets into a fixed schema of product, severity and root cause.
- · Pulling line items, totals and dates off scanned invoices into JSON.
- · Condensing every meeting transcript into decisions taken and owners assigned.
Model & inference profile IDs
Base model ID — direct on-demand invoke
Cross-region inference profiles (3)
Regions marked Inference profile only require a profile ID rather than the base model ID — invoking the base ID there returns a validation error. Regional (us., eu.) profiles carry a 10% premium over global.
Pricing by region
USD per 1M tokens| Region | Input | Output | Cache read | Cache write | Batch in | Batch out | Source |
|---|---|---|---|---|---|---|---|
| us-east-1 US East (N. Virginia) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| us-east-2 US East (Ohio) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| us-west-1 US West (N. California) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| us-west-2 US West (Oregon) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ca-central-1 Canada (Central) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ca-west-1 Canada West (Calgary) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-west-1 EU (Ireland) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-west-2 EU (London) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-west-3 EU (Paris) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-central-1 EU (Frankfurt) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-central-2 Europe (Zurich) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-north-1 EU (Stockholm) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-south-1 EU (Milan) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| eu-south-2 Europe (Spain) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-east-2 Asia Pacific (Taipei) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-northeast-1 Asia Pacific (Tokyo) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-northeast-2 Asia Pacific (Seoul) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-northeast-3 Asia Pacific (Osaka) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-south-1 Asia Pacific (Mumbai) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-south-2 Asia Pacific (Hyderabad) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-1 Asia Pacific (Singapore) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-2 Asia Pacific (Sydney) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-3 Asia Pacific (Jakarta) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-4 Asia Pacific (Melbourne) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-5 Asia Pacific (Malaysia) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| ap-southeast-7 Asia Pacific (Thailand) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| il-central-1 Israel (Tel Aviv) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
| sa-east-1 South America (Sao Paulo) | $2.00 | $12.00 | $0.200 | $2.50 | — | — | curated |
Curated pricing. The AWS Price List API publishes no SKU for this model, so these figures are hand-maintained from AWS's public pricing page and verified on Aug 27, 2026. They are not machine-derived — confirm in the AWS console before committing spend. Global endpoint, requests up to 272K tokens. Beyond that AWS bills 2x input and 1.5x output.
Region availability
TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.
Where this model actually runs
Under a strict residency constraint, GPT-5.6 Terra can be served without the request leaving 2 of the 25 jurisdictions with a Bedrock region.
GPT-5.6 Terra compared
Side-by-side pages against every model people weigh this one against.
GPT-5.6 Terra appears in
Common questions
- How much does GPT-5.6 Terra cost on Amazon Bedrock?
- $2.00 per 1M input tokens and $12.00 per 1M output tokens in us-east-1. Cached input reads at $0.200, roughly a tenth of the input rate, which dominates the bill on agentic workloads that re-send a long prefix.
- Which AWS regions support GPT-5.6 Terra?
- 28 of 31 regions: us-east-1, us-east-2, us-west-1, us-west-2, ca-central-1, ca-west-1, eu-west-1, eu-west-2, eu-west-3, eu-central-1, eu-central-2, eu-north-1, eu-south-1, eu-south-2, ap-east-2, ap-northeast-1, ap-northeast-2, ap-northeast-3, ap-south-1, ap-south-2, ap-southeast-1, ap-southeast-2, ap-southeast-3, ap-southeast-4, ap-southeast-5, ap-southeast-7, il-central-1, sa-east-1. Regions marked profile-only need a cross-region inference profile ID such as global.openai.gpt-5.6-terra rather than the bare model ID.
- What is the context window of GPT-5.6 Terra?
- 1,000,000 tokens (1M).
- What is the model ID for GPT-5.6 Terra on Bedrock?
- openai.gpt-5.6-terra. In regions where it is only reachable through a cross-region inference profile, use a prefixed ID instead: global.openai.gpt-5.6-terra, in.openai.gpt-5.6-terra, us.openai.gpt-5.6-terra.