Claude Sonnet 5

Anthropic · Active · available in 31 of 31 regions

Tool useVisionPrompt cachingBatch inferenceStreaming EmbeddingsFine-tuning
Context window
1M
1,000,000 tokens
Max output
128K
Input / 1M
$2.00
us-east-1
Output / 1M
$10.00

What Claude Sonnet 5 is good for

editorial

Claude Sonnet 5 offers a 1M-token context window with tool use and prompt caching at $2.00 per 1M input tokens. Callable in 31 of 31 regions.

Suits

  • Very long inputs. A 1M-token window holds an entire codebase or a day of transcripts in one call, so you can skip chunking.
  • Repeated prompts. Supports prompt caching, so a re-sent prefix costs about 10× less than fresh input. On agentic loops that resend a long system prompt every step, this moves the bill more than the headline price does.
  • Tool use. Can call functions, so it can drive retrieval, look things up, and take actions rather than only answering.
  • Images. Accepts image input: screenshots, scanned documents, charts and UI captures.
  • Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.

Think twice if

  • Invocation. Reachable only through a cross-region inference profile — calling the bare model ID returns a validation error. Use an ID such as au.anthropic.claude-sonnet-5.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Agentic workflows

Multi-step work where the model calls tools, reads the results and decides what to do next.

  • · A support agent that looks up the order, checks the refund policy and drafts the reply before a human approves it.
  • · An on-call assistant that reads logs and metrics, forms a hypothesis, and proposes the one command to run.
  • · A migration bot that opens a pull request per file and reruns the test suite between each change.

Coding

Writing, reviewing and refactoring code, usually across more than one file.

  • · A review pass that flags real defects with a severity and leaves style alone.
  • · A framework upgrade applied across a repository, one module at a time, tests green between each.
  • · Turning a failing bug report into a reproducing test and then a fix.

Chat and assistants

Conversational work where responsiveness matters as much as depth.

  • · A customer-facing assistant where the first token needs to appear in under a second.
  • · An in-product copilot that explains what the user is looking at and answers follow-ups.
  • · A triage bot that qualifies an incoming request before routing it to the right team.

RAG and retrieval

Answering from your own documents, with the retrieved passages in the prompt.

  • · An internal search over contracts and policies that quotes the clause it answered from.
  • · A product assistant grounded in the current documentation rather than the training data.
  • · A research summariser that reads twenty retrieved papers and reconciles where they disagree.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Cross-region inference profiles (4)

Regions marked Inference profile only require a profile ID rather than the base model ID — invoking the base ID there returns a validation error. Regional (us., eu.) profiles carry a 10% premium over global.

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
us-east-2 US East (Ohio) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
us-west-1 US West (N. California) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
us-west-2 US West (Oregon) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ca-central-1 Canada (Central) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ca-west-1 Canada West (Calgary) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
mx-central-1 Mexico (Central) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-west-1 EU (Ireland) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-west-2 EU (London) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-west-3 EU (Paris) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-central-1 EU (Frankfurt) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-central-2 Europe (Zurich) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-north-1 EU (Stockholm) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-south-1 EU (Milan) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
eu-south-2 Europe (Spain) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-east-2 Asia Pacific (Taipei) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-northeast-1 Asia Pacific (Tokyo) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-northeast-2 Asia Pacific (Seoul) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-northeast-3 Asia Pacific (Osaka) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-south-1 Asia Pacific (Mumbai) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-south-2 Asia Pacific (Hyderabad) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-1 Asia Pacific (Singapore) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-2 Asia Pacific (Sydney) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-3 Asia Pacific (Jakarta) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-4 Asia Pacific (Melbourne) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-5 Asia Pacific (Malaysia) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-6 Asia Pacific (New Zealand) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
ap-southeast-7 Asia Pacific (Thailand) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
il-central-1 Israel (Tel Aviv) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
af-south-1 Africa (Cape Town) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated
sa-east-1 South America (Sao Paulo) $2.00 $10.00 $0.200 $2.50 $1.00 $5.00 curated

Curated pricing. The AWS Price List API publishes no SKU for this model, so these figures are hand-maintained from AWS's public pricing page and verified on Sep 15, 2026. They are not machine-derived — confirm in the AWS console before committing spend.

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Claude Sonnet 5 can be served without the request leaving 3 of the 25 jurisdictions with a Bedrock region.

Claude Sonnet 5 compared

Side-by-side pages against every model people weigh this one against.

Claude Sonnet 5 vs Claude Fable 5Claude Sonnet 5 costs 5.0× less per input token than Claude Fable 5.Claude Sonnet 5 vs Claude Fable 5.1Claude Sonnet 5 costs 5.0× less per input token than Claude Fable 5.1.Claude Sonnet 5 vs Claude Haiku 4.5Claude Haiku 4.5 costs 2.0× less per input token than Claude Sonnet 5, and Claude Sonnet 5 holds 1M tokens against 200K for Claude Haiku 4.5.Claude Sonnet 5 vs Claude Opus 4.5Claude Sonnet 5 costs 2.5× less per input token than Claude Opus 4.5, and Claude Sonnet 5 holds 1M tokens against 200K for Claude Opus 4.5.Claude Sonnet 5 vs Claude Opus 4.6Claude Sonnet 5 costs 2.5× less per input token than Claude Opus 4.6.Claude Sonnet 5 vs Claude Opus 4.7Claude Sonnet 5 costs 2.5× less per input token than Claude Opus 4.7.Claude Sonnet 5 vs Claude Opus 4.8Claude Sonnet 5 costs 2.5× less per input token than Claude Opus 4.8.Claude Sonnet 5 vs Claude Opus 5Claude Sonnet 5 costs 2.5× less per input token than Claude Opus 5.Claude Sonnet 5 vs Claude Sonnet 4.5Claude Sonnet 5 costs 1.5× less per input token than Claude Sonnet 4.5, and Claude Sonnet 5 holds 1M tokens against 200K for Claude Sonnet 4.5.Claude Sonnet 5 vs Claude Sonnet 4.6Claude Sonnet 5 costs 1.5× less per input token than Claude Sonnet 4.6.Claude Sonnet 5 vs GPT-6 AstraClaude Sonnet 5 costs 5.0× less per input token than GPT-6 Astra.Claude Sonnet 5 vs GPT-5.6 LunaGPT-5.6 Luna costs 10× less per input token than Claude Sonnet 5.Claude Sonnet 5 vs GPT-5.6 SolClaude Sonnet 5 costs 2.0× less per input token than GPT-5.6 Sol.Claude Sonnet 5 vs GPT-5.6 TerraSame input price.Claude Sonnet 5 vs Grok 4.6Same input price, and Claude Sonnet 5 holds 1M tokens against 500K for Grok 4.6.

Claude Sonnet 5 appears in

Common questions

How much does Claude Sonnet 5 cost on Amazon Bedrock?
$2.00 per 1M input tokens and $10.00 per 1M output tokens in us-east-1. Cached input reads at $0.200, roughly a tenth of the input rate, which dominates the bill on agentic workloads that re-send a long prefix.
Which AWS regions support Claude Sonnet 5?
31 of 31 regions: us-east-1, us-east-2, us-west-1, us-west-2, ca-central-1, ca-west-1, mx-central-1, eu-west-1, eu-west-2, eu-west-3, eu-central-1, eu-central-2, eu-north-1, eu-south-1, eu-south-2, ap-east-2, ap-northeast-1, ap-northeast-2, ap-northeast-3, ap-south-1, ap-south-2, ap-southeast-1, ap-southeast-2, ap-southeast-3, ap-southeast-4, ap-southeast-5, ap-southeast-6, ap-southeast-7, il-central-1, af-south-1, sa-east-1. Regions marked profile-only need a cross-region inference profile ID such as au.anthropic.claude-sonnet-5 rather than the bare model ID.
What is the context window of Claude Sonnet 5?
1,000,000 tokens (1M), with up to 128K output tokens per request.
What is the model ID for Claude Sonnet 5 on Bedrock?
anthropic.claude-sonnet-5. In regions where it is only reachable through a cross-region inference profile, use a prefixed ID instead: au.anthropic.claude-sonnet-5, eu.anthropic.claude-sonnet-5, global.anthropic.claude-sonnet-5.

Other Anthropic models