Qwen3 Coder 480B A35B Instruct

Qwen · Active · available in 8 of 31 regions

Batch inferenceStreaming Tool useVisionPrompt cachingEmbeddingsFine-tuning
Context window
Max output
Input / 1M
$0.450
us-west-2
Output / 1M
$1.80

What Qwen3 Coder 480B A35B Instruct is good for

editorial

Qwen3 Coder 480B A35B Instruct at $0.450 per 1M input tokens. Callable in 8 of 31 regions.

Suits

  • Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.

Think twice if

  • Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
  • Function calling. No tool use, so it cannot drive an agent loop or call your APIs — it answers from what is in the prompt.
  • Documents with layout. Text only. Scanned PDFs, screenshots and charts need a vision model or an OCR step first.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Coding

Writing, reviewing and refactoring code, usually across more than one file.

  • · A review pass that flags real defects with a severity and leaves style alone.
  • · A framework upgrade applied across a repository, one module at a time, tests green between each.
  • · Turning a failing bug report into a reproducing test and then a fix.

Agentic workflows

Multi-step work where the model calls tools, reads the results and decides what to do next.

  • · A support agent that looks up the order, checks the refund policy and drafts the reply before a human approves it.
  • · An on-call assistant that reads logs and metrics, forms a hypothesis, and proposes the one command to run.
  • · A migration bot that opens a pull request per file and reruns the test suite between each change.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-2 US East (Ohio) $0.450 $1.80 $0.225 $0.900 API
us-west-2 US West (Oregon) $0.450 $1.80 $0.225 $0.900 API
eu-west-2 EU (London) $0.700 $2.79 $0.350 $1.40 API
eu-north-1 EU (Stockholm) $0.450 $1.80 $0.225 $0.900 API
ap-northeast-1 Asia Pacific (Tokyo) $0.540 $2.18 $0.270 $1.09 API
ap-south-1 Asia Pacific (Mumbai) $0.530 $2.12 $0.265 $1.06 API
ap-southeast-2 Asia Pacific (Sydney) $0.463 $1.85 $0.232 $0.927 API
ap-southeast-3 Asia Pacific (Jakarta) $0.470 $1.87 $0.230 $0.930 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Qwen3 Coder 480B A35B Instruct can be served without the request leaving 8 of the 25 jurisdictions with a Bedrock region.

Qwen3 Coder 480B A35B Instruct appears in

Common questions

How much does Qwen3 Coder 480B A35B Instruct cost on Amazon Bedrock?
$0.450 per 1M input tokens and $1.80 per 1M output tokens in us-west-2. The cheapest region is us-east-2 at $0.450.
Which AWS regions support Qwen3 Coder 480B A35B Instruct?
8 of 31 regions: us-east-2, us-west-2, eu-west-2, eu-north-1, ap-northeast-1, ap-south-1, ap-southeast-2, ap-southeast-3.
What is the model ID for Qwen3 Coder 480B A35B Instruct on Bedrock?
qwen.qwen3-coder-480b-a35b-v1:0.

Other Qwen models