Qwen3 VL 235B A22B

Qwen · Active · available in 10 of 31 regions

VisionBatch inferenceStreaming Tool usePrompt cachingEmbeddingsFine-tuning
Context window
Max output
Input / 1M
$0.530
us-east-1
Output / 1M
$2.66

What Qwen3 VL 235B A22B is good for

editorial

Qwen3 VL 235B A22B offers image input, at the premium end of this provider’s range at $0.530 per 1M input tokens. Callable in 10 of 31 regions.

Suits

  • Images. Accepts image input: screenshots, scanned documents, charts and UI captures.
  • Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.

Think twice if

  • Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
  • Function calling. No tool use, so it cannot drive an agent loop or call your APIs — it answers from what is in the prompt.
  • High volume. At $0.530 per 1M input tokens it sits at the expensive end of the Qwen range. Worth it for hard tasks, wasteful for bulk extraction.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Chat and assistants

Conversational work where responsiveness matters as much as depth.

  • · A customer-facing assistant where the first token needs to appear in under a second.
  • · An in-product copilot that explains what the user is looking at and answers follow-ups.
  • · A triage bot that qualifies an incoming request before routing it to the right team.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $0.530 $2.66 $0.260 $1.33 API
us-east-2 US East (Ohio) $0.530 $2.66 $0.260 $1.33 API
us-west-2 US West (Oregon) $0.530 $2.66 $0.260 $1.33 API
eu-west-1 EU (Ireland) $0.620 $3.12 $0.310 $1.56 API
eu-west-2 EU (London) $0.820 $4.12 $0.410 $2.06 API
eu-south-1 EU (Milan) $0.620 $3.12 $0.310 $1.56 API
ap-northeast-1 Asia Pacific (Tokyo) $0.640 $3.22 $0.320 $1.61 API
ap-south-1 Asia Pacific (Mumbai) $0.620 $3.13 $0.310 $1.56 API
ap-southeast-2 Asia Pacific (Sydney) $0.546 $2.74 $0.273 $1.37 API
sa-east-1 South America (Sao Paulo) $0.640 $3.22 $0.320 $1.61 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Qwen3 VL 235B A22B can be served without the request leaving 9 of the 25 jurisdictions with a Bedrock region.

Qwen3 VL 235B A22B appears in

Common questions

How much does Qwen3 VL 235B A22B cost on Amazon Bedrock?
$0.530 per 1M input tokens and $2.66 per 1M output tokens in us-east-1.
Which AWS regions support Qwen3 VL 235B A22B?
10 of 31 regions: us-east-1, us-east-2, us-west-2, eu-west-1, eu-west-2, eu-south-1, ap-northeast-1, ap-south-1, ap-southeast-2, sa-east-1.
What is the model ID for Qwen3 VL 235B A22B on Bedrock?
qwen.qwen3-vl-235b-a22b.

Other Qwen models