Llama 4 Maverick 17B Instruct
Meta · Active · available in 4 of 31 regions
- Context window
- —
- Max output
- —
- Input / 1M
- $0.240
- us-east-1
- Output / 1M
- $0.970
What Llama 4 Maverick 17B Instruct is good for
editorialLlama 4 Maverick 17B Instruct offers image input at $0.240 per 1M input tokens. Callable in 4 of 31 regions.
Suits
- Images. Accepts image input: screenshots, scanned documents, charts and UI captures.
- Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.
Think twice if
- Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
- Function calling. No tool use, so it cannot drive an agent loop or call your APIs — it answers from what is in the prompt.
- Invocation. Reachable only through a cross-region inference profile — calling the bare model ID returns a validation error. Use an ID such as us.meta.llama4-maverick-17b-instruct-v1:0.
This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.
Model & inference profile IDs
Base model ID — direct on-demand invoke
Cross-region inference profiles (1)
Regions marked Inference profile only require a profile ID rather than the base model ID — invoking the base ID there returns a validation error. Regional (us., eu.) profiles carry a 10% premium over global.
Pricing by region
USD per 1M tokens| Region | Input | Output | Cache read | Cache write | Batch in | Batch out | Source |
|---|---|---|---|---|---|---|---|
| us-east-1 US East (N. Virginia) | $0.240 | $0.970 | — | — | $0.120 | $0.485 | API |
| us-east-2 US East (Ohio) | $0.240 | $0.970 | — | — | $0.120 | $0.485 | API |
| us-west-1 US West (N. California) | $0.240 | $0.970 | — | — | $0.120 | $0.485 | API |
| us-west-2 US West (Oregon) | $0.240 | $0.970 | — | — | $0.120 | $0.485 | API |
Region availability
TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.
Where this model actually runs
Under a strict residency constraint, Llama 4 Maverick 17B Instruct can be served without the request leaving 1 of the 25 jurisdictions with a Bedrock region.
Llama 4 Maverick 17B Instruct appears in
Common questions
- How much does Llama 4 Maverick 17B Instruct cost on Amazon Bedrock?
- $0.240 per 1M input tokens and $0.970 per 1M output tokens in us-east-1.
- Which AWS regions support Llama 4 Maverick 17B Instruct?
- 4 of 31 regions: us-east-1, us-east-2, us-west-1, us-west-2. Regions marked profile-only need a cross-region inference profile ID such as us.meta.llama4-maverick-17b-instruct-v1:0 rather than the bare model ID.
- What is the model ID for Llama 4 Maverick 17B Instruct on Bedrock?
- meta.llama4-maverick-17b-instruct-v1:0. In regions where it is only reachable through a cross-region inference profile, use a prefixed ID instead: us.meta.llama4-maverick-17b-instruct-v1:0.