Mistral 7B Instruct

Mistral AI · Active · available in 9 of 31 regions

Streaming Tool useVisionPrompt cachingBatch inferenceEmbeddingsFine-tuning
Context window
32K
32,000 tokens
Max output
Input / 1M
$0.150
us-east-1
Output / 1M
$0.200

What Mistral 7B Instruct is good for

editorial

Mistral 7B Instruct offers a 32K-token context window at $0.150 per 1M input tokens. Callable in 9 of 31 regions.

Suits

  • Nothing specific stands out from the published capabilities.

Think twice if

  • Long context. Only 32K tokens — long documents will need chunking, and agentic histories will hit the ceiling.
  • Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
  • Function calling. No tool use, so it cannot drive an agent loop or call your APIs — it answers from what is in the prompt.
  • Documents with layout. Text only. Scanned PDFs, screenshots and charts need a vision model or an OCR step first.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Summarisation and extraction

High-volume, structured output from unstructured input. Usually the cheapest workload to run well.

  • · Turning a million support tickets into a fixed schema of product, severity and root cause.
  • · Pulling line items, totals and dates off scanned invoices into JSON.
  • · Condensing every meeting transcript into decisions taken and owners assigned.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $0.150 $0.200 API
us-west-2 US West (Oregon) $0.150 $0.200 API
ca-central-1 Canada (Central) $0.170 $0.230 API
eu-west-1 EU (Ireland) $0.160 $0.220 API
eu-west-2 EU (London) $0.200 $0.260 API
eu-west-3 EU (Paris) $0.200 $0.260 API
ap-south-1 Asia Pacific (Mumbai) $0.180 $0.240 API
ap-southeast-2 Asia Pacific (Sydney) $0.200 $0.260 API
sa-east-1 South America (Sao Paulo) $0.250 $0.340 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Mistral 7B Instruct can be served without the request leaving 9 of the 25 jurisdictions with a Bedrock region.

Common questions

How much does Mistral 7B Instruct cost on Amazon Bedrock?
$0.150 per 1M input tokens and $0.200 per 1M output tokens in us-east-1.
Which AWS regions support Mistral 7B Instruct?
9 of 31 regions: us-east-1, us-west-2, ca-central-1, eu-west-1, eu-west-2, eu-west-3, ap-south-1, ap-southeast-2, sa-east-1.
What is the context window of Mistral 7B Instruct?
32,000 tokens (32K).
What is the model ID for Mistral 7B Instruct on Bedrock?
mistral.mistral-7b-instruct-v0:2.

Other Mistral AI models