Mistral Large (24.07)

Mistral AI · Active · available in 1 of 31 regions

Tool useBatch inferenceStreaming VisionPrompt cachingEmbeddingsFine-tuning
Context window
128K
128,000 tokens
Max output
Input / 1M
$2.00
us-west-2
Output / 1M
$6.00

What Mistral Large (24.07) is good for

editorial

Mistral Large (24.07) offers a 128K-token context window with tool use at $2.00 per 1M input tokens. Callable in 1 of 31 regions.

Suits

  • Tool use. Can call functions, so it can drive retrieval, look things up, and take actions rather than only answering.
  • Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.

Think twice if

  • Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
  • Documents with layout. Text only. Scanned PDFs, screenshots and charts need a vision model or an OCR step first.
  • Data residency. Available in only 1 of 31 regions, so it may not clear a residency requirement.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Coding

Writing, reviewing and refactoring code, usually across more than one file.

  • · A review pass that flags real defects with a severity and leaves style alone.
  • · A framework upgrade applied across a repository, one module at a time, tests green between each.
  • · Turning a failing bug report into a reproducing test and then a fix.

Chat and assistants

Conversational work where responsiveness matters as much as depth.

  • · A customer-facing assistant where the first token needs to appear in under a second.
  • · An in-product copilot that explains what the user is looking at and answers follow-ups.
  • · A triage bot that qualifies an incoming request before routing it to the right team.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-west-2 US West (Oregon) $2.00 $6.00 $1.50 $4.50 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Mistral Large (24.07) can be served without the request leaving 1 of the 25 jurisdictions with a Bedrock region.

Mistral Large (24.07) appears in

Common questions

How much does Mistral Large (24.07) cost on Amazon Bedrock?
$2.00 per 1M input tokens and $6.00 per 1M output tokens in us-west-2.
Which AWS regions support Mistral Large (24.07)?
1 of 31 regions: us-west-2.
What is the context window of Mistral Large (24.07)?
128,000 tokens (128K).
What is the model ID for Mistral Large (24.07) on Bedrock?
mistral.mistral-large-2407-v1:0.

Other Mistral AI models