Mistral Small (24.02)
Mistral AI · Active · available in 1 of 31 regions
- Context window
- 32K
- 32,000 tokens
- Max output
- —
- Input / 1M
- $1.00
- us-east-1
- Output / 1M
- $3.00
What Mistral Small (24.02) is good for
editorialMistral Small (24.02) offers a 32K-token context window with tool use at $1.00 per 1M input tokens. Callable in 1 of 31 regions.
Suits
- Tool use. Can call functions, so it can drive retrieval, look things up, and take actions rather than only answering.
- Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.
Think twice if
- Long context. Only 32K tokens — long documents will need chunking, and agentic histories will hit the ceiling.
- Agentic loops. No prompt caching published, so every step re-pays full price for the prompt prefix. Expensive for agents that resend a long context.
- Documents with layout. Text only. Scanned PDFs, screenshots and charts need a vision model or an OCR step first.
- Data residency. Available in only 1 of 31 regions, so it may not clear a residency requirement.
This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.
Model & inference profile IDs
Base model ID — direct on-demand invoke
Pricing by region
USD per 1M tokens| Region | Input | Output | Cache read | Cache write | Batch in | Batch out | Source |
|---|---|---|---|---|---|---|---|
| us-east-1 US East (N. Virginia) | $1.00 | $3.00 | — | — | $0.500 | $1.50 | API |
Region availability
TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.
Where this model actually runs
Under a strict residency constraint, Mistral Small (24.02) can be served without the request leaving 1 of the 25 jurisdictions with a Bedrock region.
Mistral Small (24.02) appears in
Common questions
- How much does Mistral Small (24.02) cost on Amazon Bedrock?
- $1.00 per 1M input tokens and $3.00 per 1M output tokens in us-east-1.
- Which AWS regions support Mistral Small (24.02)?
- 1 of 31 regions: us-east-1.
- What is the context window of Mistral Small (24.02)?
- 32,000 tokens (32K).
- What is the model ID for Mistral Small (24.02) on Bedrock?
- mistral.mistral-small-2402-v1:0.