Llama 3 70B Instruct

Meta · Active · available in 5 of 16 regions

Streaming Tool useVisionPrompt cachingBatch inferenceEmbeddingsFine-tuning
Context window
8K
8,192 tokens
Max output
Input / 1M
$2.65
us-east-1
Output / 1M
$3.50

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $2.65 $3.50 API
us-west-2 US West (Oregon) $2.65 $3.50 API
ca-central-1 Canada (Central) $3.05 $4.03 API
eu-west-2 EU (London) $3.45 $4.55 API
ap-south-1 Asia Pacific (Mumbai) $3.18 $4.20 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Common questions

How much does Llama 3 70B Instruct cost on Amazon Bedrock?
$2.65 per 1M input tokens and $3.50 per 1M output tokens in us-east-1.
Which AWS regions support Llama 3 70B Instruct?
5 of 16 regions: us-east-1, us-west-2, ca-central-1, eu-west-2, ap-south-1.
What is the context window of Llama 3 70B Instruct?
8,192 tokens (8K).
What is the model ID for Llama 3 70B Instruct on Bedrock?
meta.llama3-70b-instruct-v1:0.

Other Meta models