Nova 2 Lite

Amazon · Active · available in 25 of 31 regions

VisionPrompt cachingBatch inferenceStreaming Tool useEmbeddingsFine-tuning
Context window
Max output
Input / 1M
$0.300
us-east-1
Output / 1M
$2.50

What Nova 2 Lite is good for

editorial

Nova 2 Lite offers prompt caching with image input at $0.300 per 1M input tokens. Callable in 25 of 31 regions.

Suits

  • Repeated prompts. Supports prompt caching, so a re-sent prefix costs about 4× less than fresh input. On agentic loops that resend a long system prompt every step, this moves the bill more than the headline price does.
  • Images. Accepts image input: screenshots, scanned documents, charts and UI captures.
  • Bulk jobs. Batch inference is available, typically at half the on-demand rate, for work that can wait — backfills, nightly enrichment, evaluation runs.

Think twice if

  • Function calling. No tool use, so it cannot drive an agent loop or call your APIs — it answers from what is in the prompt.
  • Invocation. Reachable only through a cross-region inference profile — calling the bare model ID returns a validation error. Use an ID such as eu.amazon.nova-2-lite-v1:0.

This section is judgement, not data from AWS. It is composed from the capabilities, context window, price position and region coverage shown elsewhere on this page — so it stays in step with the daily snapshot rather than going stale.

Where teams typically use it

Chat and assistants

Conversational work where responsiveness matters as much as depth.

  • · A customer-facing assistant where the first token needs to appear in under a second.
  • · An in-product copilot that explains what the user is looking at and answers follow-ups.
  • · A triage bot that qualifies an incoming request before routing it to the right team.

RAG and retrieval

Answering from your own documents, with the retrieved passages in the prompt.

  • · An internal search over contracts and policies that quotes the clause it answered from.
  • · A product assistant grounded in the current documentation rather than the training data.
  • · A research summariser that reads twenty retrieved papers and reconciles where they disagree.

Model & inference profile IDs

Base model ID — direct on-demand invoke

Cross-region inference profiles (4)

Regions marked Inference profile only require a profile ID rather than the base model ID — invoking the base ID there returns a validation error. Regional (us., eu.) profiles carry a 10% premium over global.

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $0.300 $2.50 $0.075 $0 $0.150 $1.25 API
us-east-2 US East (Ohio) $0.300 $2.50 $0.075 $0 $0.150 $1.25 API
us-west-1 US West (N. California) $0.390 $3.21 $0.098 $0 $0.195 $1.60 API
us-west-2 US West (Oregon) $0.300 $2.50 $0.075 $0 $0.150 $1.25 API
ca-central-1 Canada (Central) $0.320 $2.67 $0.080 $0 $0.160 $1.33 API
ca-west-1 Canada West (Calgary) $0.340 $2.81 $0.850 $0 $0.170 $1.41 API
eu-west-1 EU (Ireland) $0.340 $2.87 $0.085 $0 $0.170 $1.44 API
eu-west-2 EU (London) $0.420 $3.52 $0.105 $0 $0.210 $1.76 API
eu-west-3 EU (Paris) $0.440 $3.69 $0.110 $0 $0.220 $1.84 API
eu-central-1 EU (Frankfurt) $0.390 $3.27 $0.098 $0 $0.195 $1.64 API
eu-north-1 EU (Stockholm) $0.330 $2.72 $0.083 $0 $0.195 $1.36 API
eu-south-1 EU (Milan) $0.481 $4.10 $0.083 $0 $0.240 $2.00 API
eu-south-2 Europe (Spain) $0.330 $2.75 $0.083 $0 $0.165 $1.38 API
ap-east-2 Asia Pacific (Taipei) $0.360 $3.01 $0.090 $0 $0.180 $1.51 API
ap-northeast-1 Asia Pacific (Tokyo) $0.360 $3.01 $0.090 $0 $0.180 $1.51 API
ap-northeast-2 Asia Pacific (Seoul) $0.360 $2.96 $0.090 $0 $0.180 $1.48 API
ap-south-1 Asia Pacific (Mumbai) $0.350 $2.95 $0.087 $0 $0.175 $1.47 API
ap-southeast-1 Asia Pacific (Singapore) $0.410 $3.39 $0.102 $0 $0.205 $1.69 API
ap-southeast-2 Asia Pacific (Sydney) $0.320 $2.63 $0.080 $0 $0.160 $1.31 API
ap-southeast-3 Asia Pacific (Jakarta) $0.320 $2.64 $0.080 $0 $0.160 $1.31 API
ap-southeast-4 Asia Pacific (Melbourne) $0.330 $2.71 $0.083 $0 $0.165 $1.35 API
ap-southeast-5 Asia Pacific (Malaysia) $0.410 $3.39 $0.102 $0 $0.205 $1.69 API
ap-southeast-6 Asia Pacific (New Zealand) $0.320 $2.63 $0.080 $0 $0.160 $1.31 API
ap-southeast-7 Asia Pacific (Thailand) $0.410 $3.39 $0.102 $0 $0.205 $1.69 API
il-central-1 Israel (Tel Aviv) $0.380 $3.13 $0.095 $0 $0.190 $1.66 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Where this model actually runs

Under a strict residency constraint, Nova 2 Lite can be served without the request leaving 3 of the 25 jurisdictions with a Bedrock region.

Nova 2 Lite appears in

Common questions

How much does Nova 2 Lite cost on Amazon Bedrock?
$0.300 per 1M input tokens and $2.50 per 1M output tokens in us-east-1. Cached input reads at $0.075, roughly a tenth of the input rate, which dominates the bill on agentic workloads that re-send a long prefix.
Which AWS regions support Nova 2 Lite?
25 of 31 regions: us-east-1, us-east-2, us-west-1, us-west-2, ca-central-1, ca-west-1, eu-west-1, eu-west-2, eu-west-3, eu-central-1, eu-north-1, eu-south-1, eu-south-2, ap-east-2, ap-northeast-1, ap-northeast-2, ap-south-1, ap-southeast-1, ap-southeast-2, ap-southeast-3, ap-southeast-4, ap-southeast-5, ap-southeast-6, ap-southeast-7, il-central-1. Regions marked profile-only need a cross-region inference profile ID such as eu.amazon.nova-2-lite-v1:0 rather than the bare model ID.
What is the model ID for Nova 2 Lite on Bedrock?
amazon.nova-2-lite-v1:0. In regions where it is only reachable through a cross-region inference profile, use a prefixed ID instead: eu.amazon.nova-2-lite-v1:0, global.amazon.nova-2-lite-v1:0, jp.amazon.nova-2-lite-v1:0.

Other Amazon models