GLM 5

Z.AI · Active · available in 9 of 16 regions

Batch inferenceStreaming Tool useVisionPrompt cachingEmbeddingsFine-tuning
Context window
Max output
Input / 1M
$1.00
us-east-1
Output / 1M
$3.20

Model & inference profile IDs

Base model ID — direct on-demand invoke

Pricing by region

USD per 1M tokens
Region Input Output Cache read Cache write Batch in Batch out Source
us-east-1 US East (N. Virginia) $1.00 $3.20 $0.500 $1.60 API
us-east-2 US East (Ohio) $1.00 $3.20 $0.500 $1.60 API
us-west-2 US West (Oregon) $1.00 $3.20 $0.500 $1.60 API
eu-west-2 EU (London) $1.55 $4.96 $0.775 $2.48 API
eu-north-1 EU (Stockholm) $1.20 $3.84 $0.600 $1.92 API
ap-southeast-2 Asia Pacific (Sydney) $1.03 $3.30 $0.515 $1.65 API
ap-northeast-1 Asia Pacific (Tokyo) $1.20 $3.84 $0.600 $1.92 API
ap-south-1 Asia Pacific (Mumbai) $1.20 $3.84 $0.600 $1.92 API
sa-east-1 South America (Sao Paulo) $1.20 $3.84 $0.600 $1.92 API

Region availability

TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.

Common questions

How much does GLM 5 cost on Amazon Bedrock?
$1.00 per 1M input tokens and $3.20 per 1M output tokens in us-east-1.
Which AWS regions support GLM 5?
9 of 16 regions: us-east-1, us-east-2, us-west-2, eu-west-2, eu-north-1, ap-southeast-2, ap-northeast-1, ap-south-1, sa-east-1.
What is the model ID for GLM 5 on Bedrock?
zai.glm-5.

Other Z.AI models