GLM 4.7 Flash
Z.AI · Active · available in 11 of 16 regions
Batch inferenceStreaming Tool useVisionPrompt cachingEmbeddingsFine-tuning
- Context window
- —
- Max output
- —
- Input / 1M
- $0.070
- us-east-1
- Output / 1M
- $0.400
Model & inference profile IDs
Base model ID — direct on-demand invoke
Pricing by region
USD per 1M tokens| Region | Input | Output | Cache read | Cache write | Batch in | Batch out | Source |
|---|---|---|---|---|---|---|---|
| us-east-1 US East (N. Virginia) | $0.070 | $0.400 | — | — | $0.035 | $0.200 | API |
| us-east-2 US East (Ohio) | $0.070 | $0.400 | — | — | $0.035 | $0.200 | API |
| us-west-2 US West (Oregon) | $0.070 | $0.400 | — | — | $0.035 | $0.200 | API |
| eu-west-1 EU (Ireland) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
| eu-west-2 EU (London) | $0.110 | $0.620 | — | — | $0.055 | $0.310 | API |
| eu-central-1 EU (Frankfurt) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
| eu-north-1 EU (Stockholm) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
| ap-southeast-2 Asia Pacific (Sydney) | $0.070 | $0.410 | — | — | $0.035 | $0.205 | API |
| ap-northeast-1 Asia Pacific (Tokyo) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
| ap-south-1 Asia Pacific (Mumbai) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
| sa-east-1 South America (Sao Paulo) | $0.080 | $0.480 | — | — | $0.040 | $0.240 | API |
Region availability
On-demand us-east-1 On-demand On-demand us-east-2 On-demand On-demand us-west-2 On-demand Not available ca-central-1 Not available On-demand eu-west-1 On-demand On-demand eu-west-2 On-demand Not available eu-west-3 Not available On-demand eu-central-1 On-demand On-demand eu-north-1 On-demand Not available ap-southeast-1 Not available On-demand ap-southeast-2 On-demand On-demand ap-northeast-1 On-demand Not available ap-northeast-2 Not available Not available ap-northeast-3 Not available On-demand ap-south-1 On-demand On-demand sa-east-1 On-demand
TPM / RPM are default account quotas from AWS Service Quotas where a model-specific limit is published. They are per-account defaults and adjustable on request.
Common questions
- How much does GLM 4.7 Flash cost on Amazon Bedrock?
- $0.070 per 1M input tokens and $0.400 per 1M output tokens in us-east-1.
- Which AWS regions support GLM 4.7 Flash?
- 11 of 16 regions: us-east-1, us-east-2, us-west-2, eu-west-1, eu-west-2, eu-central-1, eu-north-1, ap-southeast-2, ap-northeast-1, ap-south-1, sa-east-1.
- What is the model ID for GLM 4.7 Flash on Bedrock?
- zai.glm-4.7-flash.