Amazon Bedrock GLM 4.7 Flash — status, pricing and lifecycle
No. GLM 4.7 Flash is active at Amazon Bedrock and has no announced end date.
What kind of model is this?
Tags describe what GLM 4.7 Flash takes in, what it is built for and how you can get it. Modality, access and language tags follow Amazon Bedrock's own documentation; tier and use-case tags are our reading of the line-up.
- Modality
- Use case
- Tier
- Access
Specifications
- Input price
- $0.07 / M tokens
- Output price
- $0.40 / M tokens
- Price band
- Budget
- Context window
- 203K tokens
- Max output
- unknown
- Input
- Text
- Output
- Text
- Knowledge cutoff
- unknown
- Tool use
- unknown
- Structured output
- unknown
- Weights licence
- MIT License
- Model maker
- Z.ai
“unknown” means the figure has not been published, or we have not verified it yet — never that the answer is no. “not published” under the weights licence means this model is not offered with downloadable weights. The price band is ours: it bands the blended price (three parts input to one part output) as budget (≤ $0.50), standard (≤ $3), premium (≤ $12) or flagship.
Lifecycle
Amazon Bedrock has not announced a deprecation or shutdown date for this model. That is a statement about what has been published, not a guarantee — providers typically announce a shutdown months ahead, and this page updates when they do.
- Released unknown
- Deprecated not announced
- Retired not announced
What is it good for?
No editorial profile for this model yet. The specifications above come from Amazon Bedrock's own documentation and the suggestions below are derived from the data — shared tags, tier and price band — not written by hand.
Alternatives to consider
-
Claude Haiku 4.5 Derived
Same provider, same tier (small), close in price; shares chat, coding, reasoning, agents and long-context.
$1.00 in · $5.00 out / M tokens 200K context
-
Qwen3 235B A22B 2507 Derived
Same provider, same tier (small), similar price, also open weights; shares chat, coding, agents and long-context.
$0.22 in · $0.88 out / M tokens 256K context
-
Qwen3 Coder 30B A3B Instruct Derived
Same provider, same tier (small), similar price, also open weights; shares chat, coding, agents and long-context.
$0.15 in · $0.60 out / M tokens 256K context
-
Qwen3 Next 80B A3B Derived
Same provider, same tier (small), similar price, also open weights; shares chat, coding, agents and long-context.
$0.14 in · $1.20 out / M tokens 256K context
-
GLM 5.3 Flash Derived
Alternative from Cloudflare Workers AI, same tier (small), similar price, also open weights; shares chat, coding, reasoning, agents and long-context.
$0.15 in · $0.50 out / M tokens 1M context
Derived suggestions are ranked from the data — shared tags, the same tier and a comparable price band — not from benchmark scores. Check the capability fit yourself: we track lifecycle, price and specifications.
Sources
- pricing.us-east-1.amazonaws.com https://pricing.us-east-1.amazonaws.com/offers/v1.0/aws/AmazonBedrock/current/us-east-1/index.json
- docs.aws.amazon.com https://docs.aws.amazon.com/bedrock/latest/userguide/model-card-zai-glm-4-7-flash.html
Verified against the source on . Dataset snapshot . Providers can change a date without notice — the link above is always the final word.
Common questions
- What does GLM 4.7 Flash cost?
- $0.07 per million input tokens and $0.40 per million output tokens, as last verified on 2026-09-10.
- What is the context window of GLM 4.7 Flash?
- 203K tokens.
- Is GLM 4.7 Flash deprecated?
- No. GLM 4.7 Flash is active at Amazon Bedrock and has no announced end date.