Together AI GLM-5.3 Flash — status, pricing and lifecycle
No. GLM-5.3 Flash is active at Together AI and has no announced end date.
What kind of model is this?
Tags describe what GLM-5.3 Flash takes in, what it is built for and how you can get it. Modality, access and language tags follow Together AI's own documentation; tier and use-case tags are our reading of the line-up.
- Modality
- Use case
- Tier
- Access
Specifications
- Input price
- $0.15 / M tokens
- Output price
- $0.50 / M tokens
- Price band
- Budget
- Context window
- 1M tokens
- Max output
- unknown
- Input
- Text, image
- Output
- Text
- Knowledge cutoff
- unknown
- Tool use
- unknown
- Structured output
- unknown
- Weights licence
- GLM Model License (version-specific)
- Model maker
- Z.ai
“unknown” means the figure has not been published, or we have not verified it yet — never that the answer is no. “not published” under the weights licence means this model is not offered with downloadable weights. The price band is ours: it bands the blended price (three parts input to one part output) as budget (≤ $0.50), standard (≤ $3), premium (≤ $12) or flagship.
Lifecycle
Together AI has not announced a deprecation or shutdown date for this model. That is a statement about what has been published, not a guarantee — providers typically announce a shutdown months ahead, and this page updates when they do.
- Released unknown
- Deprecated not announced
- Retired not announced
What is it good for?
No editorial profile for this model yet. The specifications above come from Together AI's own documentation and the suggestions below are derived from the data — shared tags, tier and price band — not written by hand.
Alternatives to consider
-
MiniMax M3 Derived
Alternative from DeepInfra, same tier (small), similar price, also open weights; shares chat, coding, reasoning, agents and long-context.
$0.28 in · $1.10 out / M tokens 524.3K context
-
Gemini 2.5 Flash-Lite Derived
Alternative from Google (Gemini), same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.
$0.10 in · $0.40 out / M tokens 1M context
-
GPT-5 nano Derived
Alternative from Microsoft Azure AI Foundry, same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.
$0.05 in · $0.40 out / M tokens 400K context
-
GPT-5.4 nano Derived
Alternative from Microsoft Azure AI Foundry, same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.
$0.20 in · $1.25 out / M tokens 400K context
-
GPT-5.6 luna Derived
Alternative from Microsoft Azure AI Foundry, same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.
$0.20 in · $1.20 out / M tokens 1.1M context
Derived suggestions are ranked from the data — shared tags, the same tier and a comparable price band — not from benchmark scores. Check the capability fit yourself: we track lifecycle, price and specifications.
Sources
- together.ai https://www.together.ai/pricing
- docs.together.ai https://docs.together.ai/docs/serverless-models
Verified against the source on . Dataset snapshot . Providers can change a date without notice — the link above is always the final word.
Common questions
- What does GLM-5.3 Flash cost?
- $0.15 per million input tokens and $0.50 per million output tokens, as last verified on 2026-09-10.
- What is the context window of GLM-5.3 Flash?
- 1M tokens.
- Is GLM-5.3 Flash deprecated?
- No. GLM-5.3 Flash is active at Together AI and has no announced end date.