Cloudflare Workers AI GLM 5.3 Flash — status, pricing and lifecycle

Active

No. GLM 5.3 Flash is active at Cloudflare Workers AI and has no announced end date.

@cf/zai-org/glm-5.3-flash last verified

What kind of model is this?

Tags describe what GLM 5.3 Flash takes in, what it is built for and how you can get it. Modality, access and language tags follow Cloudflare Workers AI's own documentation; tier and use-case tags are our reading of the line-up.

Modality
  • Text
  • Vision
Use case
  • Chat
  • Coding
  • Reasoning
  • Agents & tool use
  • Long context
Tier
  • Small
Access
  • Hosted API
  • Open weights

Specifications

Input price
$0.15 / M tokens
Output price
$0.50 / M tokens
Price band
Budget
Context window
1M tokens
Max output
unknown
Input
Text, image
Output
Text
Knowledge cutoff
unknown
Tool use
unknown
Structured output
unknown
Weights licence
GLM Model License (version-specific)
Model maker
Z.ai

“unknown” means the figure has not been published, or we have not verified it yet — never that the answer is no. “not published” under the weights licence means this model is not offered with downloadable weights. The price band is ours: it bands the blended price (three parts input to one part output) as budget (≤ $0.50), standard (≤ $3), premium (≤ $12) or flagship.

Lifecycle

Cloudflare Workers AI has not announced a deprecation or shutdown date for this model. That is a statement about what has been published, not a guarantee — providers typically announce a shutdown months ahead, and this page updates when they do.

  1. Released unknown
  2. Deprecated not announced
  3. Retired not announced

What is it good for?

No editorial profile for this model yet. The specifications above come from Cloudflare Workers AI's own documentation and the suggestions below are derived from the data — shared tags, tier and price band — not written by hand.

Alternatives to consider

  • MiniMax M3 Derived

    DeepInfra Active

    Alternative from DeepInfra, same tier (small), similar price, also open weights; shares chat, coding, reasoning, agents and long-context.

    $0.28 in · $1.10 out / M tokens 524.3K context

  • Kimi K2.6 Derived

    Cloudflare Workers AI Active

    Same provider, mid-tier, close in price, also open weights; shares chat, coding, reasoning, agents and long-context.

    $0.95 in · $4.00 out / M tokens 262.1K context

  • Kimi K2.7 Code Derived

    Cloudflare Workers AI Active

    Same provider, mid-tier, close in price, also open weights; shares chat, coding, reasoning, agents and long-context.

    $0.95 in · $4.00 out / M tokens 262.1K context

  • Gemini 2.5 Flash-Lite Derived

    Google (Gemini) Active

    Alternative from Google (Gemini), same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.

    $0.10 in · $0.40 out / M tokens 1M context

  • GPT-5 nano Derived

    Microsoft Azure AI Foundry Active

    Alternative from Microsoft Azure AI Foundry, same tier (small), similar price; shares chat, coding, reasoning, agents and long-context.

    $0.05 in · $0.40 out / M tokens 400K context

Derived suggestions are ranked from the data — shared tags, the same tier and a comparable price band — not from benchmark scores. Check the capability fit yourself: we track lifecycle, price and specifications.

Sources

Verified against the source on . Dataset snapshot . Providers can change a date without notice — the link above is always the final word.

Common questions

What does GLM 5.3 Flash cost?
$0.15 per million input tokens and $0.50 per million output tokens, as last verified on 2026-09-10.
What is the context window of GLM 5.3 Flash?
1M tokens.
Is GLM 5.3 Flash deprecated?
No. GLM 5.3 Flash is active at Cloudflare Workers AI and has no announced end date.

All Cloudflare Workers AI models and their end dates →