GLM 5.3 Flash pricing
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while... This page tracks 7 listings in total. Highlighted lows are $0.090 per million input and $0.300 per million output (see table for which seller matches each).
Pricing across providers
Use this table to read GLM 5.3 Flash list prices. We show 7 sources right now. Lowest input in the grid: Openrouter. The chart below the table helps when output prices are much higher than input prices.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
A Aihubmix | $0.113 | $0.394 | $0.028 | — |
F Friendliai | $0.150 | $0.500 | $0.030 | — |
N Nebius | $0.150 | $0.500 | — | — |
O Openrouter | $0.090 | $0.300 | $0.018 | — |
TA Together AI | $0.150 | $0.500 | $0.030 | — |
W Wandb | $0.150 | $0.500 | $0.050 | — |
ZA Z Ainative | $0.150 | $0.500 | $0.030 | — |
Input vs output · per provider
Cost calculator
Use this block to stress test GLM 5.3 Flash cost without a spreadsheet. All estimates come from public list rates in this page.
Provider
0.011270¢ / req
0.019720¢ / req
Model specifications
These fields describe GLM 5.3 Flash as we store it (source: Z Ai). They sit next to price so buyers can check limits and tools in one place.
- Context window
- 1,048,576 tokens
- Max output
- 131,072 tokens
- Vision (images)
- Yes
- Tool / function calling
- Yes
- Streaming
- Yes
- Released
- Aug 2026
- Primary provider
- Z Ai
- Model family
- N/A
Compare GLM 5.3 Flash
Jump into a comparison when you want one table for two models instead of two tabs. 6 curated matches for GLM 5.3 Flash.
Locked
Compare with
Pick a model on both sides.
Popular GLM 5.3 Flash comparisons
- GLM 5.3 Flash vs GPT-4o
GLM 5.3 Flash 95% cheaper on output
- GLM 5.3 Flash vs Claude Sonnet 4.6
GLM 5.3 Flash 96% cheaper on output
- GLM 5.3 Flash vs Gemini 2.0 Flash
Gemini 2.0 Flash 19% cheaper on output
- GLM 5.3 Flash vs Llama 3.1 70B
Compare pricing side by side
- GLM 5.3 Flash vs Mistral 7B
Mistral 7B 50% cheaper on output
- GLM 5.3 Flash vs Amazon Nova Lite V1 0
Compare pricing side by side
Frequently asked questions
Quick frequently asked items for GLM 5.3 Flash pricing and limits. The short model note from our index: GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Related models
Models close to GLM 5.3 Flash by family, provider, and context window — each with live pricing in our catalog.