GLM 4.7 Flash pricing
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards. Live index: 5 priced offers. Best input $0.0000 per million tokens from Z Ai. Best output $0.0000 per million tokens from Z Ai.
Pricing across providers
Every row is a seller of GLM 4.7 Flash with token pricing we track. The cheapest input in this snapshot is from Z Ai. The bar chart shows the same input and output dollars per million for a quick scan.
| Provider | Input / 1M | Output / 1M | Cached input | Batch |
|---|---|---|---|---|
C Cloudflare | $0.060 | $0.400 | — | — |
D DeepInfra | $0.060 | $0.400 | $0.010 | — |
N Novita | $0.070 | $0.400 | $0.010 | — |
O Openrouter | $0.060 | $0.400 | $0.010 | — |
ZA Z Ainative | $0.0000 | $0.0000 | $0.0000 | — |
Input vs output · per provider
Cost calculator
The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how GLM 4.7 Flash cost scales with traffic.
Provider
0.006050¢ / req
0.020000¢ / req
Model specifications
Context length, caps, and capability flags for GLM 4.7 Flash. Values follow the main provider (Z Ai) record in our index.
- Context window
- 200,000 tokens
- Max output
- 32,000 tokens
- Vision (images)
- Yes
- Tool / function calling
- Yes
- Streaming
- No
- Released
- Jan 2026
- Primary provider
- Z Ai
- Model family
- N/A
Compare GLM 4.7 Flash
Open a pair page to see GLM 4.7 Flash next to another model with a shared provider matrix. 6 shortcuts below.
Locked
Compare with
Pick a model on both sides.
Popular GLM 4.7 Flash comparisons
- GLM 4.7 Flash vs GPT-4o
Compare pricing side by side
- GLM 4.7 Flash vs Claude Sonnet 4.6
Compare pricing side by side
- GLM 4.7 Flash vs Gemini 2.0 Flash
Compare pricing side by side
- GLM 4.7 Flash vs Llama 3.1 70B
Compare pricing side by side
- GLM 4.7 Flash vs Mistral 7B
Compare pricing side by side
- GLM 4.7 Flash vs Amazon Nova Lite V1 0
Compare pricing side by side
Frequently asked questions
Answers pull from the same numbers you see on this page. The short model note from our index: As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, ...
Related models
Models close to GLM 4.7 Flash by family, provider, and context window — each with live pricing in our catalog.