Z AiVisionTool use

GLM 4.7 Flash pricing

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, and tool collaboration, and has achieved leading performance among open-source models of the same size on several current public benchmark leaderboards. Live index: 12 priced offers. Best input $0.0000 per million tokens from Z Ai. Best output $0.0000 per million tokens from Z Ai.

200K context·12 providers·verified Aug 10, 2026
Best inputFreeper 1M tokens · Z Ai
Best outputFreeper 1M tokens · Z Ai

Pricing across providers

Every row is a seller of GLM 4.7 Flash with token pricing we track. The cheapest input in this snapshot is from Z Ai. The bar chart shows the same input and output dollars per million for a quick scan.

O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.060
Output / 1M
$0.400
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.060
Output / 1M
$0.400
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.060
Output / 1M
$0.400
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
O
Openrouter
Input / 1M
$0.070
Output / 1M
$0.400
Cached in: $0.0000
ZA
Z Ainative
Input / 1M
$0.0000
Output / 1M
$0.0000
Cached in: $0.0000
C
Cloudflare
Input / 1M
$0.060
Output / 1M
$0.400

Input vs output · per provider

Cost calculator

The calculator uses the same dollars per million tokens as the table. Adjust sliders to see how GLM 4.7 Flash cost scales with traffic.

Provider

In: $0.070/M·Out: $0.400/M0

0.007000¢ / req

0.020000¢ / req

Daily
$2.70
Monthly
$81
Annual
$986

Model specifications

Context length, caps, and capability flags for GLM 4.7 Flash. Values follow the main provider (Z Ai) record in our index.

Context window
200,000 tokens
Max output
32,000 tokens
Vision (images)
Yes
Tool / function calling
Yes
Streaming
No
Released
Jan 2026
Primary provider
Z Ai
Model family
N/A

Compare GLM 4.7 Flash

Open a pair page to see GLM 4.7 Flash next to another model with a shared provider matrix. 6 shortcuts below.

Locked

Compare with

Pick a model on both sides.

Popular GLM 4.7 Flash comparisons

Frequently asked questions

Answers pull from the same numbers you see on this page. The short model note from our index: As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning, ...

GLM 4.7 Flash costs $0.00 per million input tokens and $0.00 per million output tokens via the native API.

Also from Z Ai

Other models by Z Ai with live pricing in our catalog.