Z AiVisionTool use

GLM 5.3 Flash pricing

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while... This page tracks 7 listings in total. Highlighted lows are $0.090 per million input and $0.300 per million output (see table for which seller matches each).

1.0M context·7 providers·save up to 40% vs native via Openrouter·verified Sep 20, 2026
Best input$0.090per 1M tokens · Openrouter
Best output$0.300per 1M tokens · Openrouter

Pricing across providers

Use this table to read GLM 5.3 Flash list prices. We show 7 sources right now. Lowest input in the grid: Openrouter. The chart below the table helps when output prices are much higher than input prices.

A
Aihubmix
Input / 1M
$0.113
Output / 1M
$0.394
Cached in: $0.028
F
Friendliai
Input / 1M
$0.150
Output / 1M
$0.500
Cached in: $0.030
N
Nebius
Input / 1M
$0.150
Output / 1M
$0.500
O
Openrouter
Input / 1M
$0.090
Output / 1M
$0.300
Cached in: $0.018
TA
Together AI
Input / 1M
$0.150
Output / 1M
$0.500
Cached in: $0.030
W
Wandb
Input / 1M
$0.150
Output / 1M
$0.500
Cached in: $0.050
ZA
Z Ainative
Input / 1M
$0.150
Output / 1M
$0.500
Cached in: $0.030

Input vs output · per provider

Cost calculator

Use this block to stress test GLM 5.3 Flash cost without a spreadsheet. All estimates come from public list rates in this page.

Provider

In: $0.113/M·Out: $0.394/M·Cache: $0.028/M

0.011270¢ / req

0.019720¢ / req

Daily
$3.10
Monthly
$93
Annual
$1.1K
23% cheaper than the native API at this usage level

Model specifications

These fields describe GLM 5.3 Flash as we store it (source: Z Ai). They sit next to price so buyers can check limits and tools in one place.

Context window
1,048,576 tokens
Max output
131,072 tokens
Vision (images)
Yes
Tool / function calling
Yes
Streaming
Yes
Released
Aug 2026
Primary provider
Z Ai
Model family
N/A

Compare GLM 5.3 Flash

Jump into a comparison when you want one table for two models instead of two tabs. 6 curated matches for GLM 5.3 Flash.

Locked

Compare with

Pick a model on both sides.

Popular GLM 5.3 Flash comparisons

Frequently asked questions

Quick frequently asked items for GLM 5.3 Flash pricing and limits. The short model note from our index: GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

GLM 5.3 Flash costs $0.15 per million input tokens and $0.50 per million output tokens via the native API. Prompt caching reduces input costs to $0.03/M tokens.

Related models

Models close to GLM 5.3 Flash by family, provider, and context window — each with live pricing in our catalog.