Z AiTool use

GLM 5.1 pricing

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on... Live index: 7 priced offers. Best input $1.05 per million tokens from DeepInfra. Best output $3.50 per million tokens from DeepInfra.

203K context·7 providers·save up to 25% vs native via DeepInfra·verified Sep 4, 2026
Best input$1.05per 1M tokens · DeepInfra
Best output$3.50per 1M tokens · DeepInfra

Pricing across providers

All figures are list prices per million tokens unless a column says otherwise. 7 offers are listed for GLM 5.1. Best input in this view: DeepInfra.

AC
Alibaba Cloud
Input / 1M
$1.40
Output / 1M
$4.40
Cached in: $0.260
D
DeepInfra
Input / 1M
$1.05
Output / 1M
$3.50
Cached in: $0.205
N
Novita
Input / 1M
$1.38
Output / 1M
$4.40
Cached in: $0.260
O
Openrouter
Input / 1M
$1.05
Output / 1M
$3.50
Cached in: $0.525
QA
Qwen Ai Platform
Input / 1M
$1.40
Output / 1M
$4.40
Cached in: $0.260
Q
Qwencloud
Input / 1M
$1.40
Output / 1M
$4.40
Cached in: $0.260
ZA
Z Ainative
Input / 1M
$1.40
Output / 1M
$4.40
Cached in: $0.260

Input vs output · per provider

Cost calculator

Pick any of the providers above and type how many tokens you expect per day, week, or year. We turn that into rough dollar totals for GLM 5.1.

Provider

In: $1.40/M·Out: $4.40/M·Cache: $0.260/M

0.140000¢ / req

0.220000¢ / req

Daily
$36
Monthly
$1.1K
Annual
$13.1K

Model specifications

Quick spec sheet for GLM 5.1 before you dive back into pricing. Reported under Z Ai.

Context window
202,752 tokens
Max output
202,752 tokens
Vision (images)
No
Tool / function calling
Yes
Streaming
Yes
Released
Apr 2026
Primary provider
Z Ai
Model family
N/A

Compare GLM 5.1

Jump into a comparison when you want one table for two models instead of two tabs. 6 curated matches for GLM 5.1.

Locked

Compare with

Pick a model on both sides.

Popular GLM 5.1 comparisons

Frequently asked questions

Answers pull from the same numbers you see on this page. The short model note from our index: GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and co...

GLM 5.1 costs $1.40 per million input tokens and $4.40 per million output tokens via the native API. Prompt caching reduces input costs to $0.26/M tokens.

Related models

Models close to GLM 5.1 by family, provider, and context window — each with live pricing in our catalog.