Z AiTool use

GLM 5 Turbo pricing

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading closed-source models. With advanced agentic planning, deep backend reasoning, and iterative self-correction, GLM-5 moves beyond code generation to full-system construction and autonomous execution. This page tracks 5 listings in total. Highlighted lows are $0.600 per million input and $2.08 per million output (see table for which seller matches each).

203K context·5 providers·save up to 40% vs native via DeepInfra·verified Sep 5, 2026
Best input$0.600per 1M tokens · DeepInfra
Best output$2.08per 1M tokens · DeepInfra

Pricing across providers

All figures are list prices per million tokens unless a column says otherwise. 5 offers are listed for GLM 5 Turbo. Best input in this view: DeepInfra.

B
Baseten
Input / 1M
$0.950
Output / 1M
$3.15
D
DeepInfra
Input / 1M
$0.600
Output / 1M
$2.08
Cached in: $0.120
N
Novita
Input / 1M
$1.20
Output / 1M
$4.00
Cached in: $0.240
O
Openrouter
Input / 1M
$1.20
Output / 1M
$4.00
ZA
Z Ainative
Input / 1M
$1.00
Output / 1M
$3.20
Cached in: $0.200

Input vs output · per provider

Cost calculator

Use this block to stress test GLM 5 Turbo cost without a spreadsheet. All estimates come from public list rates in this page.

Provider

In: $0.950/M·Out: $3.15/M

0.095000¢ / req

0.157500¢ / req

Daily
$25
Monthly
$758
Annual
$9.2K
3% cheaper than the native API at this usage level

Model specifications

Context length, caps, and capability flags for GLM 5 Turbo. Values follow the main provider (Z Ai) record in our index.

Context window
202,752 tokens
Max output
128,000 tokens
Vision (images)
No
Tool / function calling
Yes
Streaming
No
Released
Feb 2026
Primary provider
Z Ai
Model family
N/A

Compare GLM 5 Turbo

Jump into a comparison when you want one table for two models instead of two tabs. 6 curated matches for GLM 5 Turbo.

Locked

Compare with

Pick a model on both sides.

Popular GLM 5 Turbo comparisons

Frequently asked questions

Read these after the table if you want plain language around GLM 5 Turbo rates. The short model note from our index: GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programmi...

GLM 5 Turbo costs $1.00 per million input tokens and $3.20 per million output tokens via the native API. Prompt caching reduces input costs to $0.20/M tokens.

Related models

Models close to GLM 5 Turbo by family, provider, and context window — each with live pricing in our catalog.