GLM-5.3 (max) pricing in 2026

What one task costs on GLM-5.3 (max)

Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with GLM-5.3 (max). Figures are USD at the rates above and exclude prompt caching and batch discounts.

TaskTokens (in / out)Cost
Code review of a 500-line pull request60,000 / 4,000$0.102
Debug one failing test40,000 / 6,000$0.082
Generate a landing page in one shot3,000 / 30,000$0.136
Draft a 1,200-word article2,000 / 2,500$0.014
Answer a question from a 20-page document25,000 / 800$0.039
Multi-step agent run (10 tool calls)250,000 / 25,000$0.460
All of the above, once$0.833

Cheaper at the same or better capability: GLM-5.3 FlashX runs the same set of tasks for $0.226 instead of $0.833.

GLM-5.3 (max) by Z.ai costs $1.40 per 1M input tokens and $4.40 per 1M output tokens — about 274.0k tokens per $1. Prices last verified 2026-09-19.

The verdict

Z.ai's flagship, index 45 at $1.40 / $4.40 per 1M tokens — DeepSeek-adjacent value with genuinely strong front-end design. Its big weakness is serving: Z.ai doesn't yet have the compute to run it fast, so expect slow responses and quick usage burn until it reaches more providers.

A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with GLM-5.3 (max) costs about $0.23 at current prices.

Who is GLM-5.3 (max) for?

GLM-5.3 is Z.ai's flagship — an index of 45 at $1.40 / $4.40 per 1M, DeepSeek-adjacent value with genuinely strong front-end design. A cheaper GLM-5.3 Flash tier ($0.15 / $0.50) now covers high-volume work.

The caveat is serving: Z.ai doesn't yet have the compute to run it fast, so expect slow responses and quick usage burn until it reaches more providers. If you can tolerate latency, it's a standout open-weight design model at a value price.

Strengths & weaknesses

  • Strength: Strong design at a frontier value price
  • Best for: Design, Open-weight
  • Weakness: Very slow — Z.ai lacks the compute to serve it reliably and usage runs out quickly.

GLM-5.3 (max) key facts

ProviderGLM pricing
Input price$1.40 / 1M tokens
Output price$4.40 / 1M tokens
Tokens / $1274.0k
Intelligence / $1123,288
Context window1.0M tokens
Tierfrontier
SpeedSlowest
Intelligence index45 / 100
Released2026-08-18
Last verified2026-09-19

What GLM-5.3 (max) delivers for your budget

For a coding workload, here is roughly what each monthly budget buys:

BudgetTokens / monthMonths of workload
$209.6M14.7x
$5023.9M36.8x
$10047.8M73.5x

Other GLM models

How GLM-5.3 (max) compares to its closest rivals

Rather than repeating the full catalog on every page, here are the three current models nearest to GLM-5.3 (max) on the intelligence index — and why you might still pick one instead:

  • Qwen3.8 Max (Qwen) — $2 input / $6 output per 1M tokens. Why you might pick it: Multimodal frontier capability from Alibaba.
  • Grok 4.6 (Grok) — $2 input / $6 output per 1M tokens. Why you might pick it: Frontier index at mid price.
  • Kimi K3 (Kimi) — $3 input / $15 output per 1M tokens. Why you might pick it: Top-tier reasoning from Moonshot.

GLM-5.3 (max) — frequently asked questions

How much does GLM-5.3 (max) cost?

GLM-5.3 (max) by Z.ai costs $1.40 per 1M input tokens and $4.40 per 1M output tokens.

How many tokens per dollar does GLM-5.3 (max) give you?

On a typical 1:3 input-to-output mix, GLM-5.3 (max) delivers around 274.0k tokens per dollar — an intelligence-per-dollar score of 123288.

What is GLM-5.3 (max)'s context window?

GLM-5.3 (max) supports a context window of 1.0M tokens.

Is GLM-5.3 (max) worth it in 2026?

GLM-5.3 (max) is GLM's frontier tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.

Advertisement