Models / GLM 5.3 pricing
GLM 5.3 pricing
Z.ai · 1310k context · open-weight — what it really costs, per token and per task.
| Per 1M tokens | Price |
|---|---|
| Input | $1.40 |
| Output | $4.40 |
| Cached input | $0.140 |
List price last verified 2026-09-02.
The same task on GLM 5.3 vs 3 alternatives
Representative call: 1,500 input + 500 output tokens.
- • Gemini 3.7 Flash: intro pricing ends — list price doubles — this task costs 2x from 2027-01-01 ($1.5/M in · $7.5/M out).
- • Gemini 3.6 Flash: intro pricing ends — list price doubles — this task costs 2x from 2027-01-01 ($1.5/M in · $7.5/M out).
- • DeepSeek V4 Pro: list price shown is off-peak — peak (UTC 00:30–16:30) bills 2x.
How much does GLM 5.3 cost per 1M tokens?
GLM 5.3 costs $1.40 per 1M input tokens and $4.40 per 1M output tokens (cached input: $0.140/M). Last verified 2026-09-02.
How much does a typical GLM 5.3 call cost?
A representative call (1,500 input + 500 output tokens) costs about $0.0043 on GLM 5.3.
What are cheaper alternatives to GLM 5.3?
Comparable models: Gemini 3.7 Flash ($0.0030/call), Gemini 3.6 Flash ($0.0030/call), DeepSeek V4 Pro ($0.0020/call). The right pick depends on the task — GLM 5.3 scores strongest on code / review.
Is GLM 5.3 the right model for your work?
Scan your real usage in 60 seconds and see what every task would cost on every model. Nothing leaves your machine.
Run the free scanProvider list prices · capability scores are seed priors, refined from real usage.