Models / GLM 5.3 Flash pricing
GLM 5.3 Flash pricing
Z.ai · 1310k context · open-weight — what it really costs, per token and per task.
| Per 1M tokens | Price |
|---|---|
| Input | $0.090 |
| Output | $0.300 |
| Cached input | $0.009 |
List price last verified 2026-09-16.
The same task on GLM 5.3 Flash vs 3 alternatives
Representative call: 1,500 input + 500 output tokens.
How much does GLM 5.3 Flash cost per 1M tokens?
GLM 5.3 Flash costs $0.090 per 1M input tokens and $0.300 per 1M output tokens (cached input: $0.009/M). Last verified 2026-09-16.
How much does a typical GLM 5.3 Flash call cost?
A representative call (1,500 input + 500 output tokens) costs about $0.0003 on GLM 5.3 Flash.
What are cheaper alternatives to GLM 5.3 Flash?
Comparable models: Muse Spark 1.3 ($0.0040/call), Claude Haiku 4.5 ($0.0040/call), Qwen 3 Max ($0.0031/call). The right pick depends on the task — GLM 5.3 Flash scores strongest on classify / route.
Is GLM 5.3 Flash the right model for your work?
Scan your real usage in 60 seconds and see what every task would cost on every model. Nothing leaves your machine.
Run the free scanProvider list prices · capability scores are seed priors, refined from real usage.