Models / GLM 5.3 Flash pricing

GLM 5.3 Flash pricing

Z.ai · 1310k context · open-weight — what it really costs, per token and per task.

Per 1M tokensPrice
Input$0.090
Output$0.300
Cached input$0.009

List price last verified 2026-09-16.

The same task on GLM 5.3 Flash vs 3 alternatives

Representative call: 1,500 input + 500 output tokens.

GLM 5.3 Flash
$0.0003
Qwen 3 Max
$0.0031

How much does GLM 5.3 Flash cost per 1M tokens?

GLM 5.3 Flash costs $0.090 per 1M input tokens and $0.300 per 1M output tokens (cached input: $0.009/M). Last verified 2026-09-16.

How much does a typical GLM 5.3 Flash call cost?

A representative call (1,500 input + 500 output tokens) costs about $0.0003 on GLM 5.3 Flash.

What are cheaper alternatives to GLM 5.3 Flash?

Comparable models: Muse Spark 1.3 ($0.0040/call), Claude Haiku 4.5 ($0.0040/call), Qwen 3 Max ($0.0031/call). The right pick depends on the task — GLM 5.3 Flash scores strongest on classify / route.

Is GLM 5.3 Flash the right model for your work?

Scan your real usage in 60 seconds and see what every task would cost on every model. Nothing leaves your machine.

Run the free scan

Provider list prices · capability scores are seed priors, refined from real usage.