GLM API pricing
Every GLM model we track, with the cheapest price per 1M tokens across 24 providers. Collected daily.
GLM API prices by model
Standard (real-time) rates per 1M tokens, newest first. “Cheapest” is the lowest input price across every provider we track; batch and cached-input rates are on each model page.
| Model | Cheapest input | Output | Cheapest on |
|---|---|---|---|
| GLM-5.3-Flash | $0.040 | $0.120 | Heabsy |
| GLM-5.3 | $0.039 | $7.00 | OpenRouter |
| GLM-5.21.0M context | $0.084 | $8.00 | OpenRouter |
| GLM-5.1200K context | $0.966 | $3.04 | OpenRouter |
| GLM-5128K context | $0.600 | $1.92 | GMI Cloud |
| GLM 5V Turbo | $1.20 | $4.00 | Novita AI |
| GLM-4.7128K context | $0.400 | $1.75 | Deep Infra |
| GLM 4.7 Flash | $0.060 | $0.400 | Runcrate |
| GLM-4.6128K context | $0.430 | $1.75 | OpenRouter |
| GLM-4.6V | $0.300 | $0.900 | Novita AI |
| GLM-4.5 Air128K context | $0.130 | $0.850 | Novita AI |
| GLM-4.5 | $0.600 | $2.20 | Prime Intellect |
| GLM-4.5V | $0.600 | $1.80 | Novita AI |
GLM price history
Average input price per 1M tokens across the providers serving each model, last 90 days.
Frequently Asked Questions
What is the cheapest GLM API?
Of the 13 GLM models we track, GLM-5.3 has the lowest input price right now: $0.039 per 1M input tokens on OpenRouter. Prices are collected daily from 24 providers — see the table above for every model.
Where is GLM-5.3-Flash cheapest?
GLM-5.3-Flash is served by 17 providers we track. The lowest standard input price is $0.040 per 1M tokens on Heabsy. Throughput, latency and context limits differ between hosts, so check the model page before committing.