GLM-5.2 API pricing on Wafer
Every Wafer GLM-5.2 rate we track, compared against 7 providers serving the same model.
The cheapest GLM-5.2 input price we currently track is $0.750/1M on Deep Infra.
Wafer GLM-5.2 rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $1.20/1M | $4.10/1M |
About GLM-5.2
- Creator
- Zhipu
- Context
- 1.0M
- Modalities
- text
GLM-5.2 on other providers
About Wafer
Wafer is a San Francisco-based inference provider serving open-source LLMs through a serverless API and dedicated endpoints. Its platform uses AI agents to profile workloads and optimize the model, serving engine, kernel, and hardware combination for each deployment.
Frequently Asked Questions
How much does GLM-5.2 cost on Wafer?
Wafer serves GLM-5.2 from $1.20 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Wafer the cheapest way to run GLM-5.2?
Wafer ranks #3 of 7 providers we track serving GLM-5.2. The cheapest input price right now is on Deep Infra. Price is only one factor — throughput, latency, and context limits differ between providers.
Does Wafer offer batch or cached pricing for GLM-5.2?
We track 1 GLM-5.2 offering from Wafer. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.