Skip to main content
Wafer logo

GLM-5.1 API pricing on Wafer

Every Wafer GLM-5.1 rate we track, compared against 5 providers serving the same model.

Input from /1M
$1.00
Output /1M
$3.20
Price rank
#2 of 5
Last updated
August 4, 2026

The cheapest GLM-5.1 input price we currently track is $0.966/1M on OpenRouter.

Wafer GLM-5.1 rates

Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.

Wafer GLM-5.1 pricing by mode
ModeInputOutput
Standard$1.00/1M$3.20/1M

About GLM-5.1

Creator
Zhipu
Context
200K
Modalities
text
Full GLM-5.1 details and every provider →

GLM-5.1 on other providers

About Wafer

Wafer is a San Francisco-based inference provider serving open-source LLMs through a serverless API and dedicated endpoints. Its platform uses AI agents to profile workloads and optimize the model, serving engine, kernel, and hardware combination for each deployment.

Frequently Asked Questions

How much does GLM-5.1 cost on Wafer?

Wafer serves GLM-5.1 from $1.00 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.

Is Wafer the cheapest way to run GLM-5.1?

Wafer ranks #2 of 5 providers we track serving GLM-5.1. The cheapest input price right now is on OpenRouter. Price is only one factor — throughput, latency, and context limits differ between providers.

Does Wafer offer batch or cached pricing for GLM-5.1?

We track 1 GLM-5.1 offering from Wafer. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.