GLM 4.7 Flash API pricing on Heabsy
Every Heabsy GLM 4.7 Flash rate we track, compared against 6 providers serving the same model.
The cheapest GLM 4.7 Flash input price we currently track is $0.060/1M on OpenRouter.
Heabsy GLM 4.7 Flash rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.090/1M | $0.620/1M |
About GLM 4.7 Flash
- Creator
- Zhipu
- Modalities
- text
GLM 4.7 Flash on other providers
About Heabsy
Heabsy is an OpenAI- and Anthropic-compatible inference API for open-weight models, operated by FEYA, s.r.o. in Bratislava, Slovakia, an EU company with no US parent. Models in the EEA tier, including Qwen3.8 27B, run on GPUs that Heabsy owns and operates inside the European Economic Area, with zero data retention: prompts and completions are processed in memory and never written to logs. A routed tier adds more than 25 open models (DeepSeek, Kimi, GLM, gpt-oss, Llama, Gemma, Qwen and others) bought from global providers under a European contract; every model in the catalog is labelled with where it is processed. One API key, one EU invoice, per-token pricing with no minimum spend or subscription.
Frequently Asked Questions
How much does GLM 4.7 Flash cost on Heabsy?
Heabsy serves GLM 4.7 Flash from $0.090 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Heabsy the cheapest way to run GLM 4.7 Flash?
Heabsy ranks #6 of 6 providers we track serving GLM 4.7 Flash. The cheapest input price right now is on OpenRouter. Price is only one factor — throughput, latency, and context limits differ between providers.
Does Heabsy offer batch or cached pricing for GLM 4.7 Flash?
We track 1 GLM 4.7 Flash offering from Heabsy. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.