# GLM-5 API prices

GLM-5 (glm) — 205k context, 203k max output.
39 metered per-token offers from 42 provider(s), in USD per 1M tokens.
Prices as of 2026-10-04.

| Provider | Input $/1M | Output $/1M | Cache read $/1M | Context |
| --- | --- | --- | --- | --- |
| Vultr | $0.4 | $1.75 | — | 203k |
| GMI Cloud | $0.6 | $1.92 | $0.12 | 203k |
| Kilo Gateway | $0.6 | $1.92 | $0.12 | 198k |
| OpenRouter | $0.6 | $1.92 | $0.12 | 205k |
| NanoGPT | $0.5 | $2.55 | $0.13 | 200k |
| Alibaba (China) | $0.573 | $2.58 | — | 203k |
| LLM Gateway | $0.573 | $2.58 | — | 203k |
| ZenMux | $0.58 | $2.6 | $0.14 | 200k |
| 302.AI | $0.6 | $2.6 | — | 205k |
| DevPass (LLM Gateway) | $0.72 | $2.3 | $0.144 | 203k |
| DInference | $0.75 | $2.4 | — | 200k |
| CrossModel | $0.6 | $3 | $0.16 | 200k |
| Meganova | $0.8 | $2.56 | — | 203k |
| TokenGo | $0.89 | $3.2647 | $0.2226 | 205k |
| Baseten | $0.95 | $3.15 | $0.2 | 203k |
| FastRouter | $0.95 | $3.15 | — | 205k |
| Abacus | $1 | $3.2 | — | 205k |
| Amazon Bedrock | $1 | $3.2 | — | 203k |
| DigitalOcean | $1 | $3.2 | $0.2 | 64k |
| Eden AI | $1 | $3.2 | $0.2 | 203k |
| Hugging Face | $1 | $3.2 | $0.2 | 203k |
| Impossibl | $1 | $3.2 | $0.2 | 205k |
| LLM Gateway | $1 | $3.2 | $0.2 | 203k |
| LLM Gateway | $1 | $3.2 | $0.2 | 200k |
| LLM Gateway | $1 | $3.2 | $0.1 | 203k |
| LLM Gateway | $1 | $3.2 | $0.2 | 203k |
| Merge Gateway | $1 | $3.2 | $0.2 | 200k |
| NovitaAI | $1 | $3.2 | $0.2 | 203k |
| Ofox | $1 | $3.2 | $0.2 | 205k |
| OpenCode Zen | $1 | $3.2 | $0.2 | 205k |
| OrcaRouter | $1 | $3.2 | $0.26 | 205k |
| Poe | $1 | $3.2 | $0.2 | 205k |
| Tempr Gateway | $1 | $3.2 | $0.2 | 205k |
| TensorX | $1 | $3.2 | $0.25 | 203k |
| Venice AI | $1 | $3.2 | $0.2 | 198k |
| Vercel AI Gateway | $1 | $3.2 | — | 203k |
| Z.AI | $1 | $3.2 | $0.2 | 205k |
| Zhipu AI | $1 | $3.2 | $0.2 | 205k |
| Melious | $1.1012 | $3.3617 | $0.2666 | 203k |

## Summary

- Cheapest input: $0.4 per 1M tokens (Vultr)
- Cheapest output: $1.75 per 1M tokens (Vultr)
- First-party: $3.2 per 1M output tokens (Z.AI)
- Free offers: Alibaba Coding Plan, Alibaba Coding Plan (China), Alibaba Token Plan, Alibaba Token Plan (China), SCNet Token Plan, Tencent Coding Plan (China)
- Subscription plans (not per-token): Alibaba Coding Plan, Alibaba Coding Plan (China), Alibaba Token Plan, Alibaba Token Plan (China), SCNet Token Plan, Tencent Coding Plan (China)
- Released: 2026-02-12
- Inputs: text

## FAQ

### What is the cheapest GLM-5 API?

As of Oct 4, 2026, Vultr has the lowest GLM-5 output price at $1.75 per 1M tokens, and Vultr has the lowest input price at $0.4 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

### How much does GLM-5 cost on Z.AI?

Z.AI charges $1 per 1M input tokens and $3.2 per 1M output tokens for GLM-5.

### How many providers offer GLM-5?

42 providers list GLM-5 on Sovyron; 39 of them sell it at a metered per-token price.

### Is GLM-5 free?

No provider in the Sovyron catalog lists a free tier for GLM-5.

### Is GLM-5 included in a subscription plan?

Yes. 3 flat-rate plan(s) include it; the cheapest is Qwen (Alibaba Cloud) Token Plan Lite at $8/month.

### What is the context window of GLM-5?

GLM-5 supports a 205k-token context window and up to 203k output tokens.

HTML page: https://sovyron.com/models/glm-5/
Full dataset: https://sovyron.com/data/catalog.json
