GLM-5V-Turbo
As of Oct 4, 2026, the cheapest GLM-5V-Turbo API price is $0.72 per 1M input tokens (302.AI) and $3.1946 per 1M output tokens (ZenMux). GLM-5V-Turbo is offered by 16 providers in the glm family, reasoning enabled. Z.AI's own rate is $4 per 1M output tokens.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| 302.AI | Cheapest$0.72 | $3.2 | — | — | 200k |
| ZenMux | $0.726 | Cheapest$3.1946 | $0.1743 | — | 200k |
| Cortecs | $1.186 | $3.955 | $0.296 | $1.544 | 203k |
| Eden AI | $1.2 | $4 | $0.24 | — | 203k |
| Kilo Gateway | $1.2 | $4 | $0.24 | — | 203k |
| DevPass (LLM Gateway) | $1.2 | $4 | $0.24 | — | 200k |
| LLM Gateway | $1.2 | $4 | $0.24 | — | 200k |
| NanoGPT | $1.2 | $4 | $0.24 | — | 203k |
| Ofox | $1.2 | $4 | $0.24 | — | 200k |
| OpenRouter | $1.2 | $4 | $0.24 | — | 203k |
| Tempr Gateway | $1.2 | $4 | $0.24 | free | 200k |
| TensorX | $1.2 | $4 | $0.3 | $1.5 | 203k |
| Vercel AI Gateway | $1.2 | $4 | $0.24 | — | 200k |
| Z.AI | $1.2 | $4 | $0.24 | free | 200k |
| Venice AI | $1.5 | $5 | $0.3 | — | 200k |
| Zhipu AI | $5 | $22 | $1.2 | free | 200k |
GLM-5V-Turbo pricing FAQ
What is the cheapest GLM-5V-Turbo API?
As of Oct 4, 2026, ZenMux has the lowest GLM-5V-Turbo output price at $3.1946 per 1M tokens, and 302.AI has the lowest input price at $0.72 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How much does GLM-5V-Turbo cost on Z.AI?
Z.AI charges $1.2 per 1M input tokens and $4 per 1M output tokens for GLM-5V-Turbo.
How many providers offer GLM-5V-Turbo?
16 providers list GLM-5V-Turbo on Sovyron; 16 of them sell it at a metered per-token price.
Is GLM-5V-Turbo free?
No provider in the Sovyron catalog lists a free tier for GLM-5V-Turbo.
What is the context window of GLM-5V-Turbo?
GLM-5V-Turbo supports a 203k-token context window and up to 203k output tokens.