Skip to content
Sovyron

GLM-5V-Turbo

glm203k context203k max outputreleased Apr 1, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest GLM-5V-Turbo API price is $0.72 per 1M input tokens (302.AI) and $3.1946 per 1M output tokens (ZenMux). GLM-5V-Turbo is offered by 16 providers in the glm family, reasoning enabled. Z.AI's own rate is $4 per 1M output tokens.

Cheapest input
$0.72
Cheapest output
$3.1946
Official (Z.AI)
$4
output per 1M tokens
Providers
16
offering this model
Spread
6.9×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

GLM-5V-Turbo API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
302.AICheapest$0.72$3.2——200k
ZenMux$0.726Cheapest$3.1946$0.1743—200k
Cortecs$1.186$3.955$0.296$1.544203k
Eden AI$1.2$4$0.24—203k
Kilo Gateway$1.2$4$0.24—203k
DevPass (LLM Gateway)$1.2$4$0.24—200k
LLM Gateway$1.2$4$0.24—200k
NanoGPT$1.2$4$0.24—203k
Ofox$1.2$4$0.24—200k
OpenRouter$1.2$4$0.24—203k
Tempr Gateway$1.2$4$0.24free200k
TensorX$1.2$4$0.3$1.5203k
Vercel AI Gateway$1.2$4$0.24—200k
Z.AI$1.2$4$0.24free200k
Venice AI$1.5$5$0.3—200k
Zhipu AI$5$22$1.2free200k

GLM-5V-Turbo pricing FAQ

What is the cheapest GLM-5V-Turbo API?

As of Oct 4, 2026, ZenMux has the lowest GLM-5V-Turbo output price at $3.1946 per 1M tokens, and 302.AI has the lowest input price at $0.72 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does GLM-5V-Turbo cost on Z.AI?

Z.AI charges $1.2 per 1M input tokens and $4 per 1M output tokens for GLM-5V-Turbo.

How many providers offer GLM-5V-Turbo?

16 providers list GLM-5V-Turbo on Sovyron; 16 of them sell it at a metered per-token price.

Is GLM-5V-Turbo free?

No provider in the Sovyron catalog lists a free tier for GLM-5V-Turbo.

What is the context window of GLM-5V-Turbo?

GLM-5V-Turbo supports a 203k-token context window and up to 203k output tokens.