Skip to content
Sovyron

GLM-4.6V

glm200k context128k max outputreleased Dec 8, 2025Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest GLM-4.6V API price is $0.137 per 1M input tokens (AIHubMix) and $0.411 per 1M output tokens (AIHubMix). GLM-4.6V is offered by 15 providers in the glm family, open weights, reasoning enabled. Z.AI's own rate is $0.9 per 1M output tokens.

Cheapest input
$0.137
Cheapest output
$0.411
Official (Z.AI)
$0.9
output per 1M tokens
Providers
15
offering this model
Spread
2.2×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

GLM-4.6V API prices by provider in USD per 1M tokens, updated Oct 4, 2026
AIHubMixopenCheapest$0.137Cheapest$0.411$0.0274—128k—
302.AIopen$0.145$0.43——128k—
ZenMuxopen$0.1456$0.4367$0.0291—200k14d agoout +4%
Eden AIopen$0.3$0.9$0.05—131k—
Kilo Gatewayopen$0.3$0.9$0.05—131k—
DevPass (LLM Gateway)open$0.3$0.9$0.05—131k—
LLM Gatewayopen$0.3$0.9$0.055—131k—
LLM Gatewayopen$0.3$0.9$0.05—128k—
NanoGPTopen$0.3$0.9$0.15—128k—
NovitaAIopen$0.3$0.9$0.055—131k—
OpenRouteropen$0.3$0.9$0.05—131k—
Tempr Gatewayopen$0.3$0.9——128k—
Z.AIopen$0.3$0.9——128k—
Zhipu AIopen$0.3$0.9——128k—

Subscription plans (not per-token)

Billed per month, not per token — never counted as the cheapest offer.

Zhipu AI Coding Planplan$0.3$0.9128k

GLM-4.6V pricing FAQ

What is the cheapest GLM-4.6V API?

As of Oct 4, 2026, AIHubMix has the lowest GLM-4.6V output price at $0.411 per 1M tokens, and AIHubMix has the lowest input price at $0.137 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does GLM-4.6V cost on Z.AI?

Z.AI charges $0.3 per 1M input tokens and $0.9 per 1M output tokens for GLM-4.6V.

How many providers offer GLM-4.6V?

15 providers list GLM-4.6V on Sovyron; 14 of them sell it at a metered per-token price.

Is GLM-4.6V free?

No provider in the Sovyron catalog lists a free tier for GLM-4.6V.

What is the context window of GLM-4.6V?

GLM-4.6V supports a 200k-token context window and up to 128k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—