Skip to content
Sovyron

GLM 5.3 Fast

glm1.0M context262k max outputreleased Aug 14, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest GLM 5.3 Fast API price is $2.1 per 1M input tokens (Baseten) and $6.6 per 1M output tokens (Baseten). GLM 5.3 Fast is offered by 4 providers in the glm family, open weights, reasoning enabled.

Cheapest input
$2.1
Cheapest output
$6.6
Official
—
output per 1M tokens
Providers
4
offering this model
Spread
1.3×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

GLM 5.3 Fast API prices by provider in USD per 1M tokens, updated Oct 4, 2026
BasetenopenCheapest$2.1Cheapest$6.6——1.0M
Fireworks AIopen$2.1$6.6$0.39—1.0M
Fireworks AIopen$2.1$6.6$0.39—1.0M
Vercel AI Gatewayopen$2.1$6.6$0.21—1.0M
Incoopen$2.8$8.8——1.0M

GLM 5.3 Fast pricing FAQ

What is the cheapest GLM 5.3 Fast API?

As of Oct 4, 2026, Baseten has the lowest GLM 5.3 Fast output price at $6.6 per 1M tokens, and Baseten has the lowest input price at $2.1 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer GLM 5.3 Fast?

4 providers list GLM 5.3 Fast on Sovyron; 5 of them sell it at a metered per-token price.

Is GLM 5.3 Fast free?

No provider in the Sovyron catalog lists a free tier for GLM 5.3 Fast.

What is the context window of GLM 5.3 Fast?

GLM 5.3 Fast supports a 1.0M-token context window and up to 262k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—