Skip to content
Sovyron

GLM-5.3-FlashX

glm-flash1.0M context131k max outputreleased Sep 18, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest GLM-5.3-FlashX API price is $0.15 per 1M input tokens (Ofox) and $0.5 per 1M output tokens (Ofox). GLM-5.3-FlashX is offered by 8 providers in the glm-flash family, open weights, reasoning enabled. Z.AI's own rate is $1.25 per 1M output tokens.

Cheapest input
$0.15
Cheapest output
$0.5
Official (Z.AI)
$1.25
output per 1M tokens
Providers
8
offering this model
Spread
2.5×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

GLM-5.3-FlashX API prices by provider in USD per 1M tokens, updated Oct 4, 2026
OfoxopenCheapest$0.15Cheapest$0.5$0.03—1.0M
Kilo Gateway$0.37$1.25$0.09—1.0M
OpenRouter$0.37$1.25$0.09—1.0M
Tempr Gatewayopen$0.37$1.25$0.075free1.0M
Vercel AI Gatewayopen$0.37$1.25$0.075—1.0M
Z.AIopen$0.37$1.25$0.075free1.0M
Zhipu AIopen$0.37$1.25$0.075free1.0M
ZenMuxopen$0.375$1.25$0.075—1.0M

GLM-5.3-FlashX pricing FAQ

What is the cheapest GLM-5.3-FlashX API?

As of Oct 4, 2026, Ofox has the lowest GLM-5.3-FlashX output price at $0.5 per 1M tokens, and Ofox has the lowest input price at $0.15 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does GLM-5.3-FlashX cost on Z.AI?

Z.AI charges $0.37 per 1M input tokens and $1.25 per 1M output tokens for GLM-5.3-FlashX.

How many providers offer GLM-5.3-FlashX?

8 providers list GLM-5.3-FlashX on Sovyron; 8 of them sell it at a metered per-token price.

Is GLM-5.3-FlashX free?

No provider in the Sovyron catalog lists a free tier for GLM-5.3-FlashX.

What is the context window of GLM-5.3-FlashX?

GLM-5.3-FlashX supports a 1.0M-token context window and up to 131k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—