Skip to content
Sovyron

Qwen3.5 Flash

qwen1.0M context1.0M max outputreleased Feb 23, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3.5 Flash API price is $0.0282 per 1M input tokens (AIHubMix) and $0.26 per 1M output tokens (OpenRouter). Qwen3.5 Flash is offered by 10 providers in the qwen family, reasoning enabled. Alibaba's own rate is $0.4 per 1M output tokens.

Cheapest input
$0.0282
Cheapest output
$0.26
Official (Alibaba)
$0.4
output per 1M tokens
Providers
10
offering this model
Spread
1.5×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.5 Flash API prices by provider in USD per 1M tokens, updated Oct 4, 2026
AIHubMixCheapest†$0.0282$0.282$0.0028$0.03521.0M—
Alibaba (China)†$0.029$0.287——1.0M12d agoout −72.2%
Merge Gateway$0.029$0.287$0.0058—1.0M—
OpenRouter$0.065Cheapest$0.26——1.0M—
EmpirioLabs AI$0.09$0.368$0.09—1.0M—
Alibaba$0.1$0.4$0.01$0.1251.0M—
NanoGPT$0.1$0.4$0.05—992k—
Ofox$0.1$0.4$0.01$0.1251.0M—
Ofox$0.1$0.4$0.01$0.1251.0M—
OrcaRouter$0.1$0.4——1.0M—
ZenMux$0.1$0.4——1.0M—

Qwen3.5 Flash pricing FAQ

What is the cheapest Qwen3.5 Flash API?

As of Oct 4, 2026, OpenRouter has the lowest Qwen3.5 Flash output price at $0.26 per 1M tokens, and AIHubMix has the lowest input price at $0.0282 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3.5 Flash cost on Alibaba?

Alibaba charges $0.1 per 1M input tokens and $0.4 per 1M output tokens for Qwen3.5 Flash.

How many providers offer Qwen3.5 Flash?

10 providers list Qwen3.5 Flash on Sovyron; 11 of them sell it at a metered per-token price.

Is Qwen3.5 Flash free?

No provider in the Sovyron catalog lists a free tier for Qwen3.5 Flash.

What is the context window of Qwen3.5 Flash?

Qwen3.5 Flash supports a 1.0M-token context window and up to 1.0M output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—