Skip to content
Sovyron

Qwen Flash

qwen1.0M context250k max outputreleased Jul 28, 2025ReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen Flash API price is $0.022 per 1M input tokens (Alibaba (China)) and $0.216 per 1M output tokens (Alibaba (China)). Qwen Flash is offered by 7 providers in the qwen family, reasoning enabled. Alibaba's own rate is $0.4 per 1M output tokens.

Cheapest input
$0.022
Cheapest output
$0.216
Official (Alibaba)
$0.4
output per 1M tokens
Providers
7
offering this model
Spread
1.9×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen Flash API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Alibaba (China)Cheapest$0.022Cheapest$0.216——1.0M
Merge Gateway$0.022$0.216$0.0044—1.0M
Ofox$0.022$0.22$0.0043$0.0271.0M
Ofox$0.022$0.22$0.0043$0.0271.0M
Alibaba$0.05$0.4——1.0M
DevPass (LLM Gateway)$0.05$0.4$0.01$0.06251.0M
LLM Gateway$0.05$0.4$0.01$0.06251.0M
LLMTR$0.05$0.4——1.0M

Qwen Flash pricing FAQ

What is the cheapest Qwen Flash API?

As of Oct 4, 2026, Alibaba (China) has the lowest Qwen Flash output price at $0.216 per 1M tokens, and Alibaba (China) has the lowest input price at $0.022 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen Flash cost on Alibaba?

Alibaba charges $0.05 per 1M input tokens and $0.4 per 1M output tokens for Qwen Flash.

How many providers offer Qwen Flash?

7 providers list Qwen Flash on Sovyron; 8 of them sell it at a metered per-token price.

Is Qwen Flash free?

No provider in the Sovyron catalog lists a free tier for Qwen Flash.

What is the context window of Qwen Flash?

Qwen Flash supports a 1.0M-token context window and up to 250k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—