Skip to content
Sovyron

Qwen3.7 Flash

qwen1.0M context1.0M max outputreleased Jul 15, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3.7 Flash API price is $0.0282 per 1M input tokens (AIHubMix) and $0.1128 per 1M output tokens (AIHubMix). Qwen3.7 Flash is offered by 12 providers in the qwen family, reasoning enabled. Alibaba's own rate is $0.13 per 1M output tokens.

Cheapest input
$0.0282
Cheapest output
$0.1128
Official (Alibaba)
$0.13
output per 1M tokens
Providers
12
offering this model
Spread
7.1×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.7 Flash API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
AIHubMixCheapest$0.0282Cheapest$0.1128$0.0056$0.0352991k
Alibaba (China)†$0.0296$0.1185$0.003$0.0371.0M
Alibaba†$0.03$0.13$0.003$0.03751.0M
EmpirioLabs AI†$0.03$0.13$0.006—1.0M
Kilo Gateway$0.03$0.13$0.006$0.0381.0M
DevPass (LLM Gateway)$0.03$0.13$0.006$0.0375984k
LLM Gateway$0.03$0.13$0.006$0.0375984k
NanoGPT$0.03$0.13$0.006$0.038992k
OpenRouter†$0.03$0.13$0.006$0.0381.0M
OrcaRouter$0.03$0.13$0.006$0.0381.0M
CrossModel†$0.04$0.13$0.01$0.041.0M
Charm Hyper$0.2$0.8$0.04—1.0M

Qwen3.7 Flash pricing FAQ

What is the cheapest Qwen3.7 Flash API?

As of Oct 4, 2026, AIHubMix has the lowest Qwen3.7 Flash output price at $0.1128 per 1M tokens, and AIHubMix has the lowest input price at $0.0282 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3.7 Flash cost on Alibaba?

Alibaba charges $0.03 per 1M input tokens and $0.13 per 1M output tokens for Qwen3.7 Flash.

How many providers offer Qwen3.7 Flash?

12 providers list Qwen3.7 Flash on Sovyron; 12 of them sell it at a metered per-token price.

Is Qwen3.7 Flash free?

No provider in the Sovyron catalog lists a free tier for Qwen3.7 Flash.

What is the context window of Qwen3.7 Flash?

Qwen3.7 Flash supports a 1.0M-token context window and up to 1.0M output tokens.