Skip to content
Sovyron

Qwen3.8 Flash

qwen1.0M context131k max outputreleased Aug 26, 2026ReasoningTool calling4 subscription plans →

As of Oct 4, 2026, the cheapest Qwen3.8 Flash API price is $0.11 per 1M input tokens (Ofox) and $0.38 per 1M output tokens (Vancine). Qwen3.8 Flash is offered by 25 providers in the qwen family, reasoning enabled. Alibaba's own rate is $0.47 per 1M output tokens. One provider lists it as a free-tier offer. It is also included in 4 flat-rate subscription plans, Qwen (Alibaba Cloud) Token Plan Lite from $8/mo.

Cheapest input
$0.11
Cheapest output
$0.38
Official (Alibaba)
$0.47
output per 1M tokens
Providers
25
offering this model
Spread
1.5×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.8 Flash API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContextChanged
AIHubMix$0.1126$0.38$0.0141$0.17591.0M—
OfoxCheapest$0.11$0.39$0.011$0.141.0M22d agoout −17%
Deep Infra$0.113$0.382$0.0141—1.0M—
Vancine$0.12Cheapest$0.38$0.013—1.0M—
Alibaba (China)$0.1187$0.4007$0.0119$0.14841.0M—
CrossModel$0.13$0.43$0.016$0.21.0M—
NanoGPT$0.14$0.42$0.016$0.2992k—
Alibaba$0.15$0.47$0.016$0.21.0M—
Eden AI$0.15$0.47$0.016$0.21.0M—
Charm Hyper$0.15$0.47$0.016—1.0M—
Kilo Gateway$0.15$0.47$0.016$0.21.0M—
DevPass (LLM Gateway)$0.15$0.47$0.016$0.21.0M—
LLM Gateway$0.15$0.47$0.016$0.2984k—
LLM Gateway$0.15$0.47$0.016—1.0M—
Merge Gateway$0.15$0.47$0.016—977k—
Ofox$0.15$0.47$0.016$0.21.0M—
OpenCode Zen$0.15$0.47$0.016$0.21.0M—
OpenCode Go$0.15$0.47$0.016$0.21.0M—
OpenRouter$0.15$0.47$0.016$0.21.0M—
EmpirioLabs AI$0.16$0.47$0.16—1.0M—
GMI Cloud$0.16$0.47$0.016$0.21.0M—
Requesty$0.16$0.47$0.016$0.21.0M—
302.AI$0.18$0.564——1.0M—
NaNfreefreefree——262k—

Subscription plans (not per-token)

Billed per month, not per token — never counted as the cheapest offer.

Provider1M input1M outputContext
Alibaba Token Planplanfreefree1.0M
Alibaba Token Plan (China)planfreefree1.0M
SCNet Token Planplanfreefree1.0M

Qwen3.8 Flash pricing FAQ

What is the cheapest Qwen3.8 Flash API?

As of Oct 4, 2026, Vancine has the lowest Qwen3.8 Flash output price at $0.38 per 1M tokens, and Ofox has the lowest input price at $0.11 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3.8 Flash cost on Alibaba?

Alibaba charges $0.15 per 1M input tokens and $0.47 per 1M output tokens for Qwen3.8 Flash.

How many providers offer Qwen3.8 Flash?

25 providers list Qwen3.8 Flash on Sovyron; 23 of them sell it at a metered per-token price.

Is Qwen3.8 Flash free?

1 provider(s) list a free-tier offer: NaN. Free tiers usually have rate limits.

Is Qwen3.8 Flash included in a subscription plan?

Yes. 4 flat-rate plan(s) include it; the cheapest is Qwen (Alibaba Cloud) Token Plan Lite at $8/month.

What is the context window of Qwen3.8 Flash?

Qwen3.8 Flash supports a 1.0M-token context window and up to 131k output tokens.