Qwen3.8 Flash
As of Oct 4, 2026, the cheapest Qwen3.8 Flash API price is $0.11 per 1M input tokens (Ofox) and $0.38 per 1M output tokens (Vancine). Qwen3.8 Flash is offered by 25 providers in the qwen family, reasoning enabled. Alibaba's own rate is $0.47 per 1M output tokens. One provider lists it as a free-tier offer. It is also included in 4 flat-rate subscription plans, Qwen (Alibaba Cloud) Token Plan Lite from $8/mo.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context | Changed |
|---|---|---|---|---|---|---|
| AIHubMix | $0.1126 | $0.38 | $0.0141 | $0.1759 | 1.0M | — |
| Ofox | Cheapest$0.11 | $0.39 | $0.011 | $0.14 | 1.0M | 22d agoout −17% |
| Deep Infra | $0.113 | $0.382 | $0.0141 | — | 1.0M | — |
| Vancine | $0.12 | Cheapest$0.38 | $0.013 | — | 1.0M | — |
| Alibaba (China) | $0.1187 | $0.4007 | $0.0119 | $0.1484 | 1.0M | — |
| CrossModel | $0.13 | $0.43 | $0.016 | $0.2 | 1.0M | — |
| NanoGPT | $0.14 | $0.42 | $0.016 | $0.2 | 992k | — |
| Alibaba | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| Eden AI | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| Charm Hyper | $0.15 | $0.47 | $0.016 | — | 1.0M | — |
| Kilo Gateway | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| DevPass (LLM Gateway) | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| LLM Gateway | $0.15 | $0.47 | $0.016 | $0.2 | 984k | — |
| LLM Gateway | $0.15 | $0.47 | $0.016 | — | 1.0M | — |
| Merge Gateway | $0.15 | $0.47 | $0.016 | — | 977k | — |
| Ofox | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| OpenCode Zen | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| OpenCode Go | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| OpenRouter | $0.15 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| EmpirioLabs AI | $0.16 | $0.47 | $0.16 | — | 1.0M | — |
| GMI Cloud | $0.16 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| Requesty | $0.16 | $0.47 | $0.016 | $0.2 | 1.0M | — |
| 302.AI | $0.18 | $0.564 | — | — | 1.0M | — |
| NaNfree | free | free | — | — | 262k | — |
Subscription plans (not per-token)
Billed per month, not per token — never counted as the cheapest offer.
| Provider | 1M input | 1M output | Context |
|---|---|---|---|
| Alibaba Token Planplan | free | free | 1.0M |
| Alibaba Token Plan (China)plan | free | free | 1.0M |
| SCNet Token Planplan | free | free | 1.0M |
Qwen3.8 Flash pricing FAQ
What is the cheapest Qwen3.8 Flash API?
As of Oct 4, 2026, Vancine has the lowest Qwen3.8 Flash output price at $0.38 per 1M tokens, and Ofox has the lowest input price at $0.11 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How much does Qwen3.8 Flash cost on Alibaba?
Alibaba charges $0.15 per 1M input tokens and $0.47 per 1M output tokens for Qwen3.8 Flash.
How many providers offer Qwen3.8 Flash?
25 providers list Qwen3.8 Flash on Sovyron; 23 of them sell it at a metered per-token price.
Is Qwen3.8 Flash free?
1 provider(s) list a free-tier offer: NaN. Free tiers usually have rate limits.
Is Qwen3.8 Flash included in a subscription plan?
Yes. 4 flat-rate plan(s) include it; the cheapest is Qwen (Alibaba Cloud) Token Plan Lite at $8/month.
What is the context window of Qwen3.8 Flash?
Qwen3.8 Flash supports a 1.0M-token context window and up to 131k output tokens.