Skip to content
Sovyron

Qwen3 VL Flash

qwen262k context33k max outputreleased Oct 9, 2025Tool calling

As of Oct 4, 2026, the cheapest Qwen3 VL Flash API price is $0.05 per 1M input tokens (DevPass (LLM Gateway)) and $0.4 per 1M output tokens (DevPass (LLM Gateway)). Qwen3 VL Flash is offered by 2 providers in the qwen family.

Cheapest input
$0.05
Cheapest output
$0.4
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
1.0×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3 VL Flash API prices by provider in USD per 1M tokens, updated Oct 4, 2026
DevPass (LLM Gateway)Cheapest$0.05Cheapest$0.4$0.01—262k
LLM Gateway$0.05$0.4$0.01—262k

Qwen3 VL Flash pricing FAQ

What is the cheapest Qwen3 VL Flash API?

As of Oct 4, 2026, DevPass (LLM Gateway) has the lowest Qwen3 VL Flash output price at $0.4 per 1M tokens, and DevPass (LLM Gateway) has the lowest input price at $0.05 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3 VL Flash?

2 providers list Qwen3 VL Flash on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen3 VL Flash free?

No provider in the Sovyron catalog lists a free tier for Qwen3 VL Flash.

What is the context window of Qwen3 VL Flash?

Qwen3 VL Flash supports a 262k-token context window and up to 33k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—