Skip to content
Sovyron

Qwen3-VL Plus

qwen262k context66k max outputreleased Sep 23, 2025ReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3-VL Plus API price is $0.137 per 1M input tokens (AIHubMix) and $1.37 per 1M output tokens (AIHubMix). Qwen3-VL Plus is offered by 8 providers in the qwen family, reasoning enabled. Alibaba's own rate is $1.6 per 1M output tokens. One provider lists it as a free-tier offer.

Cheapest input
$0.137
Cheapest output
$1.37
Official (Alibaba)
$1.6
output per 1M tokens
Providers
8
offering this model
Spread
1.2×
max / min output
Free offers
1
free-tier offer

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3-VL Plus API prices by provider in USD per 1M tokens, updated Oct 4, 2026
AIHubMixCheapest†$0.137Cheapest$1.37free—262k
Merge Gateway$0.143$1.434$0.0286—262k
Alibaba (China)$0.1434$1.4335——262k
Alibaba$0.2$1.6——262k
DevPass (LLM Gateway)$0.2$1.6$0.04$0.25262k
LLM Gateway$0.2$1.6$0.04$0.25262k
LLMTR$0.2$1.6——256k
iFlowfreefreefree——256k

Qwen3-VL Plus pricing FAQ

What is the cheapest Qwen3-VL Plus API?

As of Oct 4, 2026, AIHubMix has the lowest Qwen3-VL Plus output price at $1.37 per 1M tokens, and AIHubMix has the lowest input price at $0.137 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3-VL Plus cost on Alibaba?

Alibaba charges $0.2 per 1M input tokens and $1.6 per 1M output tokens for Qwen3-VL Plus.

How many providers offer Qwen3-VL Plus?

8 providers list Qwen3-VL Plus on Sovyron; 7 of them sell it at a metered per-token price.

Is Qwen3-VL Plus free?

1 provider(s) list a free-tier offer: iFlow. Free tiers usually have rate limits.

What is the context window of Qwen3-VL Plus?

Qwen3-VL Plus supports a 262k-token context window and up to 66k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—