Skip to content
Sovyron

Qwen/Qwen3-VL-8B-Instruct

qwen262k context262k max outputreleased Oct 15, 2025Open weightsTool calling

As of Oct 4, 2026, the cheapest Qwen/Qwen3-VL-8B-Instruct API price is $0.08 per 1M input tokens (NovitaAI) and $0.5 per 1M output tokens (NovitaAI). Qwen/Qwen3-VL-8B-Instruct is offered by 3 providers in the qwen family, open weights.

Cheapest input
$0.08
Cheapest output
$0.5
Official
—
output per 1M tokens
Providers
3
offering this model
Spread
1.4×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen/Qwen3-VL-8B-Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
NovitaAIopenCheapest$0.08Cheapest$0.5——131k
SiliconFlow$0.18$0.68——262k
SiliconFlow (China)$0.18$0.68——262k

Qwen/Qwen3-VL-8B-Instruct pricing FAQ

What is the cheapest Qwen/Qwen3-VL-8B-Instruct API?

As of Oct 4, 2026, NovitaAI has the lowest Qwen/Qwen3-VL-8B-Instruct output price at $0.5 per 1M tokens, and NovitaAI has the lowest input price at $0.08 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen/Qwen3-VL-8B-Instruct?

3 providers list Qwen/Qwen3-VL-8B-Instruct on Sovyron; 3 of them sell it at a metered per-token price.

Is Qwen/Qwen3-VL-8B-Instruct free?

No provider in the Sovyron catalog lists a free tier for Qwen/Qwen3-VL-8B-Instruct.

What is the context window of Qwen/Qwen3-VL-8B-Instruct?

Qwen/Qwen3-VL-8B-Instruct supports a 262k-token context window and up to 262k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—