Skip to content
Sovyron

Qwen2.5-VL 7B Instruct

qwen131k context8k max outputreleased Sep 1, 2024Open weightsTool calling

As of Oct 4, 2026, the cheapest Qwen2.5-VL 7B Instruct API price is $0.287 per 1M input tokens (Alibaba (China)) and $0.717 per 1M output tokens (Alibaba (China)). Qwen2.5-VL 7B Instruct is offered by 2 providers in the qwen family, open weights. Alibaba's own rate is $1.05 per 1M output tokens.

Cheapest input
$0.287
Cheapest output
$0.717
Official (Alibaba)
$1.05
output per 1M tokens
Providers
2
offering this model
Spread
1.5×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen2.5-VL 7B Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
Alibaba (China)openCheapest$0.287Cheapest$0.717——131k
Alibabaopen$0.35$1.05——131k

Qwen2.5-VL 7B Instruct pricing FAQ

What is the cheapest Qwen2.5-VL 7B Instruct API?

As of Oct 4, 2026, Alibaba (China) has the lowest Qwen2.5-VL 7B Instruct output price at $0.717 per 1M tokens, and Alibaba (China) has the lowest input price at $0.287 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen2.5-VL 7B Instruct cost on Alibaba?

Alibaba charges $0.35 per 1M input tokens and $1.05 per 1M output tokens for Qwen2.5-VL 7B Instruct.

How many providers offer Qwen2.5-VL 7B Instruct?

2 providers list Qwen2.5-VL 7B Instruct on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen2.5-VL 7B Instruct free?

No provider in the Sovyron catalog lists a free tier for Qwen2.5-VL 7B Instruct.

What is the context window of Qwen2.5-VL 7B Instruct?

Qwen2.5-VL 7B Instruct supports a 131k-token context window and up to 8k output tokens.