Skip to content
Sovyron

Qwen2.5-VL 72B Instruct

qwen131k context115k max outputreleased Sep 1, 2024Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen2.5-VL 72B Instruct API price is $0.8 per 1M input tokens (NovitaAI) and $0.8 per 1M output tokens (NovitaAI). Qwen2.5-VL 72B Instruct is offered by 6 providers in the qwen family, open weights, reasoning enabled. Alibaba's own rate is $8.4 per 1M output tokens.

Cheapest input
$0.8
Cheapest output
$0.8
Official (Alibaba)
$8.4
output per 1M tokens
Providers
6
offering this model
Spread
10.5×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen2.5-VL 72B Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
NovitaAIopenCheapest$0.8Cheapest$0.8——33k
OpenRouteropen$0.8$1$0.4—128k
OVHcloud AI Endpointsopen$1.01$1.01——33k
Cortecs$1.014$1.014——32k
Alibaba (China)open$2.294$6.881——131k
Alibabaopen$2.8$8.4——131k

Qwen2.5-VL 72B Instruct pricing FAQ

What is the cheapest Qwen2.5-VL 72B Instruct API?

As of Oct 4, 2026, NovitaAI has the lowest Qwen2.5-VL 72B Instruct output price at $0.8 per 1M tokens, and NovitaAI has the lowest input price at $0.8 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen2.5-VL 72B Instruct cost on Alibaba?

Alibaba charges $2.8 per 1M input tokens and $8.4 per 1M output tokens for Qwen2.5-VL 72B Instruct.

How many providers offer Qwen2.5-VL 72B Instruct?

6 providers list Qwen2.5-VL 72B Instruct on Sovyron; 6 of them sell it at a metered per-token price.

Is Qwen2.5-VL 72B Instruct free?

No provider in the Sovyron catalog lists a free tier for Qwen2.5-VL 72B Instruct.

What is the context window of Qwen2.5-VL 72B Instruct?

Qwen2.5-VL 72B Instruct supports a 131k-token context window and up to 115k output tokens.