Qwen2.5-VL 72B Instruct
As of Oct 4, 2026, the cheapest Qwen2.5-VL 72B Instruct API price is $0.8 per 1M input tokens (NovitaAI) and $0.8 per 1M output tokens (NovitaAI). Qwen2.5-VL 72B Instruct is offered by 6 providers in the qwen family, open weights, reasoning enabled. Alibaba's own rate is $8.4 per 1M output tokens.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| NovitaAIopen | Cheapest$0.8 | Cheapest$0.8 | — | — | 33k |
| OpenRouteropen | $0.8 | $1 | $0.4 | — | 128k |
| OVHcloud AI Endpointsopen | $1.01 | $1.01 | — | — | 33k |
| Cortecs | $1.014 | $1.014 | — | — | 32k |
| Alibaba (China)open | $2.294 | $6.881 | — | — | 131k |
| Alibabaopen | $2.8 | $8.4 | — | — | 131k |
Qwen2.5-VL 72B Instruct pricing FAQ
What is the cheapest Qwen2.5-VL 72B Instruct API?
As of Oct 4, 2026, NovitaAI has the lowest Qwen2.5-VL 72B Instruct output price at $0.8 per 1M tokens, and NovitaAI has the lowest input price at $0.8 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How much does Qwen2.5-VL 72B Instruct cost on Alibaba?
Alibaba charges $2.8 per 1M input tokens and $8.4 per 1M output tokens for Qwen2.5-VL 72B Instruct.
How many providers offer Qwen2.5-VL 72B Instruct?
6 providers list Qwen2.5-VL 72B Instruct on Sovyron; 6 of them sell it at a metered per-token price.
Is Qwen2.5-VL 72B Instruct free?
No provider in the Sovyron catalog lists a free tier for Qwen2.5-VL 72B Instruct.
What is the context window of Qwen2.5-VL 72B Instruct?
Qwen2.5-VL 72B Instruct supports a 131k-token context window and up to 115k output tokens.