Qwen/Qwen3-VL-30B-A3B-Thinking
As of Oct 4, 2026, the cheapest Qwen/Qwen3-VL-30B-A3B-Thinking API price is $0.2 per 1M input tokens (NovitaAI) and $1 per 1M output tokens (NovitaAI). Qwen/Qwen3-VL-30B-A3B-Thinking is offered by 3 providers in the qwen family, open weights, reasoning enabled.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| NovitaAIopen | Cheapest$0.2 | Cheapest$1 | — | — | 131k |
| SiliconFlow | $0.29 | $1 | — | — | 262k |
| SiliconFlow (China) | $0.29 | $1 | — | — | 262k |
Qwen/Qwen3-VL-30B-A3B-Thinking pricing FAQ
What is the cheapest Qwen/Qwen3-VL-30B-A3B-Thinking API?
As of Oct 4, 2026, NovitaAI has the lowest Qwen/Qwen3-VL-30B-A3B-Thinking output price at $1 per 1M tokens, and NovitaAI has the lowest input price at $0.2 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How many providers offer Qwen/Qwen3-VL-30B-A3B-Thinking?
3 providers list Qwen/Qwen3-VL-30B-A3B-Thinking on Sovyron; 3 of them sell it at a metered per-token price.
Is Qwen/Qwen3-VL-30B-A3B-Thinking free?
No provider in the Sovyron catalog lists a free tier for Qwen/Qwen3-VL-30B-A3B-Thinking.
What is the context window of Qwen/Qwen3-VL-30B-A3B-Thinking?
Qwen/Qwen3-VL-30B-A3B-Thinking supports a 262k-token context window and up to 262k output tokens.