Qwen3.5 9B
As of Oct 4, 2026, the cheapest Qwen3.5 9B API price is $0.04 per 1M input tokens (CrofAI) and $0.13 per 1M output tokens (EmpirioLabs AI). Qwen3.5 9B is offered by 19 providers in the qwen family, open weights, reasoning enabled. One provider lists it as a free-tier offer.
API pricing by provider
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| CrofAIopen | Cheapest$0.04 | $0.15 | $0.008 | — | 262k |
| NanoGPTopen | $0.05 | $0.15 | $0.025 | — | 256k |
| EmpirioLabs AIopen | $0.09 | Cheapest$0.13 | $0.045 | — | 262k |
| Merge Gatewayopen | $0.09 | $0.13 | — | — | 262k |
| Deep Infraopen | $0.1 | $0.15 | — | — | 262k |
| Kilo Gatewayopen | $0.1 | $0.15 | — | — | 256k |
| DevPass (LLM Gateway)open | $0.1 | $0.15 | — | — | 262k |
| LLM Gatewayopen | $0.1 | $0.15 | — | — | 262k |
| OpenRouteropen | $0.1 | $0.15 | — | — | 262k |
| Cortecsopen | $0.111 | $0.167 | — | — | 262k |
| OVHcloud AI Endpointsopen | $0.12 | $0.18 | — | — | 262k |
| TensorXopen | $0.15 | $0.2 | $0.0375 | $0.1875 | 262k |
| Mixlayeropen | $0.1 | $0.4 | — | — | 262k |
| Hugging Faceopen | $0.17 | $0.25 | — | — | 262k |
| Together AIopen | $0.17 | $0.25 | — | — | 262k |
| routing.runopen | $0.16 | $0.48 | — | — | 262k |
| Regolo AIopen | $0.15 | $0.6 | — | — | 262k |
| Pioneeropen | $0.3 | $0.3 | $0.3 | $0.3 | 33k |
| QVACfree | free | free | — | — | 33k |
Qwen3.5 9B pricing FAQ
What is the cheapest Qwen3.5 9B API?
As of Oct 4, 2026, EmpirioLabs AI has the lowest Qwen3.5 9B output price at $0.13 per 1M tokens, and CrofAI has the lowest input price at $0.04 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.
How many providers offer Qwen3.5 9B?
19 providers list Qwen3.5 9B on Sovyron; 18 of them sell it at a metered per-token price.
Is Qwen3.5 9B free?
1 provider(s) list a free-tier offer: QVAC. Free tiers usually have rate limits.
What is the context window of Qwen3.5 9B?
Qwen3.5 9B supports a 262k-token context window and up to 262k output tokens.