Skip to content
Sovyron

Qwen3.5 9B

qwen262k context262k max outputreleased Feb 23, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3.5 9B API price is $0.04 per 1M input tokens (CrofAI) and $0.13 per 1M output tokens (EmpirioLabs AI). Qwen3.5 9B is offered by 19 providers in the qwen family, open weights, reasoning enabled. One provider lists it as a free-tier offer.

Cheapest input
$0.04
Cheapest output
$0.13
Official
—
output per 1M tokens
Providers
19
offering this model
Spread
4.6×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.5 9B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
CrofAIopenCheapest$0.04$0.15$0.008—262k
NanoGPTopen$0.05$0.15$0.025—256k
EmpirioLabs AIopen$0.09Cheapest$0.13$0.045—262k
Merge Gatewayopen$0.09$0.13——262k
Deep Infraopen$0.1$0.15——262k
Kilo Gatewayopen$0.1$0.15——256k
DevPass (LLM Gateway)open$0.1$0.15——262k
LLM Gatewayopen$0.1$0.15——262k
OpenRouteropen$0.1$0.15——262k
Cortecsopen$0.111$0.167——262k
OVHcloud AI Endpointsopen$0.12$0.18——262k
TensorXopen$0.15$0.2$0.0375$0.1875262k
Mixlayeropen$0.1$0.4——262k
Hugging Faceopen$0.17$0.25——262k
Together AIopen$0.17$0.25——262k
routing.runopen$0.16$0.48——262k
Regolo AIopen$0.15$0.6——262k
Pioneeropen$0.3$0.3$0.3$0.333k
QVACfreefreefree——33k

Qwen3.5 9B pricing FAQ

What is the cheapest Qwen3.5 9B API?

As of Oct 4, 2026, EmpirioLabs AI has the lowest Qwen3.5 9B output price at $0.13 per 1M tokens, and CrofAI has the lowest input price at $0.04 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3.5 9B?

19 providers list Qwen3.5 9B on Sovyron; 18 of them sell it at a metered per-token price.

Is Qwen3.5 9B free?

1 provider(s) list a free-tier offer: QVAC. Free tiers usually have rate limits.

What is the context window of Qwen3.5 9B?

Qwen3.5 9B supports a 262k-token context window and up to 262k output tokens.