Skip to content
Sovyron

Qwen/Qwen3.5-9B

qwen262k context262k max outputreleased Mar 3, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen/Qwen3.5-9B API price is $0.1 per 1M input tokens (SiliconFlow) and $0.15 per 1M output tokens (SiliconFlow). Qwen/Qwen3.5-9B is offered by 2 providers in the qwen family, open weights, reasoning enabled.

Cheapest input
$0.1
Cheapest output
$0.15
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
11.6×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen/Qwen3.5-9B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
SiliconFlowCheapest$0.1Cheapest$0.15——262k
SiliconFlow (China)open$0.22$1.74——262k

Qwen/Qwen3.5-9B pricing FAQ

What is the cheapest Qwen/Qwen3.5-9B API?

As of Oct 4, 2026, SiliconFlow has the lowest Qwen/Qwen3.5-9B output price at $0.15 per 1M tokens, and SiliconFlow has the lowest input price at $0.1 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen/Qwen3.5-9B?

2 providers list Qwen/Qwen3.5-9B on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen/Qwen3.5-9B free?

No provider in the Sovyron catalog lists a free tier for Qwen/Qwen3.5-9B.

What is the context window of Qwen/Qwen3.5-9B?

Qwen/Qwen3.5-9B supports a 262k-token context window and up to 262k output tokens.