Skip to content
Sovyron

Qwen3.5 4B

qwen262k context33k max outputreleased Nov 1, 2025Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3.5 4B API price is $0.04 per 1M input tokens (EmpirioLabs AI) and $0.07 per 1M output tokens (EmpirioLabs AI). Qwen3.5 4B is offered by 2 providers in the qwen family, open weights, reasoning enabled. One provider lists it as a free-tier offer.

Cheapest input
$0.04
Cheapest output
$0.07
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
—
max / min output
Free offers
1
free-tier offer

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.5 4B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
EmpirioLabs AIopenCheapest$0.04Cheapest$0.07$0.02—262k
QVACfreefreefree——33k

Qwen3.5 4B pricing FAQ

What is the cheapest Qwen3.5 4B API?

As of Oct 4, 2026, EmpirioLabs AI has the lowest Qwen3.5 4B output price at $0.07 per 1M tokens, and EmpirioLabs AI has the lowest input price at $0.04 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3.5 4B?

2 providers list Qwen3.5 4B on Sovyron; 1 of them sell it at a metered per-token price.

Is Qwen3.5 4B free?

1 provider(s) list a free-tier offer: QVAC. Free tiers usually have rate limits.

What is the context window of Qwen3.5 4B?

Qwen3.5 4B supports a 262k-token context window and up to 33k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—