Skip to content
Sovyron

Qwen3 8B

qwen131k context41k max outputreleased Apr 1, 2025Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3 8B API price is $0.035 per 1M input tokens (NovitaAI) and $0.138 per 1M output tokens (NovitaAI). Qwen3 8B is offered by 5 providers in the qwen family, open weights, reasoning enabled. Alibaba's own rate is $0.7 per 1M output tokens.

Cheapest input
$0.035
Cheapest output
$0.138
Official (Alibaba)
$0.7
output per 1M tokens
Providers
5
offering this model
Spread
5.1×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3 8B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
NovitaAIopenCheapest$0.035Cheapest$0.138——128k
Alibaba (China)open$0.072$0.287——131k
Pioneeropen$0.2$0.2$0.2$0.241k
OpenRouteropen$0.117$0.455——131k
Alibabaopen$0.18$0.7——131k

Qwen3 8B pricing FAQ

What is the cheapest Qwen3 8B API?

As of Oct 4, 2026, NovitaAI has the lowest Qwen3 8B output price at $0.138 per 1M tokens, and NovitaAI has the lowest input price at $0.035 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3 8B cost on Alibaba?

Alibaba charges $0.18 per 1M input tokens and $0.7 per 1M output tokens for Qwen3 8B.

How many providers offer Qwen3 8B?

5 providers list Qwen3 8B on Sovyron; 5 of them sell it at a metered per-token price.

Is Qwen3 8B free?

No provider in the Sovyron catalog lists a free tier for Qwen3 8B.

What is the context window of Qwen3 8B?

Qwen3 8B supports a 131k-token context window and up to 41k output tokens.