Skip to content
Sovyron

Qwen3.8 2.4T A95B

qwen1.0M context1.0M max outputreleased Aug 12, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3.8 2.4T A95B API price is $1.95 per 1M input tokens (IteraCompute) and $5.95 per 1M output tokens (IteraCompute). Qwen3.8 2.4T A95B is offered by 18 providers in the qwen family, open weights, reasoning enabled.

Cheapest input
$1.95
Cheapest output
$5.95
Official
—
output per 1M tokens
Providers
18
offering this model
Spread
1.1×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3.8 2.4T A95B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
IteraComputeopenCheapest$1.95Cheapest$5.95$0.2—970k
AIHubMixopen$2$6$0.5—262k
Deep Infraopen$2$6$0.2—262k
Eden AIopen$2$6$0.25$2.51.0M
Fireworks AIopen$2$6$0.25—262k
Charm Hyperopen$2$6$0.25—1.0M
Kilo Gatewayopen$2$6$0.25$2.51.0M
DevPass (LLM Gateway)open$2$6$0.25—1.0M
LLM Gatewayopen$2$6$0.2—262k
LLM Gatewayopen$2$6$0.25—1.0M
LLM Gatewayopen$2$6$0.25—1.0M
OpenRouteropen$2$6$0.25—1.0M
Requestyopen$2$6$0.2—262k
SiliconFlowopen$2$6$0.25—1.0M
Vercel AI Gatewayopen$2$6$0.25—262k
Cortecsopen$2.5$6$0.625—262k
Opperopen$2.5$6$0.63—262k
Requestyeuopen$2.5$6$0.63—1.0M
TensorXopen$2.5$6$0.63—262k
Hugging Faceopen$2.5$6.25——262k
Merge Gatewayopen$2.5$6.25$0.5—262k

Qwen3.8 2.4T A95B pricing FAQ

What is the cheapest Qwen3.8 2.4T A95B API?

As of Oct 4, 2026, IteraCompute has the lowest Qwen3.8 2.4T A95B output price at $5.95 per 1M tokens, and IteraCompute has the lowest input price at $1.95 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3.8 2.4T A95B?

18 providers list Qwen3.8 2.4T A95B on Sovyron; 21 of them sell it at a metered per-token price.

Is Qwen3.8 2.4T A95B free?

No provider in the Sovyron catalog lists a free tier for Qwen3.8 2.4T A95B.

What is the context window of Qwen3.8 2.4T A95B?

Qwen3.8 2.4T A95B supports a 1.0M-token context window and up to 1.0M output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—