Skip to content
Sovyron

Qwen3 235B A22B FP8

qwen41k context20k max outputreleased Apr 28, 2025Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Qwen3 235B A22B FP8 API price is $0.2 per 1M input tokens (DevPass (LLM Gateway)) and $0.8 per 1M output tokens (DevPass (LLM Gateway)). Qwen3 235B A22B FP8 is offered by 2 providers in the qwen family, open weights, reasoning enabled.

Cheapest input
$0.2
Cheapest output
$0.8
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
1.0×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3 235B A22B FP8 API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
DevPass (LLM Gateway)openCheapest$0.2Cheapest$0.8——41k
LLM Gateway$0.2$0.8——41k

Qwen3 235B A22B FP8 pricing FAQ

What is the cheapest Qwen3 235B A22B FP8 API?

As of Oct 4, 2026, DevPass (LLM Gateway) has the lowest Qwen3 235B A22B FP8 output price at $0.8 per 1M tokens, and DevPass (LLM Gateway) has the lowest input price at $0.2 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3 235B A22B FP8?

2 providers list Qwen3 235B A22B FP8 on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen3 235B A22B FP8 free?

No provider in the Sovyron catalog lists a free tier for Qwen3 235B A22B FP8.

What is the context window of Qwen3 235B A22B FP8?

Qwen3 235B A22B FP8 supports a 41k-token context window and up to 20k output tokens.