Skip to content
Sovyron

Qwen3 Embedding 8B

qwen41k context33k max outputreleased Jun 5, 2025Open weights

As of Oct 4, 2026, the cheapest Qwen3 Embedding 8B API price is $0.1 per 1M input tokens (Regolo AI) and $0.1 per 1M output tokens (Regolo AI). Qwen3 Embedding 8B is offered by 6 providers in the qwen family, open weights. One provider lists it as a free-tier offer.

Cheapest input
$0.1
Cheapest output
$0.1
Official
—
output per 1M tokens
Providers
6
offering this model
Spread
1.1×
max / min output
Free offers
1
free-tier offer

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3 Embedding 8B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Regolo AIopenCheapest$0.1Cheapest$0.1——33k
evrocopen$0.115$0.115——41k
InferXfreefreefree——33k

Qwen3 Embedding 8B pricing FAQ

What is the cheapest Qwen3 Embedding 8B API?

As of Oct 4, 2026, Regolo AI has the lowest Qwen3 Embedding 8B output price at $0.1 per 1M tokens, and Regolo AI has the lowest input price at $0.1 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen3 Embedding 8B?

6 providers list Qwen3 Embedding 8B on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen3 Embedding 8B free?

1 provider(s) list a free-tier offer: InferX. Free tiers usually have rate limits.

What is the context window of Qwen3 Embedding 8B?

Qwen3 Embedding 8B supports a 41k-token context window and up to 33k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—