Skip to content
Sovyron

Qwen/Qwen2.5-7B-Instruct

qwen33k context29k max outputreleased Sep 18, 2024Tool calling

As of Oct 4, 2026, the cheapest Qwen/Qwen2.5-7B-Instruct API price is $0.05 per 1M input tokens (SiliconFlow) and $0.05 per 1M output tokens (SiliconFlow). Qwen/Qwen2.5-7B-Instruct is offered by 3 providers in the qwen family.

Cheapest input
$0.05
Cheapest output
$0.05
Official
—
output per 1M tokens
Providers
3
offering this model
Spread
4.0×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen/Qwen2.5-7B-Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
SiliconFlowCheapest$0.05Cheapest$0.05——33k
SiliconFlow (China)$0.05$0.05——33k
Kilo Gateway$0.1$0.2——33k

Qwen/Qwen2.5-7B-Instruct pricing FAQ

What is the cheapest Qwen/Qwen2.5-7B-Instruct API?

As of Oct 4, 2026, SiliconFlow has the lowest Qwen/Qwen2.5-7B-Instruct output price at $0.05 per 1M tokens, and SiliconFlow has the lowest input price at $0.05 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Qwen/Qwen2.5-7B-Instruct?

3 providers list Qwen/Qwen2.5-7B-Instruct on Sovyron; 3 of them sell it at a metered per-token price.

Is Qwen/Qwen2.5-7B-Instruct free?

No provider in the Sovyron catalog lists a free tier for Qwen/Qwen2.5-7B-Instruct.

What is the context window of Qwen/Qwen2.5-7B-Instruct?

Qwen/Qwen2.5-7B-Instruct supports a 33k-token context window and up to 29k output tokens.