Skip to content
Sovyron

Mistral Medium

mistral-medium262k context262k max outputreleased Apr 29, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Mistral Medium API price is $0.4 per 1M input tokens (Merge Gateway) and $2 per 1M output tokens (Merge Gateway). Mistral Medium is offered by 6 providers in the mistral-medium family, open weights, reasoning enabled. Mistral's own rate is $7.5 per 1M output tokens.

Cheapest input
$0.4
Cheapest output
$2
Official (Mistral)
$7.5
output per 1M tokens
Providers
6
offering this model
Spread
3.8×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Mistral Medium API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Merge GatewayopenCheapest$0.4Cheapest$2——262k
Requestyopen$0.44$2.2$0.44—131k
Requestyopen$0.44$2.2$0.44—131k
Eden AIopen$1.5$7.5$0.15—262k
Mistralopen$1.5$7.5$0.15—262k
Tempr Gatewayopen$1.5$7.5——262k
Vercel AI Gateway$1.5$7.5$0.15—262k

Mistral Medium pricing FAQ

What is the cheapest Mistral Medium API?

As of Oct 4, 2026, Merge Gateway has the lowest Mistral Medium output price at $2 per 1M tokens, and Merge Gateway has the lowest input price at $0.4 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Mistral Medium cost on Mistral?

Mistral charges $1.5 per 1M input tokens and $7.5 per 1M output tokens for Mistral Medium.

How many providers offer Mistral Medium?

6 providers list Mistral Medium on Sovyron; 7 of them sell it at a metered per-token price.

Is Mistral Medium free?

No provider in the Sovyron catalog lists a free tier for Mistral Medium.

What is the context window of Mistral Medium?

Mistral Medium supports a 262k-token context window and up to 262k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—