Skip to content
Sovyron

Mistral Large

mistral-large262k context262k max outputreleased Nov 1, 2024Open weightsTool calling

As of Oct 4, 2026, the cheapest Mistral Large API price is $0.5 per 1M input tokens (Merge Gateway) and $1.5 per 1M output tokens (Merge Gateway). Mistral Large is offered by 11 providers in the mistral-large family, open weights. Mistral's own rate is $1.5 per 1M output tokens. One provider lists it as a free-tier offer.

Cheapest input
$0.5
Cheapest output
$1.5
Official (Mistral)
$1.5
output per 1M tokens
Providers
11
offering this model
Spread
8.0×
max / min output
Free offers
1
free-tier offer

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Mistral Large API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Merge GatewayopenCheapest$0.5Cheapest$1.5——262k
Mistralopen$0.5$1.5$0.05—262k
Tempr Gatewayopen$0.5$1.5——262k
Eden AIopen$2$6$0.2—262k
Helicone$2$6——128k
Kilo Gateway$2$6$0.2—128k
OpenRouter$2$6$0.2—128k
DevPass (LLM Gateway)open$4$12——128k
LLM Gatewayopen$4$12——128k
Kenarifreefreefree——262k

Mistral Large pricing FAQ

What is the cheapest Mistral Large API?

As of Oct 4, 2026, Merge Gateway has the lowest Mistral Large output price at $1.5 per 1M tokens, and Merge Gateway has the lowest input price at $0.5 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Mistral Large cost on Mistral?

Mistral charges $0.5 per 1M input tokens and $1.5 per 1M output tokens for Mistral Large.

How many providers offer Mistral Large?

11 providers list Mistral Large on Sovyron; 9 of them sell it at a metered per-token price.

Is Mistral Large free?

1 provider(s) list a free-tier offer: Kenari. Free tiers usually have rate limits.

What is the context window of Mistral Large?

Mistral Large supports a 262k-token context window and up to 262k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—