Skip to content
Sovyron

Magnum v4 72B

llama33k context8k max outputreleased Oct 22, 2024Open weights

As of Oct 4, 2026, the cheapest Magnum v4 72B API price is $2.006 per 1M input tokens (NanoGPT) and $2.992 per 1M output tokens (NanoGPT). Magnum v4 72B is offered by 3 providers in the llama family, open weights.

Cheapest input
$2.006
Cheapest output
$2.992
Official
—
output per 1M tokens
Providers
3
offering this model
Spread
1.7×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Magnum v4 72B API prices by provider in USD per 1M tokens, updated Oct 4, 2026
NanoGPTopenCheapest$2.006Cheapest$2.992$1.003—16k
Kilo Gateway$2.5$5——33k
OpenRouteropen$2.5$5——33k

Magnum v4 72B pricing FAQ

What is the cheapest Magnum v4 72B API?

As of Oct 4, 2026, NanoGPT has the lowest Magnum v4 72B output price at $2.992 per 1M tokens, and NanoGPT has the lowest input price at $2.006 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Magnum v4 72B?

3 providers list Magnum v4 72B on Sovyron; 3 of them sell it at a metered per-token price.

Is Magnum v4 72B free?

No provider in the Sovyron catalog lists a free tier for Magnum v4 72B.

What is the context window of Magnum v4 72B?

Magnum v4 72B supports a 33k-token context window and up to 8k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—