Skip to content
Sovyron

Llama 3.1 Nemotron 70B Instruct

nemotron131k context8k max outputreleased Apr 15, 2025Open weightsTool calling

As of Oct 4, 2026, the cheapest Llama 3.1 Nemotron 70B Instruct API price is $0.6 per 1M input tokens (Eden AI) and $0.6 per 1M output tokens (Eden AI). Llama 3.1 Nemotron 70B Instruct is offered by 2 providers in the nemotron family, open weights. One provider lists it as a free-tier offer.

Cheapest input
$0.6
Cheapest output
$0.6
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
—
max / min output
Free offers
1
free-tier offer

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Llama 3.1 Nemotron 70B Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Eden AIopenCheapest$0.6Cheapest$0.6——131k
Nvidiafreefreefree——128k

Llama 3.1 Nemotron 70B Instruct pricing FAQ

What is the cheapest Llama 3.1 Nemotron 70B Instruct API?

As of Oct 4, 2026, Eden AI has the lowest Llama 3.1 Nemotron 70B Instruct output price at $0.6 per 1M tokens, and Eden AI has the lowest input price at $0.6 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Llama 3.1 Nemotron 70B Instruct?

2 providers list Llama 3.1 Nemotron 70B Instruct on Sovyron; 1 of them sell it at a metered per-token price.

Is Llama 3.1 Nemotron 70B Instruct free?

1 provider(s) list a free-tier offer: Nvidia. Free tiers usually have rate limits.

What is the context window of Llama 3.1 Nemotron 70B Instruct?

Llama 3.1 Nemotron 70B Instruct supports a 131k-token context window and up to 8k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—