Skip to content
Sovyron

Llama 4 Maverick 17B 128E Instruct

llama524k context8k max outputreleased Jan 15, 2025Open weightsTool calling

As of Oct 4, 2026, the cheapest Llama 4 Maverick 17B 128E Instruct API price is $0.15 per 1M input tokens (IO.NET) and $0.6 per 1M output tokens (IO.NET). Llama 4 Maverick 17B 128E Instruct is offered by 2 providers in the llama family, open weights.

Cheapest input
$0.15
Cheapest output
$0.6
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
1.9×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Llama 4 Maverick 17B 128E Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
IO.NETopenCheapest$0.15Cheapest$0.6$0.075$0.3430k
Vertexopen$0.35$1.15——524k

Llama 4 Maverick 17B 128E Instruct pricing FAQ

What is the cheapest Llama 4 Maverick 17B 128E Instruct API?

As of Oct 4, 2026, IO.NET has the lowest Llama 4 Maverick 17B 128E Instruct output price at $0.6 per 1M tokens, and IO.NET has the lowest input price at $0.15 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Llama 4 Maverick 17B 128E Instruct?

2 providers list Llama 4 Maverick 17B 128E Instruct on Sovyron; 2 of them sell it at a metered per-token price.

Is Llama 4 Maverick 17B 128E Instruct free?

No provider in the Sovyron catalog lists a free tier for Llama 4 Maverick 17B 128E Instruct.

What is the context window of Llama 4 Maverick 17B 128E Instruct?

Llama 4 Maverick 17B 128E Instruct supports a 524k-token context window and up to 8k output tokens.