Skip to content
Sovyron

Gemma 4 31B Instruct

gemma256k context8k max outputreleased Apr 2, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest Gemma 4 31B Instruct API price is $0.12 per 1M input tokens (Venice AI) and $0.36 per 1M output tokens (Venice AI). Gemma 4 31B Instruct is offered by 2 providers in the gemma family, open weights, reasoning enabled.

Cheapest input
$0.12
Cheapest output
$0.36
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
1.5×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Gemma 4 31B Instruct API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Venice AIopenCheapest$0.12Cheapest$0.36$0.09—256k
Berget.AIopen$0.275$0.55——128k

Gemma 4 31B Instruct pricing FAQ

What is the cheapest Gemma 4 31B Instruct API?

As of Oct 4, 2026, Venice AI has the lowest Gemma 4 31B Instruct output price at $0.36 per 1M tokens, and Venice AI has the lowest input price at $0.12 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Gemma 4 31B Instruct?

2 providers list Gemma 4 31B Instruct on Sovyron; 2 of them sell it at a metered per-token price.

Is Gemma 4 31B Instruct free?

No provider in the Sovyron catalog lists a free tier for Gemma 4 31B Instruct.

What is the context window of Gemma 4 31B Instruct?

Gemma 4 31B Instruct supports a 256k-token context window and up to 8k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—