Skip to content
Sovyron

Gemini 3.5 Flash Thinking

gemini-flash1.0M context66k max outputreleased May 19, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest Gemini 3.5 Flash Thinking API price is $1.5 per 1M input tokens (302.AI) and $9 per 1M output tokens (302.AI). Gemini 3.5 Flash Thinking is offered by 2 providers in the gemini-flash family, reasoning enabled.

Cheapest input
$1.5
Cheapest output
$9
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
1.0×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Gemini 3.5 Flash Thinking API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
302.AICheapest$1.5Cheapest$9——1.0M
NanoGPT$1.5$9$0.15$0.08331.0M

Gemini 3.5 Flash Thinking pricing FAQ

What is the cheapest Gemini 3.5 Flash Thinking API?

As of Oct 4, 2026, 302.AI has the lowest Gemini 3.5 Flash Thinking output price at $9 per 1M tokens, and 302.AI has the lowest input price at $1.5 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer Gemini 3.5 Flash Thinking?

2 providers list Gemini 3.5 Flash Thinking on Sovyron; 2 of them sell it at a metered per-token price.

Is Gemini 3.5 Flash Thinking free?

No provider in the Sovyron catalog lists a free tier for Gemini 3.5 Flash Thinking.

What is the context window of Gemini 3.5 Flash Thinking?

Gemini 3.5 Flash Thinking supports a 1.0M-token context window and up to 66k output tokens.