Skip to content
Sovyron

Gemini Flash-Lite Latest

gemini-flash-lite1.0M context66k max outputreleased Jul 21, 2026ReasoningTool calling

As of Oct 4, 2026, the cheapest Gemini Flash-Lite Latest API price is $0.25 per 1M input tokens (Vertex) and $1.5 per 1M output tokens (Vertex). Gemini Flash-Lite Latest is offered by 6 providers in the gemini-flash-lite family, reasoning enabled. Google's own rate is $2.5 per 1M output tokens.

Cheapest input
$0.25
Cheapest output
$1.5
Official (Google)
$2.5
output per 1M tokens
Providers
6
offering this model
Spread
1.7×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Gemini Flash-Lite Latest API prices by provider in USD per 1M tokens, updated Oct 4, 2026
VertexCheapest$0.25Cheapest$1.5$0.025—1.0M
Merge Gateway$0.25$1.5$0.025—1.0M
OrcaRouter$0.25$1.5$0.025—1.0M
Google$0.3$2.5$0.03—1.0M
NanoGPT$0.3$2.5$0.03$0.08331.0M
Tempr Gateway$0.3$2.5$0.03—1.0M

Gemini Flash-Lite Latest pricing FAQ

What is the cheapest Gemini Flash-Lite Latest API?

As of Oct 4, 2026, Vertex has the lowest Gemini Flash-Lite Latest output price at $1.5 per 1M tokens, and Vertex has the lowest input price at $0.25 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Gemini Flash-Lite Latest cost on Google?

Google charges $0.3 per 1M input tokens and $2.5 per 1M output tokens for Gemini Flash-Lite Latest.

How many providers offer Gemini Flash-Lite Latest?

6 providers list Gemini Flash-Lite Latest on Sovyron; 6 of them sell it at a metered per-token price.

Is Gemini Flash-Lite Latest free?

No provider in the Sovyron catalog lists a free tier for Gemini Flash-Lite Latest.

What is the context window of Gemini Flash-Lite Latest?

Gemini Flash-Lite Latest supports a 1.0M-token context window and up to 66k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—