Skip to content
Sovyron

DeepSeek V4 Flash 0731 Fast

deepseek-flash1.0M context384k max outputreleased Jul 31, 2026Open weightsReasoningTool calling

As of Oct 4, 2026, the cheapest DeepSeek V4 Flash 0731 Fast API price is $0.28 per 1M input tokens (AIHubMix) and $0.7 per 1M output tokens (Venice AI). DeepSeek V4 Flash 0731 Fast is offered by 2 providers in the deepseek-flash family, open weights, reasoning enabled.

Cheapest input
$0.28
Cheapest output
$0.7
Official
—
output per 1M tokens
Providers
2
offering this model
Spread
2.0×
max / min output
Free offers
0
free-tier offers

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

DeepSeek V4 Flash 0731 Fast API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Venice AIopen$0.35Cheapest$0.7$0.0875—1.0M
AIHubMixopenCheapest$0.28$1.4$0.07—1.0M

DeepSeek V4 Flash 0731 Fast pricing FAQ

What is the cheapest DeepSeek V4 Flash 0731 Fast API?

As of Oct 4, 2026, Venice AI has the lowest DeepSeek V4 Flash 0731 Fast output price at $0.7 per 1M tokens, and AIHubMix has the lowest input price at $0.28 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How many providers offer DeepSeek V4 Flash 0731 Fast?

2 providers list DeepSeek V4 Flash 0731 Fast on Sovyron; 2 of them sell it at a metered per-token price.

Is DeepSeek V4 Flash 0731 Fast free?

No provider in the Sovyron catalog lists a free tier for DeepSeek V4 Flash 0731 Fast.

What is the context window of DeepSeek V4 Flash 0731 Fast?

DeepSeek V4 Flash 0731 Fast supports a 1.0M-token context window and up to 384k output tokens.

Token cost calculator

What a month costs at one provider's published rates. Enter the tokens you expect to send and receive; cached tokens use the cached-input price where the provider lists one, and count as ordinary input where it does not.

Estimated

—