# DeepSeek V4.1 Flash API prices

DeepSeek V4.1 Flash (deepseek-flash) — 1.1M context, 1.0M max output.
58 metered per-token offers from 55 provider(s), in USD per 1M tokens.
Prices as of 2026-10-04.

| Provider | Input $/1M | Output $/1M | Cache read $/1M | Context |
| --- | --- | --- | --- | --- |
| engy | $0.04 | $0.08 | $0.008 | 328k |
| AMD | $0.14 | $0.28 | $0.0028 | 1.0M |
| NanoGPT | $0.13 | $0.52 | $0.006 | 1.0M |
| DevPass (LLM Gateway) | $0.15 | $0.55 | $0.005 | 1.1M |
| 302.AI | $0.15 | $0.6 | $0.003 | 1.0M |
| DeepSeek | $0.15 | $0.6 | $0.003 | 1.0M |
| Eden AI | $0.15 | $0.6 | $0.015 | 1.0M |
| LLM Gateway | $0.15 | $0.6 | $0.003 | 1.1M |
| LLM Gateway | $0.15 | $0.6 | $0.01 | 1.0M |
| Merge Gateway | $0.15 | $0.6 | $0.003 | 1.0M |
| Neuralwatt | $0.15 | $0.6 | $0.015 | 1.0M |
| Ollama Cloud | $0.15 | $0.6 | $0.003 | 1.0M |
| OpenCode Go | $0.15 | $0.6 | $0.003 | 1.0M |
| Tempr Gateway | $0.15 | $0.6 | $0.003 | 1.0M |
| Umans AI | $0.15 | $0.6 | $0.028 | 1.0M |
| Vultr | $0.15 | $0.6 | — | 1.0M |
| ZenMux | $0.15 | $0.6 | $0.003 | 1.0M |
| AIHubMix | $0.155 | $0.62 | $0.0031 | 1.0M |
| above.dev | $0.165 | $0.66 | $0.0033 | 1.0M |
| Deep Infra | $0.2 | $0.6 | $0.006 | 1.0M |
| Eden AI | $0.2 | $0.6 | $0.006 | 1.0M |
| LLM Gateway | $0.2 | $0.6 | $0.006 | 1.0M |
| Cortecs | $0.201 | $0.6 | $0.004 | 1.0M |
| CoreWeave | $0.2 | $0.65 | $0.03 | 1.0M |
| LLM Gateway | $0.22 | $0.66 | $0.007 | 1.0M |
| Ofox | $0.21 | $0.84 | $0.0042 | 1.0M |
| ai& | $0.3 | $0.6 | $0.02 | 1.0M |
| Vancine | $0.24 | $0.96 | $0.0048 | 1.0M |
| Melious | $0.2318 | $1.1592 | $0.0116 | 1.0M |
| GreenPT | $0.2556 | $1.2778 | $0.0128 | 1.0M |
| Alibaba (China) | $0.2975 | $1.1902 | $0.0149 | 1.0M |
| Baseten | $0.3 | $1.2 | $0.03 | 1.0M |
| CrossModel | $0.3 | $1.2 | $0.006 | 1.0M |
| DigitalOcean | $0.3 | $1.2 | $0.006 | 1.0M |
| Eden AI | $0.3 | $1.2 | $0.3 | 1.0M |
| Eden AI | $0.3 | $1.2 | $0.006 | 1.0M |
| EmpirioLabs AI | $0.3 | $1.2 | $0.3 | 1.0M |
| Fireworks AI | $0.3 | $1.2 | $0.006 | 1.0M |
| Hugging Face | $0.3 | $1.2 | — | 1.0M |
| Kilo Gateway | $0.3 | $1.2 | $0.006 | 1.0M |
| LLM Gateway | $0.3 | $1.2 | $0.03 | 1.0M |
| LLM Gateway | $0.3 | $1.2 | $0.006 | 1.0M |
| LLM Gateway | $0.3 | $1.2 | $0.006 | 1.0M |
| Nebius Token Factory | $0.3 | $1.2 | $0.3 | 1.0M |
| OpenCode Zen | $0.3 | $1.2 | $0.006 | 1.0M |
| Together AI | $0.3 | $1.2 | $0.006 | 1.0M |
| Venice AI | $0.3 | $1.2 | $0.0075 | 1.0M |
| Vercel AI Gateway | $0.3 | $1.2 | $0.007 | 1.0M |
| Vivgrid | $0.31 | $1.23 | $0.01 | 1.0M |
| Charm Hyper | $0.33 | $1.31 | $0.03 | 1.0M |
| OpenRouter | $0.003 | $2.4 | $0.003 | 1.0M |
| Eden AI | $0.5 | $1.5 | $0.125 | 1.0M |
| Requesty | $0.5 | $1.5 | $0.05 | 1.0M |
| Requesty (eu) | $0.5 | $1.5 | $0.05 | 1.0M |
| Synthetic | $0.6 | $1.2 | $0.03 | 524k |
| TensorX | $0.5 | $1.5 | $0.13 | 1.0M |
| Tinfoil | $0.65 | $1.45 | $0.13 | 1.0M |
| Inco | $0.6 | $2.4 | — | 1.0M |

## Summary

- Cheapest input: $0.003 per 1M tokens (OpenRouter)
- Cheapest output: $0.08 per 1M tokens (engy)
- First-party: $0.6 per 1M output tokens (DeepSeek)
- Free offers: Alibaba Token Plan, Alibaba Token Plan (China), Kenari, NaN, Nvidia, SCNet Token Plan, Umans AI Coding Plan
- Subscription plans (not per-token): Alibaba Token Plan, Alibaba Token Plan (China), ClinePass, SCNet Token Plan, Umans AI Coding Plan
- Released: 2026-09-10
- Inputs: image, text

## FAQ

### What is the cheapest DeepSeek V4.1 Flash API?

As of Oct 4, 2026, engy has the lowest DeepSeek V4.1 Flash output price at $0.08 per 1M tokens, and OpenRouter has the lowest input price at $0.003 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

### How much does DeepSeek V4.1 Flash cost on DeepSeek?

DeepSeek charges $0.15 per 1M input tokens and $0.6 per 1M output tokens for DeepSeek V4.1 Flash.

### How many providers offer DeepSeek V4.1 Flash?

55 providers list DeepSeek V4.1 Flash on Sovyron; 58 of them sell it at a metered per-token price.

### Is DeepSeek V4.1 Flash free?

3 provider(s) list a free-tier offer: Kenari, NaN, Nvidia. Free tiers usually have rate limits.

### Is DeepSeek V4.1 Flash included in a subscription plan?

Yes. 4 flat-rate plan(s) include it; the cheapest is Qwen (Alibaba Cloud) Token Plan Lite at $8/month.

### What is the context window of DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash supports a 1.1M-token context window and up to 1.0M output tokens.

HTML page: https://sovyron.com/models/deepseek-v4-1-flash/
Full dataset: https://sovyron.com/data/catalog.json
