DeepSeek · Pricing

DeepSeek V4 Flash pricing

Complete DeepSeek V4 Flash API cost reference: $0.14/M input, $0.28/M output, 1.3M token context window, $0.00/M cached reads. Verified against official sources.

Prices verified August 20, 2026
DeepSeek V4 Flash
Input / M
$0.14
Output / M
$0.28
Cache read
$0.00
Context
1.3M
Batch
−50%
Vision
Yes
Reasoning
Yes

What DeepSeek V4 Flash costs in practice

WorkloadCost / callCost @ 1K req/day
Light · 1K in / 300 out$0.0003$9.37/mo
Standard · 10K in / 2K out$0.0025$76.70/mo
Heavy · 100K in / 20K out$0.0252$767.03/mo

Estimates assume medium reasoning effort, no batch discounts and no prompt caching. Turn on prompt caching and the input side typically drops 50–90% — run your own mix in the free calculator preloaded with DeepSeek V4 Flash.

About this page

Prices were last verified on August 20, 2026 against DeepSeek's official pricing page. Providers reprice frequently — see our methodology for exactly how estimates are computed, and always confirm final rates with the provider before budgeting.

DeepSeek V4 Flash FAQ

How much does DeepSeek V4 Flash cost?

DeepSeek V4 Flash charges $0.14 per million input tokens and $0.28 per million output tokens, with cached input reads at $0.00/M. Batch mode cuts both rates by 50% for asynchronous workloads.

What is the context window of DeepSeek V4 Flash?

DeepSeek V4 Flash supports a context window of 1,310,720 tokens (1.3M) — roughly 983,040 English words or about 1,966 pages of text in a single request.

Does DeepSeek V4 Flash support images, audio or reasoning?

DeepSeek V4 Flash offers vision & image understanding, hidden reasoning tokens. Extremely cheap frontier-adjacent tier. 1M context, auto cache hits ~$0.0028/M.

How do I estimate my DeepSeek V4 Flash bill?

A typical workload of 10K input and 2K output tokens costs about $0.0025 per call with DeepSeek V4 Flash. At 1,000 requests per day that is roughly $76.70 per month. Use the free AITokenCalculator to run your exact numbers.