DeepSeek API pricing in 2026: the budget benchmark

Prices verified August 21, 2026 ยท USD per 1M tokens

DeepSeek is the price floor most teams compare everything else against. Its two current models undercut Western flagship APIs by one to two orders of magnitude on list price โ€” and its cache-hit pricing is the most aggressive in the industry.

ModelInputCached inputOutputContext
DeepSeek V4 Pro$0.435$0.0036$0.87128K
DeepSeek V4 Flash$0.14$0.0028$0.28128K

Cache-hit input bills at under 1% of the standard input rate โ€” far deeper than the ~10% cached rate typical at OpenAI, Anthropic and Google.

How it compares to GPT, Claude and Gemini

ModelInputOutputOutput vs V4 Pro
DeepSeek V4 Pro$0.435$0.871ร—
Gemini 3 Flash$0.50$3.003.4ร—
Claude Sonnet 5$2.00$10.0011ร—
GPT-5.6 Terra$2.00$12.0014ร—
Claude Fable 5$10.00$50.0057ร—

On raw list price the gap is enormous. The honest caveats: DeepSeek's 128K context window is smaller than the 1M windows now standard on flagship APIs, there is no separate batch tier, and for the hardest reasoning tasks the frontier models still lead on quality. For chat, classification, summarization, extraction and most agent plumbing, many teams find V4-class quality more than sufficient.

The caching trick that changes the math

DeepSeek bills repeated prompt prefixes at $0.0036 per million tokens โ€” effectively free. A chatbot with a long system prompt and 80% cache-hit rate pays almost nothing for its instruction overhead. If your workload is prefix-heavy, DeepSeek's effective price drops even further below the sticker price. Model your exact mix โ€” cache slider included โ€” in the calculator, or see how it ranks in the cheapest APIs list.

โ†’ What would YOUR workload cost?
Enter your token mix once โ€” compare all 28 models side by side, with caching and batch discounts.
Advertisement

Frequently asked questions

How much does the DeepSeek API cost in 2026?
DeepSeek V4 Pro costs $0.435 input / $0.87 output per million tokens; DeepSeek V4 Flash costs $0.14 / $0.28 (August 2026). Cache-hit input is billed at under 1% of the standard rate.
Is DeepSeek cheaper than GPT or Claude?
Yes, by a wide margin on list price. DeepSeek V4 Pro's $0.87 output rate is roughly 14ร— cheaper than GPT-5.6 Terra ($12) and 11ร— cheaper than Claude Sonnet 5 ($10). Whether quality is sufficient depends on your task.
Does DeepSeek support prompt caching?
Yes โ€” DeepSeek's context caching is among the most aggressive in the industry: cache-hit input bills at $0.0036 per million tokens on V4 Pro, under 1% of the normal input price.
What is DeepSeek's context window?
Both DeepSeek V4 Pro and V4 Flash offer a 128K-token context window โ€” smaller than the 1M windows on GPT-5.6, Claude 5 and Gemini 3.
Does DeepSeek offer batch discounts?
DeepSeek does not publish a separate 50% batch tier the way OpenAI, Anthropic and Google do. Its list prices are already below most competitors' batch prices.