Skip to content

DeepSeek · Efficient · Open weights

DeepSeek V4 Flash

The cheap DeepSeek tier with a 1M-token context window, among the lowest cost-per-token options with a window that large.

What that means

A worked example

Per-token pricing is hard to feel. Take a moderately chatty production workload: 100 calls a day, roughly 10,000 tokens in and 2,000 out each time. On DeepSeek V4 Flash that comes to about:

$6 / month

Before caching, batching or volume discounts, all of which move this number a long way. Illustrative only.

Strong at

  • -Lowest cost at 1M context
  • -Throughput

Typical use

  • -Bulk document processing
  • -Cheap long-context tasks

Same tier

What else to look at

Full table →
ModelDeveloperContextIn / 1MOut / 1M
Claude Haiku 4.5Anthropic200K$1$5
GPT-5.6 LunaOpenAI1M$0.20$1.20
DeepSeek V4 ProDeepSeek1M$0.43$0.87
Qwen 3.7 FlashAlibaba-$0.03$0.13
Qwen 3.6 27BAlibaba---

Verified 2026-08-07.