DeepSeek · Efficient · Open weights
DeepSeek V4 Flash
The cheap DeepSeek tier with a 1M-token context window, among the lowest cost-per-token options with a window that large.
What that means
A worked example
Per-token pricing is hard to feel. Take a moderately chatty production workload: 100 calls a day, roughly 10,000 tokens in and 2,000 out each time. On DeepSeek V4 Flash that comes to about:
$6 / month
Before caching, batching or volume discounts, all of which move this number a long way. Illustrative only.
Strong at
- -Lowest cost at 1M context
- -Throughput
Typical use
- -Bulk document processing
- -Cheap long-context tasks
Same tier
What else to look at
| Model | Developer | Context | In / 1M | Out / 1M |
|---|---|---|---|---|
| Claude Haiku 4.5 | Anthropic | 200K | $1 | $5 |
| GPT-5.6 Luna | OpenAI | 1M | $0.20 | $1.20 |
| DeepSeek V4 Pro | DeepSeek | 1M | $0.43 | $0.87 |
| Qwen 3.7 Flash | Alibaba | - | $0.03 | $0.13 |
| Qwen 3.6 27B | Alibaba | - | - | - |
Verified 2026-08-07.