DeepSeek · Efficient · Open weights
DeepSeek V4 Pro
DeepSeek's flagship, priced roughly an order of magnitude below the Western frontier tiers. The main reason teams look at it is cost per token on reasoning-heavy work.
What that means
A worked example
Per-token pricing is hard to feel. Take a moderately chatty production workload: 100 calls a day, roughly 10,000 tokens in and 2,000 out each time. On DeepSeek V4 Pro that comes to about:
$18 / month
Before caching, batching or volume discounts, all of which move this number a long way. Illustrative only.
Strong at
- -Very low cost
- -Reasoning
- -Open weights
Typical use
- -High-volume reasoning
- -Self-hosted deployments
Watch out for
- -Consider data residency and governance before sending regulated data
Same tier
What else to look at
| Model | Developer | Context | In / 1M | Out / 1M |
|---|---|---|---|---|
| Claude Haiku 4.5 | Anthropic | 200K | $1 | $5 |
| GPT-5.6 Luna | OpenAI | 1M | $0.20 | $1.20 |
| DeepSeek V4 Flash | DeepSeek | 1M | $0.14 | $0.28 |
| Qwen 3.7 Flash | Alibaba | - | $0.03 | $0.13 |
| Qwen 3.6 27B | Alibaba | - | - | - |
Verified 2026-08-07.