DeepSeek-V4-Pro
Open model, fully self-hostable; data never leaves your domain.
01
Snapshot
Overall
77
rank #8 of 22
Coding
79
rank #5 of 22
Multimodal
77
rank #13 of 22
Input / Output
$0.44 / $1.32
USD per 1M tokens
Context
1M
usable recall 94
Speed
~70 tok/s
TTFT 0.5s
Aggregated from public sources and independently weighted; methodology on the Terms page. Scores within 3 points are treated as statistically tied.
02
Specs & pricing
Every tracked field. List API prices in USD per 1M tokens; verify current vendor pricing before purchase.
| Vendor | DeepSeek (CN) |
| Released / current | 2026.04 |
| Licensing | Open weights |
| Context window | 1M |
| Max output | 128K |
| Effective-context score | 94 |
| Input / output price | $0.44 / $1.32 per 1M tokens |
| Cache discount | No published cache discount |
| Free tier | DeepSeek App free; new API users receive credits |
| Speed / TTFT | ~70 tok/s / 0.5s |
| Function calling | 82 |
| Refusal rate | ~5% |
| English / Chinese | 82 / 85 |
| Modalities | text |
| Fine-tuning | Yes |
| Private deployment | Yes (enterprise) |
| SOC2 / no-train | no / yes |
03
Strengths & watch-outs
Strengths
- Strongest open source
- Value champion
- Excellent math reasoning
- Extremely low price
Watch-outs
- Text-only, no multimodal
- Average Chinese
- Weaker ecosystem
04
Where it ranks among 22 models
Overall 77/100 (#8), coding 79 (#5), multimodal 77 (#13). See the full boards and side-by-side compare on the leaderboard.
05
Pricing in practice
Illustrative monthly bill for 100M input + 30M output tokens: $84 at list. Your mix and cache-hit ratio change this.
Illustrative model using public list rates before any enterprise agreement; cache discount: No published cache discount.
06
Frequently asked questions
How much does DeepSeek-V4-Pro cost?
List API pricing is $0.44 input and $1.32 output per 1M tokens. No published cache discount. DeepSeek App free; new API users receive credits
What is DeepSeek-V4-Pro's context window, and is it usable end to end?
It advertises a 1M window with 128K max output; its effective-context score is 94/100, which is the better predictor of whether long-document details are actually retained.
Is DeepSeek-V4-Pro good for coding?
Its coding aggregate is 79/100, rank #5 of 22. It is a strong choice for most product engineering.
Can I self-host DeepSeek-V4-Pro?
Yes — it ships open weights and supports private deployment, so you can run it on your own infrastructure for data control; budget for GPU and operations.
How strong is it in Chinese and on multimodal inputs?
Chinese score is 85 versus English 82; multimodal aggregate is 77/100 (rank #13), accepting text.