DeepSeek V4 Pro
DeepSeek V4 Pro is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, rag, and agents
- Workloads that can use a 1m context window
- Buyers comparing 4 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Family
- DeepSeek V4
- Released
- 2026-04-24
- Context
- 1m
- Max output
- 384,000
- Parameters
- 1.6T
- Architecture
- Mixture of Experts
- Specialization
- general
- Openness
- Open source
- License
- MITOSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
- Training
- Pretrained
Cheapest of 5 routes · DeepSeek Platform · cache read $0.0036
About
DeepSeek V4 Pro is DeepSeek's flagship open-weights model, released April 24 2026 under the MIT license. Architecture: 1.6T total / 49B active parameters, MoE with Compressed Sparse Attention (CSA) + Heavily Compressed Attention (HCA) hybrid — requiring only 27% of inference FLOPs vs standard 1M-context transformers — plus Manifold-Constrained Hyper-Connections (mHC) and Muon Optimizer. Context window: 1,000,000 tokens; max output: 384,000 tokens (Think Max mode requires >=384K context). Text-only (no vision/image input). Supports three reasoning modes: Non-Think, Think High, Think Max. Function calling, tool use, and structured outputs supported.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ C4 relevant benchmarks in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Agents
Q/$ A1 relevant benchmark in the decision map.
Provider price ladder
Compare all 5Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| DeepSeek Platform | $0.435 | $0.870 | read $0.0036 | Serverless |
| Vercel AI Gateway | $0.435 | $0.870 | read $0.0036 | Serverless |
| OpenRouter | $0.440 | $0.870 | - | Serverless |
| Novita AI | $1.60 | $3.20 | read $0.135 | Serverless |
Available via routers & gateways(4)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Requesty
HybridAI gateway to 400+ LLM providers with intelligent routing, caching, guardrails, and governance; flat 5% markup on model costs with no subscription fee.
Respan
HybridUnified LLM engineering platform (gateway + observability + evals + prompt management) routing across 250+ models; previously Keywords AI, rebranded February 2026.
Capabilities
Benchmark peer barsfor Coding
Benchmark scores(15)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 90.1 | diamondObserved 2026-04-24 | — | Source |
| Massive Multitask Language Understanding | 90.1 | 5-shotObserved 2026-04-24 | — | Source |
| MMLU PRO | 87.5 | Think Max mode (accuracy)Observed 2026-06-07 | — | Source |
| SWE-bench Verified | 80.6 | SWE-bench VerifiedObserved 2026-04-24 | — | Source |
| Chatbot Arena | 1456.0 | text (June 2026)Observed 2026-06-16 | — | Source |
| LiveCodeBench | 93.5 | Think Max mode (pass@1)Observed 2026-06-07 | — | Source |
| SWE-bench Pro | 55.4 | —Observed 2026-04-24 | — | Source |
| HumanEval | 76.8 | Pass@1Observed 2026-04-24 | — | Source |
| Mathematics Aptitude Test of Heuristics | 64.5 | —Observed 2026-04-24 | — | Source |
| Terminal-Bench 2.0 | 59.1 | independent evaluation (BenchLM, June 2026)Observed 2026-06-18 | — | Source |
| SWE-bench Multilingual | 76.2 | —Observed 2026-04-24 | — | Source |
| BrowseComp | 83.4 | BrowseComp (accuracy%)Observed 2026-06-07 | — | Source |
| Humanity's Last Exam | 37.7 | HLE (accuracy)Observed 2026-06-07 | — | Source |
| MCP-Atlas | 69.4 | independent evaluation (BenchLM, June 2026)Observed 2026-06-18 | — | Source |
| GeneBench-Pro | 2.4 | xhighObserved 2026-06-30 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(10)
Compare DeepSeek V4 Pro with other models
Comparison and alternatives
Browse all comparisons →Show all 75 popular comparisonssorted by 7-day search impressions
Cheapest of 5 routes · DeepSeek Platform · cache read $0.0036