DeepSeek V4 Pro
DeepSeek V4 Pro is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, rag, and agents
- Workloads that can use a 1m context window
- Buyers comparing 4 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Family
- DeepSeek V4
- Released
- 2026-04-24
- Context
- 1m
- Max output
- 384,000
- Parameters
- 1.6T
- Architecture
- Mixture of Experts
- Specialization
- general
- Openness
- Open source
- License
- MITOSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
- Training
- Pretrained
Cheapest of 5 routes · DeepSeek Platform · cache read $0.0036
About
DeepSeek V4 Pro is DeepSeek's flagship open-weights model, released April 24 2026 under the MIT license. Architecture: 1.6T total / 49B active parameters, MoE with Compressed Sparse Attention (CSA) + Heavily Compressed Attention (HCA) hybrid — requiring only 27% of inference FLOPs vs standard 1M-context transformers — plus Manifold-Constrained Hyper-Connections (mHC) and Muon Optimizer. Context window: 1,000,000 tokens; max output: 384,000 tokens (Think Max mode requires >=384K context). Text-only (no vision/image input). Supports three reasoning modes: Non-Think, Think High, Think Max. Function calling, tool use, and structured outputs supported. Key benchmarks: SWE-bench Verified 80.6%, SWE-bench Pro 55.4%, LiveCodeBench 93.5%, GPQA Diamond 90.1%, MMLU-Pro 87.5%, Terminal-Bench 2.0 59.1% on BenchLM's independent June 2026 harness, and Chatbot Arena 1456 (2026-06-16). Current API pricing: $0.435/$0.87 per 1M input/output tokens; DeepSeek made the former 75% promotional rate permanent in May 2026.
DeepSeek V4 Pro is an open-source model in the DeepSeek V4 family. The structured metadata tracks a 1m-token context window, reasoning, function calling, tool use, and structured outputs. This page tracks provider routes through DeepSeek Platform, Fireworks AI, OpenRouter, and 2 more, with the cheapest tracked route listed at $0.435 input and $0.87 output per 1M tokens. Headline tracked benchmarks include Google-Proof Q&A 90.1, Massive Multitask Language Understanding 90.1, and MMLU PRO 87.5.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ C4 relevant benchmarks in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Agents
Q/$ B1 relevant benchmark in the decision map.
Provider price ladder
Compare all 5Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| DeepSeek Platform | $0.435 | $0.870 | read $0.0036 | Serverless |
| Vercel AI Gateway | $0.435 | $0.870 | read $0.0036 | Serverless |
| OpenRouter | $0.440 | $0.870 | - | Serverless |
| Novita AI | $1.60 | $3.20 | read $0.135 | Serverless |
Available via routers & gateways(4)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Requesty
HybridAI gateway to 400+ LLM providers with intelligent routing, caching, guardrails, and governance; flat 5% markup on model costs with no subscription fee.
Respan
HybridUnified LLM engineering platform (gateway + observability + evals + prompt management) routing across 250+ models; previously Keywords AI, rebranded February 2026.
Capabilities
Benchmark peer barsfor Coding
Benchmark scores(15)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 90.1 | diamondObserved 2026-04-24 | — | Source |
| Massive Multitask Language Understanding | 90.1 | 5-shotObserved 2026-04-24 | — | Source |
| MMLU PRO | 87.5 | Think Max mode (accuracy)Observed 2026-06-07 | — | Source |
| SWE-bench Verified | 80.6 | SWE-bench VerifiedObserved 2026-04-24 | — | Source |
| Chatbot Arena | 1456.0 | text (June 2026)Observed 2026-06-16 | — | Source |
| LiveCodeBench | 93.5 | Think Max mode (pass@1)Observed 2026-06-07 | — | Source |
| SWE-bench Pro | 55.4 | —Observed 2026-04-24 | — | Source |
| HumanEval | 76.8 | Pass@1Observed 2026-04-24 | — | Source |
| Mathematics Aptitude Test of Heuristics | 64.5 | —Observed 2026-04-24 | — | Source |
| Terminal-Bench 2.0 | 59.1 | independent evaluation (BenchLM, June 2026)Observed 2026-06-18 | — | Source |
| SWE-bench Multilingual | 76.2 | —Observed 2026-04-24 | — | Source |
| BrowseComp | 83.4 | BrowseComp (accuracy%)Observed 2026-06-07 | — | Source |
| Humanity's Last Exam | 37.7 | HLE (accuracy)Observed 2026-06-07 | — | Source |
| MCP-Atlas | 69.4 | independent evaluation (BenchLM, June 2026)Observed 2026-06-18 | — | Source |
| GeneBench-Pro | 2.4 | xhighObserved 2026-06-30 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(10)
Compare DeepSeek V4 Pro with other models
Comparison and alternatives
Browse all comparisons →Show all 75 popular comparisonssorted by 7-day search impressions
Frequently asked questions
What is the context window of DeepSeek V4 Pro?
DeepSeek V4 Pro has a context window of 1m tokens.
What is the max output of DeepSeek V4 Pro?
DeepSeek V4 Pro can generate up to 384,000 output tokens.
How much does DeepSeek V4 Pro cost?
DeepSeek V4 Pro pricing ranges from $0.435/1M to $1.74/1M input tokens depending on the provider.
When was DeepSeek V4 Pro released?
DeepSeek V4 Pro was released on 2026-04-24.
Which providers offer DeepSeek V4 Pro?
DeepSeek V4 Pro is available from 5 providers: DeepSeek Platform, Fireworks AI, OpenRouter, Vercel AI Gateway, Novita AI.
What benchmarks has DeepSeek V4 Pro been tested on?
DeepSeek V4 Pro has been evaluated on 15 benchmarks, including Google-Proof Q&A, Massive Multitask Language Understanding, MMLU PRO, SWE-bench Verified, Chatbot Arena.
Cheapest of 5 routes · DeepSeek Platform · cache read $0.0036