LLM Reference

DeepSeek V4 Models by DeepSeek

DeepSeekMITOpen source
3 models2026Up to 1m ctxFrom $0.05866/1M input

Last refreshed 2026-08-24. Next refresh: weekly.

Details

ResearcherDeepSeek
LicenseMITOSI-approved
Commercial useCommercial use: permitted
Models3
Released2026
Max context1m

Capabilities

Vision1 of 3 models
Multimodal1 of 3 models
ReasoningAll models
JSON / Tool useAll models
Structured OutputsAll models

Links

Website

About

DeepSeek V4 is the April 2026 DeepSeek release for long-context reasoning and coding. The family has two models: DeepSeek V4 Pro, a 1.6T-parameter MoE with 49B active parameters, a 1M-token context window, and the family's highest SWE-bench Verified score at 80.6, and DeepSeek V4 Flash, a 284B MoE with 13B active parameters for lower-cost inference. Direct DeepSeek pricing starts at $0.435 / $0.87 per million input / output tokens for Pro and $0.14 / $0.28 for Flash.

Decision facts

Best fit
visionvision and multimodal workreasoning
Capability starting point
DeepSeek V4 Flash Vision Exp with 1m context and reasoning, JSON / Tool use, structured outputs, and multimodal inputs
Lowest tracked input
DeepSeek V4 Flash · $0.05866/1M · OpenRouter
Closest related family
Janus

Compare Against Top Competitors

Long-context coding and reasoning set for comparing DeepSeek V4 Pro and Flash against Kimi, Claude, GLM, and GPT-5.Scores come from existing benchmark seed data; "-" means this site has no local score for that benchmark yet.

DeepSeek V4 flagship benchmark comparison against top competitor models
ModelContextInput / 1MChatbot ArenaSWE-bench VerifiedLiveCodeBenchGPQA
DeepSeek V4 Profamily pick1m$0.435/1M1,45680.693.590.1
DeepSeek V4 Flash1m$0.05866/1M1,4377991.688.1
Kimi K2.6262k-1,46280.289.690.5
Claude Sonnet 4.61m-1,45979.68089.9
GLM-5.1200k-1,475--86.2
GPT-5.51.05m-1,48882.6-93.6

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

3 in view

Use when the workload needs vision, 1m context, and reasoning.

2026-08vision1m contextreasoning

Use when the workload needs 1m context, 284B parameters, and reasoning.

2026-041m context284B parametersreasoning

Use when the workload needs 1m context, 1600B parameters, and reasoning.

2026-041m context1600B parametersreasoning

Release Timeline

2 release groups
2026-08
1 current
DeepSeek V4 Flash Vision Exp
vision1m contextreasoning
Current
2026-04
2 current
DeepSeek V4 Flash
1m context284B parametersreasoning
Current
DeepSeek V4 Pro
1m context1600B parametersreasoning
Current

Specifications(3 models)

DeepSeek V4 model specifications comparison
ModelReleasedContextParametersVisionMultimodalReasoningJSON / Tool useStructured Outputs
DeepSeek V4 Flash Vision Exp2026-081mYesYesYesYesYes
DeepSeek V4 Flash2026-041m284BNoNoYesYesYes
DeepSeek V4 Pro2026-041m1.6TNoNoYesYesYes

Pricing

DeepSeek V4 model pricing by provider
ModelProviderInput / 1MOutput / 1MType
DeepSeek V4 FlashOpenRouter$0.05866$0.11732Serverless
DeepSeek V4 FlashVercel AI Gateway$0.13$0.26Serverless
DeepSeek V4 FlashNovita AI$0.14$0.28Serverless
DeepSeek V4 FlashMicrosoft Foundry$0.19$0.51Serverless
DeepSeek V4 FlashDeepSeek Platform$0.22$0.66Serverless
DeepSeek V4 Flash Vision ExpDeepSeek Platform$0.22$0.66Serverless
DeepSeek V4 ProDeepSeek Platform$0.435$0.87Serverless
DeepSeek V4 ProVercel AI Gateway$0.435$0.87Serverless
DeepSeek V4 ProOpenRouter$0.44$0.87Serverless
DeepSeek V4 ProNovita AI$1.6$3.2Serverless
DeepSeek V4 ProFireworks AI$1.74$3.48Serverless

Popular comparisons in this family