LLM Reference

DeepSeek Coder V2 Models by DeepSeek

DeepSeekDeepSeek LicenseOpen weightsCoding
5 models2024Up to 128k ctxFrom $0.12/1M input

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherDeepSeek
Commercial useCommercial use: permitted
Models5
Released2024
Max context128k

About

DeepSeek Coder V2 is an open-source family of Mixture-of-Experts (MoE) code language models crafted specifically for code-related tasks. It builds on the advancements of the DeepSeek V2 model, featuring notable improvements in code-specific tasks and reasoning capabilities. These models were trained on an additional 6 trillion tokens, enhancing their skills in coding and mathematical reasoning while maintaining strong general language performance. Key enhancements include support for over 338 programming languages, a significant jump from the 86 supported in earlier iterations, and an increased context length of up to 128K tokens. The family offers a variety of models such as smaller "Lite" versions for projects with limited computational resources, and larger models for complex tasks. Available on Hugging Face, these models can be accessed through their API or chat interface, making them easily deployable across diverse coding environments 123.

Decision facts

Best fit
codingcodemath-heavy prompts
Capability starting point
DeepSeek Coder V2 Instruct (0724) with 128k context
Lowest tracked input
DeepSeek Coder V2 Lite Instruct · $0.12/1M · Novita AI
Closest related family
Janus

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

5 in view

Use when the workload needs code, 128k context, and 236B parameters.

2024-07code128k context236B parameters

Use when the workload needs code, 128k context, and 236B parameters.

2024-06code128k context236B parameters

Use when the workload needs code, 128k context, and 16B parameters.

2024-06code128k context16B parameters

Use when the workload needs code, 128k context, and 236B parameters.

2024-06code128k context236B parameters

Use when the workload needs code, 128k context, and 16B parameters.

2024-06code128k context16B parameters

Release Timeline

2 release groups
2024-07
1 current
DeepSeek Coder V2 Instruct (0724)
code128k context236B parameters
Current
2024-06
4 current
DeepSeek Coder V2
code128k context236B parameters
Current
DeepSeek Coder V2 Instruct
code128k context236B parameters
Current
DeepSeek Coder V2 Lite
code128k context16B parameters
Current
DeepSeek Coder V2 Lite Instruct
code128k context16B parameters
Current

Specifications(5 models)

DeepSeek Coder V2 model specifications comparison
ModelReleasedContextParameters
DeepSeek Coder V2 Instruct (0724)2024-07128k236B
DeepSeek Coder V22024-06128k236B
DeepSeek Coder V2 Lite2024-06128k16B
DeepSeek Coder V2 Instruct2024-06128k236B
DeepSeek Coder V2 Lite Instruct2024-06128k16B

Available From(3 providers)

Pricing

DeepSeek Coder V2 model pricing by provider
ModelProviderInput / 1MOutput / 1MType
DeepSeek Coder V2 Lite InstructNovita AI$0.12$0.36Serverless
DeepSeek Coder V2DeepSeek Platform$0.14$0.28Serverless
DeepSeek Coder V2 Lite InstructFireworks AI$0.2$0.2Serverless
DeepSeek Coder V2 LiteFireworks AI$0.5$0.5Serverless
DeepSeek Coder V2Fireworks AI$1.2$1.2Serverless
DeepSeek Coder V2 InstructFireworks AI$1.2$1.2Serverless

Popular comparisons in this family