DeepSeek OCR Models by DeepSeek

DeepSeekMITOpen source
DeepSeek releases · 7 in the last 12 months · this family litChangelog →
2 models2025–2026Up to 8k ctxFrom $0.030/1M input

Last refreshed 2026-05-22. Next refresh: weekly.

Details

ResearcherDeepSeek
LicenseMITOSI-approved
Commercial useCommercial use: permitted
Models2
Released2025–2026
Max context8k

Capabilities

VisionAll models
MultimodalAll models

About

DeepSeek OCR is DeepSeek's family of specialized vision-language models for optical character recognition and document parsing. DeepSeek-OCR introduces Contexts Optical Compression (arXiv 2510.18234) while DeepSeek-OCR-2 introduces Visual Causal Flow for improved document understanding (arXiv 2601.20552).

Decision facts

Best fit
vision and multimodal work
Capability starting point
DeepSeek OCR 2 with 8k context and multimodal inputs
Lowest tracked input
DeepSeek OCR · $0.030/1M · Novita AI
Closest related family
Janus

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

2 in view

Use when the workload needs vision, 8k context, and multimodal inputs.

2026-01vision8k contextmultimodal inputs

Use when the workload needs vision, 8k context, and multimodal inputs.

2025-10vision8k contextmultimodal inputs

Release Timeline

2 release groups
2026-01
1 current
DeepSeek OCR 2
vision8k contextmultimodal inputs
Current
2025-10
1 current
DeepSeek OCR
vision8k contextmultimodal inputs
Current

Specifications(2 models)

DeepSeek OCR model specifications comparison
ModelReleasedContextVisionMultimodal
DeepSeek OCR 22026-018kYesYes
DeepSeek OCR2025-108kYesYes

Available From(1 provider)

Pricing

DeepSeek OCR model pricing by provider
ModelProviderInput / 1MOutput / 1MType
DeepSeek OCRNovita AI$0.03$0.03Serverless
DeepSeek OCR 2Novita AI$0.03$0.03Serverless