LLM Reference

DeepSeek R1 vs DeepSeek V4 Flash

Keep DeepSeek R1 if downloadable weights and a 128K context window fit your deployment. Choose DeepSeek V4 Flash when you need a current public-beta direct API, 1M context, or documented JSON Output and Tool Calls. Recalculate Flash from DeepSeek's official pricing page using cache state and request time; no source-qualified R1-versus-Flash price or benchmark winner exists.

As of September 7, 2026, V4 Flash is the direct-API choice for Long context and documented JSON / Tool use; its price changes with cache state and peak or off-peak timing. Keep R1 for a weights-based deployment that fits 128K, then test both exact builds on your Coding and Agents workload.

Decision scorecard

Local evidence first
SignalDeepSeek R1DeepSeek V4 Flash
Best forreasoning-heavy apps and provider-routed productionreasoning-heavy apps, tool-calling agents, and long-context analysis
Decision fitCoding, RAG, and AgentsCoding, RAG, and Agents
Context window128k1m
Shared benchmarksWithheldWithheld

Decision tradeoffs

Choose DeepSeek R1 when...
  • Local decision data tags DeepSeek R1 for Coding, RAG, and Agents.
Choose DeepSeek V4 Flash when...
  • DeepSeek V4 Flash has the larger context window for long prompts, retrieval packs, or transcript analysis.
  • DeepSeek V4 Flash uniquely exposes JSON / Tool use in local model data.
  • Local decision data tags DeepSeek V4 Flash for Coding, RAG, and Agents.

Switch friction

DeepSeek R1 -> DeepSeek V4 Flash
  • DeepSeek V4 Flash adds JSON / Tool use in local capability data.
DeepSeek V4 Flash -> DeepSeek R1
  • Check replacement coverage for JSON / Tool use before moving production traffic.

Specs

Specification
Released2025-01-202026-04-24
Context window128k1m
Parameters671B, 37B Active284B
ArchitectureDecoder OnlyMixture of Experts
LicenseMITOSI-approvedMITOSI-approved
OpennessOpen sourceOpen source
WeightsAvailableAvailable
CodeUnknownUnknown
Commercial useCommercial use: permittedCommercial use: permitted
Knowledge cutoff2023-12-

Pricing and availability

Pricing attributeDeepSeek R1DeepSeek V4 Flash
Input price$0.10/1M tokens
Off-peak
$0.22/1M tokens
All UTC hours outside DeepSeek's peak windows. Cache-hit input: $0.007 per 1M tokens; cache-miss input: $0.22 per 1M; output: $0.66 per 1M.
Peak
$0.44/1M tokens
01:00-04:00 and 06:00-10:00 UTC. Cache-hit input: $0.014 per 1M tokens; cache-miss input: $0.44 per 1M; output: $1.32 per 1M.
Output price$0.30/1M tokens
Off-peak
$0.66/1M tokens
All UTC hours outside DeepSeek's peak windows. Cache-hit input: $0.007 per 1M tokens; cache-miss input: $0.22 per 1M; output: $0.66 per 1M.
Peak
$1.32/1M tokens
01:00-04:00 and 06:00-10:00 UTC. Cache-hit input: $0.014 per 1M tokens; cache-miss input: $0.44 per 1M; output: $1.32 per 1M.

Capabilities

CapabilityDeepSeek R1DeepSeek V4 Flash
VisionNoNo
MultimodalNoNo
ReasoningYesYes
JSON / Tool useNoYes
Structured outputsYesYes
IDE integrationNoNo
Computer useNoNo
Parallel agentsNoNo

Benchmarks

No like-for-like benchmark rows are shown for this pair.

Continue comparing

Last reviewed: 2026-07-03. Data sourced from public model cards and provider documentation.