LLM Reference

DeepSeek V4.1 Flash

Released
2026-09-10
Last refreshed
2026-09-10
Status
Researched today
Open sourceCommercial use: permittedMultimodalRAGAgentsLong contextVisionJSON / Tool use

DeepSeek V4.1 Flash is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.

Use it for

  • Teams evaluating rag, agents, and long context
  • Workloads that can use a 1m context window
  • Buyers comparing 2 tracked provider routes

Do not use it for

  • Workloads where another current model has stronger sourced task evidence
Specifications
Released
2026-09-10
Context
1m
Max output
384,000
Parameters
552B
Architecture
Mixture of Experts
Specialization
general
Openness
Open source
License
MITOSI-approvedCommercial use: permitted
Weights
Available
Code
Unknown
Training
Pretrained
Created by

Advancing artificial general intelligence (AGI).

Hangzhou, Zhejiang, China
Founded 2023
Website
Pricing
Output / 1M
$0.600
Input / 1M
$0.150

Cheapest of 2 routes · DeepSeek Platform · cache read $0.003

About

DeepSeek V4.1 Flash is a 552B-parameter (8B prefill / 16B decode activated) mixture-of-experts model with Compressed Expert Dispatch (CED), released September 10, 2026 under the MIT license. It supports 1M-token context, up to 384K output tokens, multimodal vision input, reasoning, function calling, tool use, structured outputs, and prompt caching. Weights are available on Hugging Face. The DeepSeek API serves it as deepseek-flash; legacy API aliases deepseek-v4-flash and deepseek-v4-flash-vision-exp route to this model. Off-peak API pricing: $0.15/1M input, $0.60/1M output (cache read: $0.003/1M); peak hours are 2×.

Top use-case fit: coding, agents, and build tasks

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Long context

Included by capability and metadata signals in the decision map.

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsPrompt Caching

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.