DeepSeek logo

DeepSeek-V4.1-Flash

Multimodal
DeepSeek

DeepSeek-V4.1-Flash is DeepSeek's open-weights multimodal Mixture-of-Experts release of September 2026. Its 552B-parameter backbone activates 8B parameters per token during prefill and 16B during decode, and it supports up to 1M tokens of context. The architecture pairs a 20-layer causal encoder with a 20-layer decoder, Compressed Sparse Attention 2 with a Hierarchical Sparse Indexer, an FP4 main KV cache (about 890 bytes per token), SWA Bounded Replay, Single-Pass mHC and Engram conditional memory (196B parameters), plus DSpark speculative decoding. It accepts native image input through DeepSeek-ViT and was pretrained from scratch on a 45T-token multimodal corpus with reasoning effort continuously controllable from 1 to 100.

Key Specifications

Parameters
552.0B
Context
1.0M
Release Date
September 10, 2026
Average Score
60.9%

Timeline

Key dates in the model's history
Announcement
September 10, 2026
Last Update
September 11, 2026

Technical Specifications

Parameters
552.0B
Training Tokens
45.0T tokens
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.30
Output (per 1M tokens)
$1.20
Max Input Tokens
1.0M
Max Output Tokens
393.2K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Programming

Programming skills tests
HumanEval
Instruct; max reasoning effort (vendor-reported)Self-reported
79.4%

Mathematics

Mathematical problems and computations
GSM8k
Instruct; max reasoning effort (vendor-reported)Self-reported
93.0%

Reasoning

Logical reasoning and analysis
GPQA Diamond
Instruct; max reasoning effort (vendor-reported)Self-reported
90.9%

Other Tests

Specialized benchmarks
Terminal-Bench 2.1
Instruct; max reasoning effort (vendor-reported)Self-reported
90.6%
Terminal-Bench 3.0
Instruct; max reasoning effort (vendor-reported)Self-reported
30.0%
Terminal-Bench 4.0
Instruct; max reasoning effort (vendor-reported)Self-reported
31.2%
DeepSWE 1.1
Instruct; max reasoning effort (vendor-reported)Self-reported
74.2%
CyberGym
Instruct; max reasoning effort (vendor-reported)Self-reported
88.1%
AutomationBench
Instruct; max reasoning effort (vendor-reported)Self-reported
54.8%
Agents' Last Exam
Instruct; max reasoning effort (vendor-reported)Self-reported
31.8%
Humanity's Last Exam (with tools, text-only)
Instruct; max reasoning effort (vendor-reported)Self-reported
63.9%
HLE
Instruct; max reasoning effort (vendor-reported)Self-reported
36.8%
SEC-bench Pro
Instruct; max reasoning effort (vendor-reported)Self-reported
62.8%
NL2Repo
Instruct; max reasoning effort (vendor-reported)Self-reported
64.0%
Program Bench
Instruct; max reasoning effort (vendor-reported)Self-reported
20.3%
MathArena Apex
Instruct; max reasoning effort (vendor-reported)Self-reported
65.6%
MMLU-Pro
Instruct; max reasoning effort (vendor-reported)Self-reported
74.1%
BigCodeBench
Instruct; max reasoning effort (vendor-reported)Self-reported
60.6%
LongBench v2
Instruct; max reasoning effort (vendor-reported)Self-reported
45.2%

License & Metadata

License
mit
Announcement Date
September 10, 2026
Last Updated
September 11, 2026

Compare DeepSeek-V4.1-Flash

All comparisons

Articles about DeepSeek-V4.1-Flash

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.