DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use.
Added Sep 8, 2026
Model weightsContext Window
1M
Max Output
384K
Avg output tokens (7d)
1.5K tokens
Input Price (Auto)
$0.12/1M
Output Price (Auto)
$0.40/1M
Cache Read (Auto)
$0.0050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
39.5
Agentic work
GDPval-AA (unversioned / legacy)
Legacy fallback · Economically valuable tasks
55.0%
Document reasoning
AA-LCR (unversioned / legacy)
Legacy fallback · Long context reasoning evaluation
84.0%
Better than 98% of models compared
Reasoning
HLE
Humanity's Last Exam
39.2%
Better than 87% of models compared
CritPt
Research-level physics reasoning
14.3%
Coding
SciCode
Python programming for scientific computing
51.9%
Better than 63% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
46.4%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
96.5%
Last updated Sep 20, 2026, 12:01 AM
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare DeepSeek V4.1 Flash Thinking with similar models from the same provider or model family.
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use.
DeepSeek V4 Flash Vision Exp
deepseek/deepseek-v4-flash-vision-expAn experimental vision-enabled DeepSeek V4 Flash model that adds image understanding while retaining the text, reasoning, coding, tool-calling, and agent capabilities of the base model.
DeepSeek V4 Flash Latest
deepseek/deepseek-v4-flash-latestCompatibility alias for the current dated DeepSeek V4 Flash release. It currently routes to DeepSeek V4 Flash 0731 and inherits that release's limits and capabilities.
DeepSeek V4 Flash 0731 (Thinking)
deepseek/deepseek-v4-flash-0731:thinkingDeepSeek V4 Flash 0731 Thinking enables reasoning by default on the re-post-trained Mixture-of-Experts model with a 1M-token context window, built for coding, reasoning, and agent workflows.
DeepSeek V4 Flash 0731
deepseek/deepseek-v4-flash-0731DeepSeek V4 Flash 0731 is a re-post-trained Mixture-of-Experts model with a 1M-token context window, built for coding, reasoning, and agent workflows.
DeepSeek V4 Flash
deepseek/deepseek-v4-flashDeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with a 1M-token context window, built for fast inference, high-throughput workloads, reasoning, coding, and agent workflows.