Private AI
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with a 1M-token context window, built for fast inference, high-throughput workloads, reasoning, coding, and agent workflows.
Added Apr 24, 2026
Model weightsContext Window
1.0M
Max Output
384.0K
Avg output tokens (7d)
171 tokens
Input Price (Auto)
$0.094/1M
Output Price (Auto)
$0.19/1M
Cache Read (Auto)
$0.019/1M
Capabilities
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
40.3
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Coding Index
56.2
Agentic Index
31.1
GPQA Diamond
Graduate-level scientific reasoning
89.4%
Better than 91% of models compared
HLE
Humanity's Last Exam
32.1%
Better than 90% of models compared
IFBench
Instruction-following benchmark
79.2%
Better than 97% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
95.0%
Better than 94% of models compared
AA-LCR
Long context reasoning evaluation
63.0%
Better than 76% of models compared
GDPval-AA
Economically valuable tasks
34.5%
CritPt
Research-level physics reasoning
7.1%
SciCode
Python programming for scientific computing
44.9%
Better than 84% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
35.6%
AA-Omniscience Accuracy
Proportion of correctly answered questions
37.2%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
95.8%
Last updated Jul 31, 2026
Artificial AnalysisBetter than 82% of models compared