Private AI
DeepSeek V4 Flash Thinking enables DeepSeek's reasoning mode on the efficiency-optimized Mixture-of-Experts model with a 1M-token context window, built for fast inference, high-throughput workloads, reasoning, coding, and agent workflows.
Added Apr 24, 2026
Model weightsContext Window
1.0M
Max Output
384.0K
Avg output tokens (7d)
1.3K tokens
Input Price (Auto)
$0.094/1M
Output Price (Auto)
$0.19/1M
Cache Read (Auto)
$0.019/1M
Capabilities
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
49.9
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Coding Index
69.1
Agentic Index
45.7
GPQA Diamond
Graduate-level scientific reasoning
90.8%
Better than 94% of models compared
HLE
Humanity's Last Exam
36.8%
Better than 94% of models compared
AA-LCR
Long context reasoning evaluation
65.7%
Better than 82% of models compared
SciCode
Python programming for scientific computing
49.9%
Better than 91% of models compared
Last updated Jul 31, 2026
Artificial Analysis