Cheaper version of Deepseek V3/Chat. Note: may be routed through Deepseek itself. Quantized at FP8.
Added Apr 15, 2025
Context Window
128.0K
Max Output
8.2K
Avg output tokens (7d)
592 tokens
Input Price (Auto)
$0.10/1M
Output Price (Auto)
$0.42/1M
Cache Read (Auto)
$0.050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
25.1
Coding Index
21.2
Agentic Index
1.6
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
75.1%
Better than 59% of models compared
HLE
Humanity's Last Exam
11.2%
Better than 61% of models compared
IFBench
Instruction-following benchmark
49.0%
Better than 58% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
78.9%
Better than 69% of models compared
AA-LCR
Long context reasoning evaluation
42.7%
Better than 48% of models compared
GDPval-AA
Economically valuable tasks
0.0%
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
38.7%
Better than 65% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
32.6%
Better than 76% of models compared
LiveCodeBench
Contamination-free coding benchmark
59.3%
Better than 66% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
59.0%
Better than 56% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
83.7%
Better than 86% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
24.3%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
85.9%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare DeepSeek V3/Chat Cheaper with similar models from the same provider or model family.
DeepSeek V3/Deepseek Chat
deepseek-chatDeepSeek original V3 model, trained on nearly 15 trillion tokens, matches leading closed-source models at a far lower price. Quantized at FP8.
Deepseek R1 Cheaper
deepseek-reasoner-cheaperCheaper version of DeepSeek R1. Note: may be routed through Chinese providers.
DeepSeek Chat 0324
deepseek-v3-0324DeepSeek V3 0324, DeepSeek's 03 March 2025 V3 model, optimized for general-purpose tasks. Quantized at FP8.
DeepSeek V4 Flash Vision Exp
deepseek/deepseek-v4-flash-vision-expAn experimental vision-enabled DeepSeek V4 Flash model that adds image understanding while retaining the text, reasoning, coding, tool-calling, and agent capabilities of the base model. This route is served directly by DeepSeek, so privacy and logging guarantees are limited.
DeepSeek V4 Pro 0813
deepseek/deepseek-v4-pro-0813DeepSeek V4 Pro 0813 is the general-availability release of DeepSeek V4 Pro, built for coding, tool use, cybersecurity, automation, and long-horizon agent workflows. DeepSeek reports strong gains over the preview and scores above Opus 4.8 on Terminal Bench 2.1, Cybergym, DeepSWE, and AutomationBench. It supports a 1M-token context window.
DeepSeek V4 Pro 0813 Thinking
deepseek/deepseek-v4-pro-0813:thinkingDeepSeek V4 Pro 0813 Thinking enables reasoning by default on the general-availability release of DeepSeek V4 Pro, built for coding, tool use, cybersecurity, automation, and long-horizon agent workflows. DeepSeek reports strong gains over the preview and scores above Opus 4.8 on Terminal Bench 2.1, Cybergym, DeepSWE, and AutomationBench. It supports a 1M-token context window.