DeepSeek V3/Chat Cheaper

Cheaper version of Deepseek V3/Chat. Note: may be routed through Deepseek itself. Quantized at FP8.

  • Native PDF input
  • Tool Calling
  • Structured Output

Added Apr 15, 2025

Pricing

Auto routing · per 1M tokens
Input
$0.10
Output
$0.42
Compare provider prices

Specifications

Context window
128K
Max output
8.2K
Parameters
671B / 37B
Total / active
Avg output (7d)
157 tokens
Longer than 12% of models

Benchmarks

Sourced from Artificial Analysis.

Intelligence Index

16.0

Better than 58% of models compared

Coding Index

21.2

Better than 25% of models compared

Agentic Index

0.8

Better than 7% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

78.9%

Better than 69% of models compared

GDPval-AA (unversioned / legacy)

Legacy fallback · Economically valuable tasks

0.0%

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

45.7%

Better than 41% of models compared

Reasoning

HLE

Humanity's Last Exam

11.2%

Better than 54% of models compared

IFBench

Instruction-following benchmark

49.0%

Better than 58% of models compared

CritPt

Research-level physics reasoning

0.0%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

32.6%

Better than 76% of models compared

SciCode

Python programming for scientific computing

39.0%

Better than 23% of models compared

LiveCodeBench

Contamination-free coding benchmark

59.3%

Better than 66% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

59.0%

Better than 56% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

83.7%

Better than 86% of models compared

AA-Omniscience Accuracy

Proportion of correctly answered questions

24.3%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

85.9%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

75.1%

Better than 57% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

45.7%

Better than 41% of models compared

Last updated Oct 4, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…