Anthropic Claude Opus 4.8 with thinking enabled.
Added May 28, 2026
Context Window
1.0M
Max Output
128.0K
Avg output tokens (7d)
2.7K tokens
Input Price (Auto)
$5.00/1M
Output Price (Auto)
$25.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
57.3
Coding Index
74.3
Agentic Index
49.4
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
92.0%
Better than 95% of models compared
HLE
Humanity's Last Exam
48.7%
Better than 99% of models compared
IFBench
Instruction-following benchmark
62.2%
Better than 73% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
94.4%
Better than 93% of models compared
AA-LCR
Long context reasoning evaluation
73.0%
Better than 86% of models compared
GDPval-AA
Economically valuable tasks
54.3%
CritPt
Research-level physics reasoning
20.9%
Coding
SciCode
Python programming for scientific computing
53.5%
Better than 95% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
58.3%
Better than 98% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
48.8%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
39.3%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…