OpenAI's precusor to ChatGPT-4o. Great on English text and code, with significant improvements on text in non-English languages.
Context Window
128.0K
Max Output
16.4K
Avg output tokens (7d)
532 tokens
Input Price (Auto)
$2.50/1M
Output Price (Auto)
$10.00/1M
Cache Read (Auto)
$1.25/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
9.4
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
52.1%
Better than 28% of models compared
HLE
Humanity's Last Exam
2.3%
Better than 1% of models compared
IFBench
Instruction-following benchmark
36.0%
Better than 27% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
28.9%
Better than 36% of models compared
AA-LCR
Long context reasoning evaluation
39.3%
Better than 45% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
33.1%
Better than 47% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
8.3%
Better than 42% of models compared
LiveCodeBench
Contamination-free coding benchmark
31.7%
Better than 38% of models compared
Math
AIME
American Invitational Mathematics Examination
11.7%
Better than 37% of models compared
Math-500
Diverse mathematical problem solving benchmark
79.5%
Better than 45% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
19.9%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
37.9%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GPT-4o (2024-11-20) with similar models from the same provider or model family.
GPT-4o (2024-08-06)
openai/gpt-4o-2024-08-06OpenAI's precusor to ChatGPT-4o. Great on English text and code, with significant improvements on text in non-English languages.
GPT 5.6 Luna
openai/gpt-5.6-lunaGPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume chat, classification, lightweight agentic workflows, and latency-sensitive reasoning tasks.
GPT 5.6 Luna Pro
openai/gpt-5.6-luna-proGPT-5.6 Luna with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.
GPT 5.6 Sol
openai/gpt-5.6-solGPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, agentic workflows, command-line work, and multi-step professional tasks.
GPT 5.6 Sol Pro
openai/gpt-5.6-sol-proGPT-5.6 Sol with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.
GPT 5.6 Terra
openai/gpt-5.6-terraGPT-5.6 Terra is the balanced model in OpenAI's GPT-5.6 series, positioned between flagship Sol and cost-efficient Luna. It is suited for everyday coding, reasoning, agentic work, and general professional tasks.