Version-pinned snapshot of GPT-5.1 from the November 13, 2025 release. Use this when audits or regulated workflows require deterministic behavior.
Context Window
1.0M
Max Output
32.8K
Input Price (Auto)
$1.25/1M
Output Price (Auto)
$10.00/1M
Cache Read (Auto)
$0.13/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
20.7
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
64.3%
Better than 41% of models compared
HLE
Humanity's Last Exam
5.3%
Better than 36% of models compared
IFBench
Instruction-following benchmark
43.2%
Better than 47% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
46.5%
Better than 50% of models compared
AA-LCR
Long context reasoning evaluation
44.3%
Better than 49% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
36.5%
Better than 57% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
22.7%
Better than 62% of models compared
LiveCodeBench
Contamination-free coding benchmark
49.4%
Better than 56% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
38.0%
Better than 39% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
80.1%
Better than 67% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
29.5%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
90.6%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GPT-5.1 (2025-11-13) with similar models from the same provider or model family.
GPT 5.6 Luna
openai/gpt-5.6-lunaGPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume chat, classification, lightweight agentic workflows, and latency-sensitive reasoning tasks.
GPT 5.6 Luna Pro
openai/gpt-5.6-luna-proGPT-5.6 Luna with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.
GPT 5.6 Sol
openai/gpt-5.6-solGPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, agentic workflows, command-line work, and multi-step professional tasks.
GPT 5.6 Sol Pro
openai/gpt-5.6-sol-proGPT-5.6 Sol with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.
GPT 5.6 Terra
openai/gpt-5.6-terraGPT-5.6 Terra is the balanced model in OpenAI's GPT-5.6 series, positioned between flagship Sol and cost-efficient Luna. It is suited for everyday coding, reasoning, agentic work, and general professional tasks.
GPT 5.6 Terra Pro
openai/gpt-5.6-terra-proGPT-5.6 Terra with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.