Gemini 3.1 Pro preview is built for tasks where simple answers are not enough. Stronger core reasoning for complex coding, math, and long-context workflows, with multimodal support and a reported 77.1% verified score on ARC-AGI-2. NOTE: Inputs > 200k tokens are charged at 2x input and 1.5x output rates.
Added Feb 19, 2026
Context Window
1.0M
Max Output
65.5K
Avg output tokens (7d)
1.6K tokens
Input Price (Auto)
$2.00/1M
Output Price (Auto)
$12.00/1M
Cache Read (Auto)
$0.20/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
47.7
Coding Index
68.8
Agentic Index
23.0
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
94.1%
Better than 99% of models compared
HLE
Humanity's Last Exam
47.0%
Better than 98% of models compared
IFBench
Instruction-following benchmark
77.1%
Better than 96% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
95.6%
Better than 95% of models compared
AA-LCR
Long context reasoning evaluation
79.0%
Better than 97% of models compared
GDPval-AA
Economically valuable tasks
23.2%
CritPt
Research-level physics reasoning
17.7%
Coding
SciCode
Python programming for scientific computing
58.9%
Better than 99% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
53.8%
Better than 96% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
54.8%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
50.9%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Gemini 3.1 Pro (Preview) with similar models from the same provider or model family.
Gemini Pro Latest
google/gemini-pro-latestCompatibility alias that routes to the newest version of Gemini Pro. Currently routes to Gemini 3.1 Pro Preview High.
Gemini 3.1 Pro (Preview Custom Tools)
google/gemini-3.1-pro-preview-customtoolsGemini 3.1 Pro preview variant tuned for better tool selection behavior in coding agents and multi-tool workflows. It reduces overuse of generic bash tools and improves function-calling reliability while retaining Gemini 3.1 Pro's multimodal reasoning and 1M-token context. NOTE: Inputs > 200k tokens are charged at 2x input and 1.5x output rates.
Gemini 3.1 Pro (Preview High)
google/gemini-3.1-pro-preview-highGemini 3.1 Pro preview high-reasoning variant.
Gemini 3.1 Pro (Preview Low)
google/gemini-3.1-pro-preview-lowGemini 3.1 Pro preview low-reasoning variant.
Gemini 3.7 Flash
google/gemini-3.7-flashGoogle's frontier-performance Flash model for multimodal and agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Google reports 65.3% on DeepSWE v1.1, up from 49.0% for Gemini 3.6 Flash, and 34% on GDP.pdf, up from 14%.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.