Same checkpoint with thinking enabled by default for deeper reasoning and stepwise analysis on complex tasks. Routed to Gemini 2.5 Flash Thinking (stable).
Added Sep 25, 2025
Context Window
1.0M
Max Output
65.5K
Input Price (Auto)
$0.30/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.030/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
17.7
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
69.8%
Better than 50% of models compared
HLE
Humanity's Last Exam
12.1%
Better than 64% of models compared
Coding
SciCode
Python programming for scientific computing
35.9%
Better than 55% of models compared
LiveCodeBench
Contamination-free coding benchmark
50.5%
Better than 56% of models compared
Math
AIME
American Invitational Mathematics Examination
84.3%
Better than 91% of models compared
Math-500
Diverse mathematical problem solving benchmark
98.1%
Better than 92% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
80.0%
Better than 67% of models compared
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Gemini 2.5 Flash Preview (09/2025) – Thinking with similar models from the same provider or model family.
Gemini 2.5 Flash Lite Preview (09/2025)
gemini-2.5-flash-lite-preview-09-2025Deprecated compatibility alias. Requests route to the stable Gemini 2.5 Flash Lite model.
Gemini 2.5 Flash Lite Preview (09/2025) – Thinking
gemini-2.5-flash-lite-preview-09-2025-thinkingDeprecated compatibility alias. Requests route to Gemini 2.5 Flash Lite with the stable thinking path.
Gemini 2.5 Flash Preview (09/2025)
gemini-2.5-flash-preview-09-2025State-of-the-art Gemini 2.5 Flash checkpoint tuned for advanced reasoning, coding, math, and scientific tasks. Built-in thinking for higher accuracy and nuanced context handling. Routed to Gemini 2.5 Flash (stable).
Gemini 3.7 Flash
google/gemini-3.7-flashGoogle's frontier-performance Flash model for multimodal and agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Google reports 65.3% on DeepSWE v1.1, up from 49.0% for Gemini 3.6 Flash, and 34% on GDP.pdf, up from 14%.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.
Gemini 3.6 Flash
google/gemini-3.6-flashGoogle's GA Flash model delivers near-Pro coding and agentic capability at Flash-tier speed and cost, with improved token efficiency over Gemini 3.5 Flash. It is optimized for multi-step orchestration, full-stack code refactoring, general reasoning, and multimodal workflows.