Google's Gemini 3 Flash preview model optimized for speed while maintaining high capability. Features sub-second response times with strong multimodal understanding and reasoning.
Added Dec 17, 2025
Context Window
1.0M
Max Output
65.5K
Avg output tokens (7d)
752 tokens
Input Price (Auto)
$0.50/1M
Output Price (Auto)
$3.00/1M
Cache Read (Auto)
$0.050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
27.9
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
81.2%
Better than 72% of models compared
HLE
Humanity's Last Exam
15.0%
Better than 69% of models compared
IFBench
Instruction-following benchmark
55.1%
Better than 66% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
43.3%
Better than 48% of models compared
AA-LCR
Long context reasoning evaluation
53.0%
Better than 55% of models compared
CritPt
Research-level physics reasoning
1.4%
Coding
SciCode
Python programming for scientific computing
49.9%
Better than 90% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
31.8%
Better than 75% of models compared
LiveCodeBench
Contamination-free coding benchmark
79.7%
Better than 91% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
55.7%
Better than 52% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
88.2%
Better than 98% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
45.8%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
92.4%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Gemini 3 Flash (Preview) with similar models from the same provider or model family.
Gemini 3.7 Flash
google/gemini-3.7-flashGoogle's frontier-performance Flash model for multimodal and agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Google reports 65.3% on DeepSWE v1.1, up from 49.0% for Gemini 3.6 Flash, and 34% on GDP.pdf, up from 14%.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.
Gemini 3.6 Flash
google/gemini-3.6-flashGoogle's GA Flash model delivers near-Pro coding and agentic capability at Flash-tier speed and cost, with improved token efficiency over Gemini 3.5 Flash. It is optimized for multi-step orchestration, full-stack code refactoring, general reasoning, and multimodal workflows.
Gemini 3.5 Flash
google/gemini-3.5-flashGoogle's speed-focused Gemini Flash model for frontier multimodal intelligence across text, images, audio, video, PDFs, and code. Built for agentic coding, reliable tool use, structured outputs, and long-context workflows.
Gemini 3.5 Flash Thinking
google/gemini-3.5-flash-thinkingGemini 3.5 Flash with higher thinking enabled for harder reasoning, agentic coding, tool use, multimodal analysis, and long-context workflows.
Gemini Flash Latest
google/gemini-flash-latestCompatibility alias that routes to the newest version of Gemini Flash. Currently routes to Gemini 3.7 Flash.