Browse all Google text models
Provider logo

Gemini 3 Flash Thinking

google/gemini-3-flash-preview-thinking
Provider logo

Gemini 3 Flash Thinking

google/gemini-3-flash-preview-thinking

Google's Gemini 3 Flash preview model with thinking mode enabled for enhanced reasoning and chain-of-thought capabilities. Audio input costs $1.00 per million audio tokens.

Added Dec 17, 2025

Context Window

1.0M

Max Output

65.5K

Avg output tokens (7d)

2.2K tokens

92%

Input Price (Auto)

$0.50/1M

Output Price (Auto)

$3.00/1M

Cache Read (Auto)

$0.050/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

26.3

Better than 81% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

80.4%

Better than 70% of models compared

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

78.0%

Better than 83% of models compared

Reasoning

HLE

Humanity's Last Exam

36.6%

Better than 87% of models compared

IFBench

Instruction-following benchmark

78.0%

Better than 97% of models compared

CritPt

Research-level physics reasoning

8.6%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

38.6%

Better than 86% of models compared

LiveCodeBench

Contamination-free coding benchmark

90.8%

Better than 99% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

97.0%

Better than 99% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

89.0%

Better than 99% of models compared

AA-Omniscience Accuracy

Proportion of correctly answered questions

53.4%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

93.0%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

89.8%

Better than 89% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

78.0%

Better than 83% of models compared

Last updated Sep 13, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Gemini 3 Flash Thinking with similar models from the same provider or model family.

Gemini 3.8 Flash

google/gemini-3.8-flash

Google's fast multimodal model for agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Its capabilities, limits, reasoning behavior, and pricing currently mirror Gemini 3.7 Flash.

Gemini 3.7 Flash

google/gemini-3.7-flash

Google's frontier-performance Flash model for multimodal and agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Google reports 65.3% on DeepSWE v1.1, up from 49.0% for Gemini 3.6 Flash, and 34% on GDP.pdf, up from 14%.

Gemini 3.5 Flash Lite

google/gemini-3.5-flash-lite

Google's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.

Gemini 3.6 Flash

google/gemini-3.6-flash

Google's GA Flash model delivers near-Pro coding and agentic capability at Flash-tier speed and cost, with improved token efficiency over Gemini 3.5 Flash. It is optimized for multi-step orchestration, full-stack code refactoring, general reasoning, and multimodal workflows.

Gemini 3.5 Flash

google/gemini-3.5-flash

Google's speed-focused Gemini Flash model for frontier multimodal intelligence across text, images, audio, video, PDFs, and code. Built for agentic coding, reliable tool use, structured outputs, and long-context workflows. Audio input costs $3.00 per million audio tokens.

Gemini 3.5 Flash Thinking

google/gemini-3.5-flash-thinking

Gemini 3.5 Flash with higher thinking enabled for harder reasoning, agentic coding, tool use, multimodal analysis, and long-context workflows. Audio input costs $3.00 per million audio tokens.