Same checkpoint with thinking enabled by default for deeper reasoning and stepwise analysis on complex tasks. Routed to Gemini 2.5 Flash Thinking (stable).
Added Sep 25, 2025
Context Window
1.0M
Max Output
65.5K
Input Price (Auto)
$0.30/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.030/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
15.5
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
45.6%
Better than 49% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
71.0%
Better than 65% of models compared
Reasoning
HLE
Humanity's Last Exam
13.8%
Better than 60% of models compared
IFBench
Instruction-following benchmark
52.3%
Better than 62% of models compared
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
16.7%
Better than 54% of models compared
LiveCodeBench
Contamination-free coding benchmark
71.3%
Better than 82% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
78.3%
Better than 73% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
84.2%
Better than 88% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
79.3%
Better than 66% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
71.0%
Better than 65% of models compared
Last updated Oct 1, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Gemini 2.5 Flash Preview (09/2025) – Thinking with similar models from the same provider or model family.
Gemini 2.5 Flash Lite Preview (09/2025)
gemini-2.5-flash-lite-preview-09-2025Deprecated compatibility alias. Requests route to the stable Gemini 2.5 Flash Lite model.
Gemini 2.5 Flash Lite Preview (09/2025) – Thinking
gemini-2.5-flash-lite-preview-09-2025-thinkingDeprecated compatibility alias. Requests route to Gemini 2.5 Flash Lite with the stable thinking path.
Gemini 2.5 Flash Preview (09/2025)
gemini-2.5-flash-preview-09-2025State-of-the-art Gemini 2.5 Flash checkpoint tuned for advanced reasoning, coding, math, and scientific tasks. Built-in thinking for higher accuracy and nuanced context handling. Routed to Gemini 2.5 Flash (stable).
Gemini 3.8 Flash
google/gemini-3.8-flashGoogle's fast multimodal model for agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Its capabilities, limits, reasoning behavior, and pricing currently mirror Gemini 3.7 Flash.
Gemini 3.7 Flash
google/gemini-3.7-flashGoogle's frontier-performance Flash model for multimodal and agentic workloads, including coding, tool use, image understanding, PDF and document extraction, audio, and video. Google reports 65.3% on DeepSWE v1.1, up from 49.0% for Gemini 3.6 Flash, and 34% on GDP.pdf, up from 14%.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.