gemini-3.8-flash

Common Name: Gemini 3.8 Flash

Google
-10%On SaleReleased on Sep 2 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Google's workhorse model for coding and agents, arriving three weeks after 3.7 Flash with sharper multi-turn coding, rapid refactoring, code synthesis, tighter tool-execution logic, and reduced verbosity.

Specifications

Context
1000K
Maximum Output
64K
Inputtext, image, audio, video
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Availability Trend (24h)

Pricing

Standard
Batch
Flex
Priority
Input/MTokens
$0.675
$0.3375
$0.3375
$1.215
Cached Input/MTokens
$0.0675
$0.03375
$0.03375
$0.1215
Output/MTokens
$3.375
$1.6875
$1.6875
$6.075
Thinking Output/MTokens
$3.375
$1.6875
$1.6875
$6.075

Performance Metrics (24h)

Similar Models

$1.125/$9.00/M-10%
ctx1.0Mmax66Kavailtps

Google's most intelligent AI model with adaptive thinking capabilities. Among the world's best models for coding and tasks requiring advanced reasoning.

$0.45/$2.70/M-10%
ctx1.0Mmax64Kavailtps

Preview of Google's next-generation Gemini 3 Flash model, optimized for speed with frontier intelligence combined with superior search and grounding capabilities.

$0.675/$3.375/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Google's most intelligent workhorse model for coding and agents, with substantial gains over 3.6 Flash in software engineering (DeepSWE v1.1 65.3%), web development (WebDev Arena Elo 1588), and knowledge-dense document processing (GDP.pdf 34.0%), at half the introductory cost.

$1.35/$6.75/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Google's workhorse model delivering better coding, knowledge work, and multimodal performance with up to 17% token reduction compared to 3.5 Flash, at lower output cost.