gemini-3.5-flash

Common Name: Gemini 3.5 Flash

Google
-10%On SaleReleased on May 19 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Google's native multimodal reasoning model based on the Gemini 3 Flash foundation, optimized for agentic workflows, coding tasks, and long-context enterprise processes with configurable thinking levels.

Specifications

Context
1000K
Maximum Output
64K
Inputtext, image, audio, video
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Standard
Batch
Flex
Priority
Input/MTokens
$1.35
$0.675
$0.675
$2.43
Cached Input/MTokens
$0.135
$0.0675
$0.072
$0.243
Output/MTokens
$8.10
$4.05
$4.05
$14.58
Thinking Output/MTokens
$8.10
$4.05
$4.05
$14.58

Performance Metrics (24h)

Similar Models

$1.125/$9.00/M-10%
ctx1.0Mmax66Kavailtps

Google's most intelligent AI model with adaptive thinking capabilities. Among the world's best models for coding and tasks requiring advanced reasoning.

$1.80/$10.80/M-10%
ctx1.0Mmax66Kavailtps

Google's Gemini 3.1 Pro with advanced intelligence, complex problem-solving skills, and powerful agentic and vibe coding capabilities.

$0.675/$3.375/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Google's most intelligent workhorse model for coding and agents, with substantial gains over 3.6 Flash in software engineering (DeepSWE v1.1 65.3%), web development (WebDev Arena Elo 1588), and knowledge-dense document processing (GDP.pdf 34.0%), at half the introductory cost.

$1.35/$6.75/M-10%
ctx1.0Mmax64Kavailtps
InOutCap

Google's workhorse model delivering better coding, knowledge work, and multimodal performance with up to 17% token reduction compared to 3.5 Flash, at lower output cost.