google logo
google

Google: Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Overview

Model specifications

Context
1,048,576 tokens
Maximum output
65,536 tokens
Architecture
text+image+file+audio+video->text
Tokenizer
Gemini
Knowledge cutoff
Not provided
Moderated
No
Capabilities

InputOutput

Input
textimagevideofileaudio
Output
text
Reasoning

Yes · high · medium · low · minimal

API

Supported API parameters

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Available providers

8 providers

Google AI Studio

unknown
Context
1.0M
Maximum output
66K
Input
$0.300
Output
$2.50
Cache read
$0.030
Cache write
$0.083

Google AI Studio

unknown
Context
1.0M
Maximum output
66K
Input
$0.150
Output
$1.25
Cache read
$0.015
Cache write
$0.042

Google AI Studio

unknown
Context
1.0M
Maximum output
66K
Input
$0.540
Output
$4.50
Cache read
$0.054
Cache write
$0.150

Google

unknown
Context
1.0M
Maximum output
66K
Input
$0.300
Output
$2.50
Cache read
$0.030
Cache write
$0.083

Google

unknown
Context
1.0M
Maximum output
66K
Input
$0.150
Output
$1.25
Cache read
$0.015
Cache write
$0.042

Google

unknown
Context
1.0M
Maximum output
66K
Input
$0.540
Output
$4.50
Cache read
$0.054
Cache write
$0.150

Google

unknown
Context
1.0M
Maximum output
66K
Input
$0.330
Output
$2.75
Cache read
$0.033
Cache write
$0.083

Google

unknown
Context
1.0M
Maximum output
66K
Input
$0.330
Output
$2.75
Cache read
$0.033
Cache write
$0.083
google

models

All models
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per 1M tokens
View model details
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.188Output: $0.938per 1M tokens
View model details
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per 1M tokens
View model details
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.88per 1M tokens
View model details