google logo
google

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. $0.125 per million input tokens, $0.75 per million output tokens. 1,048,576 token context window, maximum output of 65,536 tokens. Higher uptime with 2 providers.

模型概览

模型规格

上下文
1,048,576 tokens
最大输出
65,536 tokens
架构
text+image+file+audio+video->text
分词器
Gemini
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

输入
$0.125
/M tokens
输出
$0.75
/M tokens
缓存读取
$0.0125
/M tokens
Image Input
$0.125
/M tokens
Input Audio
$0.25
/M tokens
Input Audio Cache
$0.025
/M tokens
网页搜索
$14
/1K calls
API

模型配置

推理能力

内容审核
默认参数
minimal
能力与模态
high, medium, low, minimal
API

快速调用

通过 OpenRouter 的 OpenAI 兼容 API 调用这个确切的模型 ID。

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-flash-lite:batch","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
textimagevideofileaudio
输出
text
推理能力

· high · medium · low · minimal

API

支持的 API 参数

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
可用提供方

8 个提供方

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

Google

unknown
上下文
1.0M
最大输出
66K
输入
$0.250
输出
$1.50
缓存读取
$0.025
缓存写入
$0.083

Google

unknown
上下文
1.0M
最大输出
66K
输入
$0.125
输出
$0.750
缓存读取
$0.012
缓存写入
$0.042

Google

unknown
上下文
1.0M
最大输出
66K
输入
$0.450
输出
$2.70
缓存读取
$0.045
缓存写入
$0.150

Google AI Studio

unknown
上下文
1.0M
最大输出
66K
输入
$0.250
输出
$1.50
缓存读取
$0.025
缓存写入
$0.083

Google AI Studio

unknown
上下文
1.0M
最大输出
66K
输入
$0.125
输出
$0.750
缓存读取
$0.012
缓存写入
$0.042

Google AI Studio

unknown
上下文
1.0M
最大输出
66K
输入
$0.450
输出
$2.70
缓存读取
$0.045
缓存写入
$0.150

Google

unknown
上下文
1.0M
最大输出
66K
输入
$0.275
输出
$1.65
缓存读取
$0.028
缓存写入
$0.083

Google

unknown
上下文
1.0M
最大输出
66K
输入
$0.275
输出
$1.65
缓存读取
$0.028
缓存写入
$0.083
google

个模型

全部模型
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.750输出: $3.75每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.188输出: $0.938每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.750输出: $3.75每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.375输出: $1.88每百万 Token
查看模型详情