google logo
google

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. $0.125 per million input tokens, $0.75 per million output tokens. 1,048,576 token context window, maximum output of 65,536 tokens. Higher uptime with 2 providers.

Überblick

Modellspezifikationen

Kontext
1.048.576 tokens
Maximale Ausgabe
65.536 tokens
Architektur
text+image+file+audio+video->text
Tokenizer
Gemini
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
$0.125
/M tokens
Ausgabe
$0.75
/M tokens
Cache-Lesen
$0.0125
/M tokens
Image Input
$0.125
/M tokens
Input Audio
$0.25
/M tokens
Input Audio Cache
$0.025
/M tokens
Websuche
$14
/1K calls
API

Modellkonfiguration

Schlussfolgern

Moderiert
Nein
Standardparameter
minimal
Fähigkeiten und Modalitäten
high, medium, low, minimal
API

Schnellstart

Dieses Modell über die OpenAI-kompatible API von OpenRouter aufrufen.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-flash-lite:batch","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
textimagevideofileaudio
Ausgabe
text
Schlussfolgern

Ja · high · medium · low · minimal

API

Unterstützte API-Parameter

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Verfügbare Anbieter

8 Anbieter

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Google

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.250
Ausgabe
$1.50
Cache-Lesen
$0.025
Cache-Schreiben
$0.083

Google

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.125
Ausgabe
$0.750
Cache-Lesen
$0.012
Cache-Schreiben
$0.042

Google

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.450
Ausgabe
$2.70
Cache-Lesen
$0.045
Cache-Schreiben
$0.150

Google AI Studio

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.250
Ausgabe
$1.50
Cache-Lesen
$0.025
Cache-Schreiben
$0.083

Google AI Studio

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.125
Ausgabe
$0.750
Cache-Lesen
$0.012
Cache-Schreiben
$0.042

Google AI Studio

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.450
Ausgabe
$2.70
Cache-Lesen
$0.045
Cache-Schreiben
$0.150

Google

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.275
Ausgabe
$1.65
Cache-Lesen
$0.028
Cache-Schreiben
$0.083

Google

unknown
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$0.275
Ausgabe
$1.65
Cache-Lesen
$0.028
Cache-Schreiben
$0.083
google

Modelle

Alle Modelle
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.188Ausgabe: $0.938pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.375Ausgabe: $1.88pro 1 Mio. Token
Modelldetails ansehen