google logo
google

Google: Gemini 3.1 Pro Preview

Beschreibung der Quelle (Englisch)

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

Überblick

Modellspezifikationen

Kontext
1.048.576 tokens
Maximale Ausgabe
65.536 tokens
Architektur
text+image+file+audio+video->text
Tokenizer
Gemini
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
$2
pro 1 Mio. Token
  • ≤200K$2
  • >200K$4
Ausgabe
$12
pro 1 Mio. Token
  • ≤200K$12
  • >200K$18
Cache-Lesen
$0.2
pro 1 Mio. Token
  • ≤200K$0.2
  • >200K$0.4
Cache-Schreiben
$0.375
pro 1 Mio. Token
Bildeingabe
$2
pro 1 Mio. Token
Audioeingabe
$2
pro 1 Mio. Token
  • ≤200K$2
  • >200K$4
Input Audio Cache
$0.2
pro 1 Mio. Token
  • ≤200K$0.2
  • >200K$0.4
Websuche
$14
pro 1.000 Aufrufe
API

Modellkonfiguration

Schlussfolgern

Reasoning erforderlich
Ja
Standardparameter
medium
Fähigkeiten und Modalitäten
high, medium, low
API

Schnellstart

Setzen Sie OPENROUTER_API_KEY lokal. Python benötigt requests; JavaScript läuft in Node.js. Bewahren Sie den Schlüssel auf dem Server auf.

API-Dokumentation

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-pro-preview","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
audiofileimagetextvideo
Ausgabe
text
Schlussfolgern

Nein · high · medium · low

API

Unterstützte API-Parameter

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Verfügbare Anbieter

6 Anbieter

Geprüft: 7. September 2026

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Google

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$1
Ausgabe
$6
Cache-Lesen
$0.1
Cache-Schreiben
$0.1875

Google AI Studio

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$1
Ausgabe
$6
Cache-Lesen
$0.1
Cache-Schreiben
$0.1875

Google

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$2
Ausgabe
$12
Cache-Lesen
$0.2
Cache-Schreiben
$0.375

Google AI Studio

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$2
Ausgabe
$12
Cache-Lesen
$0.2
Cache-Schreiben
$0.375

Google

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$3.6
Ausgabe
$21.6
Cache-Lesen
$0.36
Cache-Schreiben
$0.675

Google AI Studio

Nicht angegeben
Kontext
1.0M
Maximale Ausgabe
66K
Eingabe
$3.6
Ausgabe
$21.6
Cache-Lesen
$0.36
Cache-Schreiben
$0.675
google

4 Modelle

Alle Modelle
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.75Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.375Ausgabe: $1.875pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.75Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.375Ausgabe: $1.875pro 1 Mio. Token
Modelldetails ansehen