google logo
google

Google: Gemini 3.1 Pro Preview

Descrizione della fonte (inglese)

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

Panoramica

Specifiche del modello

Contesto
1.048.576 tokens
Output massimo
65.536 tokens
Architettura
text+image+file+audio+video->text
Tokenizer
Gemini
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$2
per milione di token
  • ≤200K$2
  • >200K$4
Output
$12
per milione di token
  • ≤200K$12
  • >200K$18
Lettura cache
$0.2
per milione di token
  • ≤200K$0.2
  • >200K$0.4
Scrittura cache
$0.375
per milione di token
Input immagine
$2
per milione di token
Input audio
$2
per milione di token
  • ≤200K$2
  • >200K$4
Input Audio Cache
$0.2
per milione di token
  • ≤200K$0.2
  • >200K$0.4
Ricerca web
$14
per 1.000 chiamate
API

Configurazione del modello

Ragionamento

Ragionamento obbligatorio
Parametri predefiniti
medium
Capacità e modalità
high, medium, low
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.1-pro-preview","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
audiofileimagetextvideo
Output
text
Ragionamento

No · high · medium · low

API

Parametri API supportati

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Provider disponibili

6 provider

Verificato: 7 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

Google

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$1
Output
$6
Lettura cache
$0.1
Scrittura cache
$0.1875

Google AI Studio

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$1
Output
$6
Lettura cache
$0.1
Scrittura cache
$0.1875

Google

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$2
Output
$12
Lettura cache
$0.2
Scrittura cache
$0.375

Google AI Studio

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$2
Output
$12
Lettura cache
$0.2
Scrittura cache
$0.375

Google

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$3.6
Output
$21.6
Lettura cache
$0.36
Scrittura cache
$0.675

Google AI Studio

Non indicato
Contesto
1.0M
Output massimo
66K
Input
$3.6
Output
$21.6
Lettura cache
$0.36
Scrittura cache
$0.675
google

4 modelli

Tutti i modelli
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.75Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.875per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.75Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.875per milione di token
Vedi i dettagli del modello