google logo
google

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. This model is free to use. 262,144 token context window, maximum output of 32,768 tokens. Higher uptime with 15 providers. Includes independent benchmarks from Artificial Analysis.

Panoramica

Specifiche del modello

Contesto
262.144 tokens
Output massimo
32.768 tokens
Architettura
text+image+video->text
Tokenizer
Gemma
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
Gratuito
/M tokens
Output
Gratuito
/M tokens
API

Configurazione del modello

Parametri predefiniti

temperature
1
top_p
0.95
top_k
64

Ragionamento

Moderato
No
Parametri predefiniti
No
API

Avvio rapido

Chiama questo modello tramite l’API compatibile OpenAI di OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-31b-it:free","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
imagetextvideo
Output
text
Ragionamento

No

API

Parametri API supportati

include_reasoningmax_tokensreasoningresponse_formatseedtemperaturetool_choicetoolstop_p
Provider disponibili

16 provider

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

fp4
Contesto
262K
Output massimo
16K
Input
$0.090
Output
$0.340
Lettura cache
$0.050
Scrittura cache
Non indicato

CoreWeave

fp4
Contesto
262K
Output massimo
236K
Input
$0.100
Output
$0.340
Lettura cache
$0.100
Scrittura cache
Non indicato

Venice

bf16
Contesto
256K
Output massimo
8K
Input
$0.120
Output
$0.360
Lettura cache
$0.090
Scrittura cache
Non indicato

Chutes

fp4
Contesto
131K
Output massimo
66K
Input
$0.120
Output
$0.370
Lettura cache
$0.012
Scrittura cache
Non indicato

DeepInfra

fp8
Contesto
262K
Output massimo
16K
Input
$0.130
Output
$0.380
Lettura cache
Non indicato
Scrittura cache
Non indicato

SiliconFlow

fp8
Contesto
262K
Output massimo
236K
Input
$0.130
Output
$0.400
Lettura cache
Non indicato
Scrittura cache
Non indicato

Crusoe

unknown
Contesto
262K
Output massimo
262K
Input
$0.140
Output
$0.400
Lettura cache
$0.140
Scrittura cache
Non indicato

Friendli

unknown
Contesto
262K
Output massimo
8K
Input
$0.140
Output
$0.400
Lettura cache
Non indicato
Scrittura cache
Non indicato

Novita

bf16
Contesto
262K
Output massimo
131K
Input
$0.140
Output
$0.400
Lettura cache
Non indicato
Scrittura cache
Non indicato

Parasail

fp8
Contesto
262K
Output massimo
236K
Input
$0.150
Output
$0.400
Lettura cache
$0.060
Scrittura cache
Non indicato

Phala

unknown
Contesto
262K
Output massimo
236K
Input
$0.150
Output
$0.460
Lettura cache
$0.075
Scrittura cache
Non indicato

DeepInfra

fp8
Contesto
131K
Output massimo
8K
Input
$0.270
Output
$0.760
Lettura cache
Non indicato
Scrittura cache
Non indicato

SambaNova

unknown
Contesto
131K
Output massimo
118K
Input
$0.380
Output
$1.15
Lettura cache
Non indicato
Scrittura cache
Non indicato

Together

unknown
Contesto
262K
Output massimo
236K
Input
$0.390
Output
$0.970
Lettura cache
Non indicato
Scrittura cache
Non indicato

ModelRun

fp4
Contesto
262K
Output massimo
236K
Input
$0.750
Output
$1.00
Lettura cache
$0.750
Scrittura cache
Non indicato

Cerebras

fp16
Contesto
131K
Output massimo
41K
Input
$0.990
Output
$1.49
Lettura cache
$0.990
Scrittura cache
Non indicato
google

modelli

Tutti i modelli
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.188Output: $0.938per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.88per milione di token
Vedi i dettagli del modello