google logo
google

Google: Gemma 4 31B

Description de la source (anglais)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Présentation

Caractéristiques du modèle

Contexte
262 144 tokens
Sortie maximale
16 384 tokens
Architecture
text+image+video->text
Tokenizer
Gemma
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.09
par million de tokens
Sortie
$0.34
par million de tokens
Lecture du cache
$0.05
par million de tokens
API

Configuration du modèle

Paramètres par défaut

temperature
1
top_p
0.95
top_k
64

Raisonnement

Raisonnement obligatoire
Non
Paramètres par défaut
Non
API

Démarrage rapide

Définissez OPENROUTER_API_KEY localement. Python nécessite requests ; JavaScript s’exécute dans Node.js. Gardez la clé côté serveur.

Documentation de l’API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-31b-it","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
imagetextvideo
Sortie
text
Raisonnement

Non

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Fournisseurs disponibles

15 fournisseurs

Vérifié: 7 septembre 2026

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

DeepInfra

fp4
Contexte
262K
Sortie maximale
16K
Entrée
$0.09
Sortie
$0.34
Lecture du cache
$0.05
Écriture du cache
Non indiqué

CoreWeave

fp4
Contexte
262K
Sortie maximale
236K
Entrée
$0.1
Sortie
$0.34
Lecture du cache
$0.1
Écriture du cache
Non indiqué

Venice

bf16
Contexte
256K
Sortie maximale
8K
Entrée
$0.12
Sortie
$0.36
Lecture du cache
$0.09
Écriture du cache
Non indiqué

Chutes

fp4
Contexte
131K
Sortie maximale
66K
Entrée
$0.12
Sortie
$0.37
Lecture du cache
$0.012
Écriture du cache
Non indiqué

DeepInfra

fp8
Contexte
262K
Sortie maximale
16K
Entrée
$0.13
Sortie
$0.38
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

SiliconFlow

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.13
Sortie
$0.4
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Crusoe

Non indiqué
Contexte
262K
Sortie maximale
262K
Entrée
$0.14
Sortie
$0.4
Lecture du cache
$0.14
Écriture du cache
Non indiqué

Friendli

Non indiqué
Contexte
262K
Sortie maximale
8K
Entrée
$0.14
Sortie
$0.4
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Novita

bf16
Contexte
262K
Sortie maximale
131K
Entrée
$0.14
Sortie
$0.4
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Parasail

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.15
Sortie
$0.4
Lecture du cache
$0.06
Écriture du cache
Non indiqué

DeepInfra

fp8
Contexte
131K
Sortie maximale
8K
Entrée
$0.27
Sortie
$0.76
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

SambaNova

Non indiqué
Contexte
131K
Sortie maximale
118K
Entrée
$0.38
Sortie
$1.15
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Together

Non indiqué
Contexte
262K
Sortie maximale
236K
Entrée
$0.39
Sortie
$0.97
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

ModelRun

fp4
Contexte
262K
Sortie maximale
236K
Entrée
$0.75
Sortie
$1
Lecture du cache
$0.75
Écriture du cache
Non indiqué

Cerebras

fp16
Contexte
131K
Sortie maximale
41K
Entrée
$0.99
Sortie
$1.49
Lecture du cache
$0.99
Écriture du cache
Non indiqué
google

4 modèles

Tous les modèles
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.75Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.875par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.75Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.875par million de tokens
Voir la fiche du modèle