google logo
google

Google: Gemma 4 26B A4B

Description de la source (anglais)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Présentation

Caractéristiques du modèle

Contexte
262 144 tokens
Sortie maximale
235 929 tokens
Architecture
text+image+video->text
Tokenizer
Gemma
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
$0.09
par million de tokens
Sortie
$0.3
par million de tokens
Lecture du cache
$0.05
par million de tokens
API

Configuration du modèle

Paramètres par défaut

temperature
1
top_p
0.95
top_k
64

Raisonnement

Raisonnement obligatoire
Non
Paramètres par défaut
Non
API

Démarrage rapide

Définissez OPENROUTER_API_KEY localement. Python nécessite requests ; JavaScript s’exécute dans Node.js. Gardez la clé côté serveur.

Documentation de l’API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-26b-a4b-it","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
imagetextvideo
Sortie
text
Raisonnement

Non

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Fournisseurs disponibles

11 fournisseurs

Vérifié: 13 septembre 2026

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

Darkbloom

Non indiqué
Contexte
131K
Sortie maximale
33K
Entrée
$0.042
Sortie
$0.22
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

DekaLLM

bf16
Contexte
262K
Sortie maximale
236K
Entrée
$0.06
Sortie
$0.33
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

DeepInfra

fp8
Contexte
262K
Sortie maximale
16K
Entrée
$0.07
Sortie
$0.34
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

NextBit

bf16
Contexte
262K
Sortie maximale
236K
Entrée
$0.09
Sortie
$0.3
Lecture du cache
$0.05
Écriture du cache
Non indiqué

Cloudflare

Non indiqué
Contexte
256K
Sortie maximale
230K
Entrée
$0.1
Sortie
$0.3
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Makora

Non indiqué
Contexte
262K
Sortie maximale
236K
Entrée
$0.1
Sortie
$0.34
Lecture du cache
$0.034
Écriture du cache
Non indiqué

Venice

bf16
Contexte
256K
Sortie maximale
8K
Entrée
$0.13
Sortie
$0.4
Lecture du cache
$0.05
Écriture du cache
Non indiqué

Parasail

bf16
Contexte
262K
Sortie maximale
236K
Entrée
$0.13
Sortie
$0.4
Lecture du cache
$0.05
Écriture du cache
Non indiqué

Novita

bf16
Contexte
262K
Sortie maximale
131K
Entrée
$0.13
Sortie
$0.4
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

SiliconFlow

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.14
Sortie
$0.4
Lecture du cache
$0.05
Écriture du cache
Non indiqué

Google

Non indiqué
Contexte
262K
Sortie maximale
236K
Entrée
$0.15
Sortie
$0.6
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
google

4 modèles

Tous les modèles
google logogoogle

Google: Gemini 3.8 Flash

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.75Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.8 Flash (batch)

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.875par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.75Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.875par million de tokens
Voir la fiche du modèle