google logo
google

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. This model is free to use. 262,144 token context window, maximum output of 32,768 tokens. Higher uptime with 10 providers. Includes independent benchmarks from Artificial Analysis.

Présentation

Caractéristiques du modèle

Contexte
262 144 tokens
Sortie maximale
32 768 tokens
Architecture
text+image+video->text
Tokenizer
Gemma
Limite des connaissances
Non indiqué
Modéré
Non
OPENROUTER

Tarification complète

Tarifs synchronisés depuis OpenRouter, par million de tokens.

Entrée
Gratuit
/M tokens
Sortie
Gratuit
/M tokens
API

Configuration du modèle

Paramètres par défaut

temperature
1
top_p
0.95
top_k
64

Raisonnement

Modéré
Non
Paramètres par défaut
Non
API

Démarrage rapide

Appelez ce modèle via l’API compatible OpenAI d’OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-26b-a4b-it:free","messages":[{"role":"user","content":"Hello!"}]}'
Capacités et modalités

EntréeSortie

Entrée
imagetextvideo
Sortie
text
Raisonnement

Non

API

Paramètres API pris en charge

include_reasoningmax_tokensreasoningresponse_formatseedtemperaturetool_choicetoolstop_p
Fournisseurs disponibles

9 fournisseurs

Fournisseurs en direct sur OpenRouter

Disponibilité, latence, débit et routage évoluent en continu. Consultez la source pour les données en temps réel.

OpenRouter

Darkbloom

unknown
Contexte
131K
Sortie maximale
33K
Entrée
$0.042
Sortie
$0.220
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

DeepInfra

fp8
Contexte
262K
Sortie maximale
16K
Entrée
$0.070
Sortie
$0.340
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Cloudflare

unknown
Contexte
256K
Sortie maximale
230K
Entrée
$0.100
Sortie
$0.300
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

NextBit

bf16
Contexte
262K
Sortie maximale
236K
Entrée
$0.100
Sortie
$0.400
Lecture du cache
$0.050
Écriture du cache
Non indiqué

SiliconFlow

fp8
Contexte
262K
Sortie maximale
236K
Entrée
$0.120
Sortie
$0.400
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Novita

bf16
Contexte
262K
Sortie maximale
131K
Entrée
$0.130
Sortie
$0.400
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué

Venice

bf16
Contexte
256K
Sortie maximale
8K
Entrée
$0.130
Sortie
$0.400
Lecture du cache
$0.050
Écriture du cache
Non indiqué

Parasail

bf16
Contexte
262K
Sortie maximale
236K
Entrée
$0.130
Sortie
$0.400
Lecture du cache
$0.050
Écriture du cache
Non indiqué

Google

unknown
Contexte
262K
Sortie maximale
236K
Entrée
$0.150
Sortie
$0.600
Lecture du cache
Non indiqué
Écriture du cache
Non indiqué
google

modèles

Tous les modèles
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.750Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.188Sortie: $0.938par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.750Sortie: $3.75par million de tokens
Voir la fiche du modèle
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contexte
1.0M
Entrée
text · image · video · file · audio
Sortie
text
Entrée: $0.375Sortie: $1.88par million de tokens
Voir la fiche du modèle