google logo
google

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. This model is free to use. 262,144 token context window, maximum output of 32,768 tokens. Higher uptime with 10 providers. Includes independent benchmarks from Artificial Analysis.

Überblick

Modellspezifikationen

Kontext
262.144 tokens
Maximale Ausgabe
32.768 tokens
Architektur
text+image+video->text
Tokenizer
Gemma
Wissensstand
Nicht angegeben
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
Kostenlos
/M tokens
Ausgabe
Kostenlos
/M tokens
API

Modellkonfiguration

Standardparameter

temperature
1
top_p
0.95
top_k
64

Schlussfolgern

Moderiert
Nein
Standardparameter
Nein
API

Schnellstart

Dieses Modell über die OpenAI-kompatible API von OpenRouter aufrufen.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-26b-a4b-it:free","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
imagetextvideo
Ausgabe
text
Schlussfolgern

Nein

API

Unterstützte API-Parameter

include_reasoningmax_tokensreasoningresponse_formatseedtemperaturetool_choicetoolstop_p
Verfügbare Anbieter

9 Anbieter

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

Darkbloom

unknown
Kontext
131K
Maximale Ausgabe
33K
Eingabe
$0.042
Ausgabe
$0.220
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

DeepInfra

fp8
Kontext
262K
Maximale Ausgabe
16K
Eingabe
$0.070
Ausgabe
$0.340
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Cloudflare

unknown
Kontext
256K
Maximale Ausgabe
230K
Eingabe
$0.100
Ausgabe
$0.300
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

NextBit

bf16
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.100
Ausgabe
$0.400
Cache-Lesen
$0.050
Cache-Schreiben
Nicht angegeben

SiliconFlow

fp8
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.120
Ausgabe
$0.400
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Novita

bf16
Kontext
262K
Maximale Ausgabe
131K
Eingabe
$0.130
Ausgabe
$0.400
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben

Venice

bf16
Kontext
256K
Maximale Ausgabe
8K
Eingabe
$0.130
Ausgabe
$0.400
Cache-Lesen
$0.050
Cache-Schreiben
Nicht angegeben

Parasail

bf16
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.130
Ausgabe
$0.400
Cache-Lesen
$0.050
Cache-Schreiben
Nicht angegeben

Google

unknown
Kontext
262K
Maximale Ausgabe
236K
Eingabe
$0.150
Ausgabe
$0.600
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben
google

Modelle

Alle Modelle
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.188Ausgabe: $0.938pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.375Ausgabe: $1.88pro 1 Mio. Token
Modelldetails ansehen