google logo
google

Google: Gemma 3 4B

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. $0.05 per million input tokens, $0.10 per million output tokens. 131,072 token context window, maximum output of 16,384 tokens. Includes independent benchmarks from Artificial Analysis.

Überblick

Modellspezifikationen

Kontext
131.072 tokens
Maximale Ausgabe
16.384 tokens
Architektur
text+image->text
Tokenizer
Gemini
Wissensstand
2024-08-31
Moderiert
Nein
OPENROUTER

Vollständige Preise

Mit OpenRouter synchronisierte Preise; Tokenpreise gelten pro eine Million Token.

Eingabe
$0.05
/M tokens
Ausgabe
$0.1
/M tokens
API

Schnellstart

Dieses Modell über die OpenAI-kompatible API von OpenRouter aufrufen.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-3-4b-it","messages":[{"role":"user","content":"Hello!"}]}'
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
textimage
Ausgabe
text
API

Unterstützte API-Parameter

frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p
Verfügbare Anbieter

1 Anbieter

Live-Anbieter bei OpenRouter

Verfügbarkeit, Latenz, Durchsatz und Routing ändern sich laufend. Aktuelle Betriebsdaten stehen auf der Quellseite.

OpenRouter

DeepInfra

bf16
Kontext
131K
Maximale Ausgabe
16K
Eingabe
$0.050
Ausgabe
$0.100
Cache-Lesen
Nicht angegeben
Cache-Schreiben
Nicht angegeben
google

Modelle

Alle Modelle
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.188Ausgabe: $0.938pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.375Ausgabe: $1.88pro 1 Mio. Token
Modelldetails ansehen