google logo
google

Google: Gemma 3 4B

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. $0.05 per million input tokens, $0.10 per million output tokens. 131,072 token context window, maximum output of 16,384 tokens. Includes independent benchmarks from Artificial Analysis.

Panoramica

Specifiche del modello

Contesto
131.072 tokens
Output massimo
16.384 tokens
Architettura
text+image->text
Tokenizer
Gemini
Limite di conoscenza
2024-08-31
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$0.05
/M tokens
Output
$0.1
/M tokens
API

Avvio rapido

Chiama questo modello tramite l’API compatibile OpenAI di OpenRouter.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-3-4b-it","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
textimage
Output
text
API

Parametri API supportati

frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p
Provider disponibili

1 Provider

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

bf16
Contesto
131K
Output massimo
16K
Input
$0.050
Output
$0.100
Lettura cache
Non indicato
Scrittura cache
Non indicato
google

modelli

Tutti i modelli
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.188Output: $0.938per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.750Output: $3.75per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.88per milione di token
Vedi i dettagli del modello