google logo
google

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. $0.042 per million input tokens, $0.22 per million output tokens. 262,144 token context window, maximum output of 32,768 tokens. Higher uptime with 10 providers. Includes independent benchmarks from Artificial Analysis.

개요

모델 사양

컨텍스트
262,144 tokens
최대 출력
16,384 tokens
아키텍처
text+image+video->text
토크나이저
Gemma
지식 기준일
정보 없음
검토됨
아니요
OPENROUTER

전체 가격

OpenRouter에서 동기화한 요금이며 토큰 가격은 100만 토큰 기준입니다.

입력
$0.042
/M tokens
출력
$0.22
/M tokens
API

모델 구성

기본 매개변수

temperature
1
top_p
0.95
top_k
64

추론

검토됨
아니요
기본 매개변수
아니요
API

빠른 시작

OpenRouter의 OpenAI 호환 API로 이 모델 ID를 호출합니다.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-26b-a4b-it","messages":[{"role":"user","content":"Hello!"}]}'
기능 및 모달리티

입력출력

입력
imagetextvideo
출력
text
추론

아니요

API

지원 API 매개변수

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
사용 가능한 제공업체

9 개 제공업체

OpenRouter 실시간 제공업체

가용성, 지연 시간, 처리량과 라우팅은 계속 바뀝니다. 최신 운영 데이터는 출처 페이지에서 확인하세요.

OpenRouter

Darkbloom

unknown
컨텍스트
131K
최대 출력
33K
입력
$0.042
출력
$0.220
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

DeepInfra

fp8
컨텍스트
262K
최대 출력
16K
입력
$0.070
출력
$0.340
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

Cloudflare

unknown
컨텍스트
256K
최대 출력
230K
입력
$0.100
출력
$0.300
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

NextBit

bf16
컨텍스트
262K
최대 출력
236K
입력
$0.100
출력
$0.400
캐시 읽기
$0.050
캐시 쓰기
정보 없음

SiliconFlow

fp8
컨텍스트
262K
최대 출력
236K
입력
$0.120
출력
$0.400
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

Novita

bf16
컨텍스트
262K
최대 출력
131K
입력
$0.130
출력
$0.400
캐시 읽기
정보 없음
캐시 쓰기
정보 없음

Venice

bf16
컨텍스트
256K
최대 출력
8K
입력
$0.130
출력
$0.400
캐시 읽기
$0.050
캐시 쓰기
정보 없음

Parasail

bf16
컨텍스트
262K
최대 출력
236K
입력
$0.130
출력
$0.400
캐시 읽기
$0.050
캐시 쓰기
정보 없음

Google

unknown
컨텍스트
262K
최대 출력
236K
입력
$0.150
출력
$0.600
캐시 읽기
정보 없음
캐시 쓰기
정보 없음
google

개 모델

모든 모델
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.750출력: $3.75백만 토큰당
모델 상세 보기
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.188출력: $0.938백만 토큰당
모델 상세 보기
google logogoogle

Google: Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.750출력: $3.75백만 토큰당
모델 상세 보기
google logogoogle

Google: Gemini 3.6 Flash (batch)

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.375출력: $1.88백만 토큰당
모델 상세 보기