Overview
Model specifications
- Context
- 1,048,576 tokens
- Maximum output
- 65,536 tokens
- Architecture
- text+image+file+audio+video->text
- Tokenizer
- Router
- Knowledge cutoff
- Not provided
- Moderated
- No
OPENROUTER
Complete pricing
Synchronized OpenRouter rates. Token prices are shown per one million tokens.
- Input
- $2 /M tokens
- ≤200K$2
- >200K$4
- Output
- $12 /M tokens
- ≤200K$12
- >200K$18
- Cache read
- $0.2 /M tokens
- ≤200K$0.2
- >200K$0.4
- Cache write
- $0.375 /M tokens
- Image Input
- $2 /M tokens
- Input Audio
- $2 /M tokens
- ≤200K$2
- >200K$4
- Input Audio Cache
- $0.2 /M tokens
- ≤200K$0.2
- >200K$0.4
- Web search
- $14 /1K calls
API
Model configuration
Reasoning
- Moderated
- Yes
- Default parameters
- medium
- Capabilities
- high, medium, low
API
Quick start
Call this exact model ID through OpenRouter’s OpenAI-compatible API.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"~google/gemini-pro-latest","messages":[{"role":"user","content":"Hello!"}]}'Capabilities
Input → Output
Input
audiofileimagetextvideo
Output
text
Reasoning
No · high · medium · low
API
Supported API parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_pAvailable providers
0 providers
Live providers on OpenRouter
OpenRouter Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.
No endpoint details were provided at the last synchronization.