Overview
Model specifications
- Context
- 1,048,576 tokens
- Maximum output
- 65,536 tokens
- Architecture
- text+image+file+audio+video->text
- Tokenizer
- Router
- Knowledge cutoff
- Not provided
- Moderated
- No
OPENROUTER
Complete pricing
Synchronized OpenRouter rates. Token prices are shown per one million tokens.
- Input
- $0.75 per 1M tokens
- Output
- $3.75 per 1M tokens
- Cache read
- $0.075 per 1M tokens
- Cache write
- $0.0417 per 1M tokens
- Image input
- $0.75 per 1M tokens
- Audio input
- $0.75 per 1M tokens
- Input Audio Cache
- $0.075 per 1M tokens
- Web search
- $14 per 1,000 calls
API
Model configuration
Reasoning
- Reasoning required
- Yes
- Default parameters
- medium
- Capabilities
- high, medium, low
API
Quick start
Set OPENROUTER_API_KEY locally. Python requires requests; JavaScript runs in Node.js. Keep the key on the server.
curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"~google/gemini-flash-latest","messages":[{"role":"user","content":"Hello!"}]}'Capabilities
Input → Output
Input
textimagevideofileaudio
Output
text
Reasoning
Yes · high · medium · low
API
Supported API parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_pAvailable providers
0 providers
Checked: September 12, 2026
Live providers on OpenRouter
OpenRouter Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.
No endpoint details were provided at the last synchronization.