Overview
Model specifications
- Context
- 1,310,720 tokens
- Maximum output
- 131,072 tokens
- Architecture
- text+image+video->text
- Tokenizer
- Router
- Knowledge cutoff
- Not provided
- Moderated
- No
OPENROUTER
Complete pricing
Synchronized OpenRouter rates. Token prices are shown per one million tokens.
- Input
- $0.0713 per 1M tokens
- Output
- $0.2375 per 1M tokens
- Cache read
- $0.0143 per 1M tokens
API
Model configuration
Default parameters
temperature- 1
top_p- 0.95
Reasoning
- Reasoning required
- Yes
- Default parameters
- max
- Capabilities
- max, high, low
API
Quick start
Set OPENROUTER_API_KEY locally. Python requires requests; JavaScript runs in Node.js. Keep the key on the server.
curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"~z-ai/glm-flash-latest","messages":[{"role":"user","content":"Hello!"}]}'Capabilities
Input → Output
Input
textimagevideo
Output
text
Reasoning
Yes · max · high · low
API
Supported API parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_pAvailable providers
0 providers
Checked: September 7, 2026
Live providers on OpenRouter
OpenRouter Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.
No endpoint details were provided at the last synchronization.