MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
631
providers
86
last synced
Sep 29, 2026
631 models
kwaivgi logokwaivgi

Kling: Video O1

Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...

Context
Not provided
Input
text · image
Output
video
Video output: $0.112 per second
View model details →
minimax logominimax

MiniMax: Hailuo 2.3

Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...

Context
Not provided
Input
text · image
Output
video
Video output: $0.0817 per second
View model details →
moonshotai logomoonshotai

MoonshotAI: Kimi K2.6

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

Context
262K
Input
text · image
Output
text
Input: $0.65Output: $3.41per 1M tokens
View model details →
mistralai logomistralai

Mistral: Voxtral Mini TTS

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Context
4K
Input
text
Output
speech
Characters: $16 per million characters
View model details →
google logogoogle

Google: Gemini Embedding 2 Preview

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...

Context
8K
Input
text · image · file · audio · video
Output
embeddings
Input: $0.2per 1M tokens
View model details →
anthropic logoanthropic

Anthropic: Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Context
1M
Input
text · image · file
Output
text
Input: $5Output: $25per 1M tokens
View model details →
anthropic logoanthropic

Anthropic: Claude Opus 4.7 (batch)

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Context
1M
Input
text · image · file
Output
text
Input: $2.5Output: $12.5per 1M tokens
View model details →
alibaba logoalibaba

Alibaba: Wan 2.7

Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...

Context
Not provided
Input
text · image
Output
video
Video output: $0.1 per second
View model details →
bytedance logobytedance

ByteDance: Seedance 2.0

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

Context
Not provided
Input
text · image · video · audio
Output
video
Video (with audio): $0.1512 per secondVideo (no audio): $0.1512 per second
View model details →
bytedance logobytedance

ByteDance: Seedance 2.0 Fast

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

Context
Not provided
Input
text · image · video · audio
Output
video
Video (with audio): $0.0907 per secondVideo (no audio): $0.0907 per second
View model details →
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Context
205K
Input
text
Output
text
Input: $1.4Output: $4.4per 1M tokens
View model details →
cohere logocohere

Cohere: Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Context
33K
Input
text
Output
rerank
Search units: $0.0025 per search
View model details →
cohere logocohere

Cohere: Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Context
33K
Input
text
Output
rerank
Search units: $0.002 per search
View model details →
cohere logocohere

Cohere: Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

Context
4K
Input
text
Output
rerank
Search units: $0.001 per search
View model details →
google logogoogle

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Context
262K
Input
image · text · video
Output
text
Input: $0.0765Output: $0.255per 1M tokens
View model details →
google logogoogle

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Context
262K
Input
image · text · video
Output
text
Input: FreeOutput: Freeper 1M tokens
View model details →
google logogoogle

Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Context
262K
Input
image · text · video
Output
text
Input: $0.09Output: $0.34per 1M tokens
View model details →
google logogoogle

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Context
262K
Input
image · text · video
Output
text
Input: FreeOutput: Freeper 1M tokens
View model details →
qwen logoqwen

Qwen: Qwen3.6 Plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

Context
1M
Input
text · image · video
Output
text
Input: $0.325Output: $1.95per 1M tokens
View model details →
z-ai logoz-ai

Z.ai: GLM 5V Turbo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Context
203K
Input
image · text · video
Output
text
Input: $1.2Output: $4per 1M tokens
View model details →
arcee-ai logoarcee-ai

Arcee AI: Trinity Large Thinking

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

Context
262K
Input
text
Output
text
Input: $0.25Output: $0.8per 1M tokens
View model details →
x-ai logox-ai

SpaceXAI: Grok 4.20 Multi-Agent

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

Context
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.5per 1M tokens
View model details →
x-ai logox-ai

SpaceXAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Context
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.5per 1M tokens
View model details →
google logogoogle

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Context
1.0M
Input
text · image
Output
text · audio
Song generation: $0.08 per song
View model details →