MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
534
providers
75
last synced
Aug 31, 2026
534 models
bytedance logobytedance

ByteDance: Seedance 2.0

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

Context
Not provided
Input
text · image · video · audio
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
bytedance logobytedance

ByteDance: Seedance 2.0 Fast

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

Context
Not provided
Input
text · image · video · audio
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Context
205K
Input
text
Output
text
Input: $0.966Output: $3.04per 1M tokens
View model details
cohere logocohere

Cohere: Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Context
33K
Input
text
Output
rerank
Input: FreeOutput: Freeper 1M tokens
View model details
cohere logocohere

Cohere: Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Context
33K
Input
text
Output
rerank
Input: FreeOutput: Freeper 1M tokens
View model details
cohere logocohere

Cohere: Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

Context
4K
Input
text
Output
rerank
Input: FreeOutput: Freeper 1M tokens
View model details
google logogoogle

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Context
262K
Input
image · text · video
Output
text
Input: $0.070Output: $0.340per 1M tokens
View model details
google logogoogle

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Context
262K
Input
image · text · video
Output
text
Input: FreeOutput: Freeper 1M tokens
View model details
google logogoogle

Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Context
262K
Input
image · text · video
Output
text
Input: $0.090Output: $0.340per 1M tokens
View model details
google logogoogle

Google: Gemma 4 31B (batch)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Context
262K
Input
image · text · video
Output
text
Input: $0.390Output: $0.970per 1M tokens
View model details
google logogoogle

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Context
262K
Input
image · text · video
Output
text
Input: FreeOutput: Freeper 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3.6 Plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

Context
1M
Input
text · image · video
Output
text
Input: $0.325Output: $1.95per 1M tokens
View model details
arcee-ai logoarcee-ai

Arcee AI: Trinity Large Thinking

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

Context
262K
Input
text
Output
text
Input: $0.250Output: $0.800per 1M tokens
View model details
x-ai logox-ai

SpaceXAI: Grok 4.20 Multi-Agent

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

Context
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.50per 1M tokens
View model details
x-ai logox-ai

SpaceXAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Context
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.50per 1M tokens
View model details
google logogoogle

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Context
1.0M
Input
text · image
Output
text · audio
Input: FreeOutput: Freeper 1M tokens
View model details
google logogoogle

Google: Lyria 3 Clip Preview

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

Context
1.0M
Input
text · image
Output
text · audio
Input: FreeOutput: Freeper 1M tokens
View model details
alibaba logoalibaba

Alibaba: Wan 2.6

Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,...

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
kwaipilot logokwaipilot

Kwaipilot: KAT-Coder-Pro V2

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

Context
262K
Input
text
Output
text
Input: $0.300Output: $1.20per 1M tokens
View model details
bytedance logobytedance

ByteDance: Seedance 1.5 Pro

ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing...

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
openai logoopenai

OpenAI: Sora 2 Pro

OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships...

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
google logogoogle

Google: Veo 3.1

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —...

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
View model details
rekaai logorekaai

Reka Edge

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

Context
16K
Input
image · text · video
Output
text
Input: $0.100Output: $0.100per 1M tokens
View model details
minimax logominimax

MiniMax: MiniMax M2.7

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

Context
205K
Input
text
Output
text
Input: $0.300Output: $1.20per 1M tokens
View model details