MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
534
providers
75
last synced
Aug 31, 2026
534 models
tencent logotencent

Tencent: Hunyuan A13B Instruct

Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...

Context
131K
Input
text
Output
text
Input: $0.140Output: $0.570per 1M tokens
View model details
morph logomorph

Morph: Morph V3 Large

Morph's high-accuracy apply model for complex code edits. 4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initialcode}</code>...

Context
262K
Input
text
Output
text
Input: $0.900Output: $1.90per 1M tokens
View model details
morph logomorph

Morph: Morph V3 Fast

Morph's fastest apply model for code edits. 10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initialcode}</code> <update>{editsnippet}</update>...

Context
82K
Input
text
Output
text
Input: $0.800Output: $1.20per 1M tokens
View model details
baidu logobaidu

Baidu: ERNIE 4.5 VL 424B A47B

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

Context
123K
Input
image · text
Output
text
Input: $0.420Output: $1.25per 1M tokens
View model details
mistralai logomistralai

Mistral: Mistral Small 3.2 24B

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...

Context
131K
Input
image · text
Output
text
Input: $0.075Output: $0.200per 1M tokens
View model details
minimax logominimax

MiniMax: MiniMax M1

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

Context
1M
Input
text
Output
text
Input: $0.550Output: $2.20per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Flash

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Context
1.0M
Input
file · image · text · audio · video
Output
text
Input: $0.300Output: $2.50per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Flash (batch)

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Context
1.0M
Input
file · image · text · audio · video
Output
text
Input: $0.150Output: $1.25per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Pro

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Context
1.0M
Input
text · image · file · audio · video
Output
text
Input: $1.25Output: $10.00per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Pro (batch)

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Context
1.0M
Input
text · image · file · audio · video
Output
text
Input: $0.625Output: $5.00per 1M tokens
View model details
openai logoopenai

OpenAI: o3 Pro

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

Context
200K
Input
text · file · image
Output
text
Input: $20.00Output: $80.00per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Pro Preview 06-05

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Context
1.0M
Input
file · image · text · audio
Output
text
Input: $1.25Output: $10.00per 1M tokens
View model details
deepseek logodeepseek

DeepSeek: R1 0528

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

Context
164K
Input
text
Output
text
Input: $0.500Output: $2.15per 1M tokens
View model details
anthropic logoanthropic

Anthropic: Claude Opus 4

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

Context
200K
Input
image · text · file
Output
text
Input: $15.00Output: $75.00per 1M tokens
View model details
anthropic logoanthropic

Anthropic: Claude Sonnet 4

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

Context
1M
Input
image · text · file
Output
text
Input: $3.00Output: $15.00per 1M tokens
View model details
mistralai logomistralai

Mistral: Mistral Medium 3

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

Context
131K
Input
text · image · file
Output
text
Input: $0.400Output: $2.00per 1M tokens
View model details
google logogoogle

Google: Gemini 2.5 Pro Preview 05-06

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

Context
1.0M
Input
text · image · file · audio · video
Output
text
Input: $1.25Output: $10.00per 1M tokens
View model details
meta-llama logometa-llama

Meta: Llama Guard 4 12B

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

Context
164K
Input
image · text
Output
text
Input: $0.180Output: $0.180per 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3 30B A3B

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

Context
131K
Input
text
Output
text
Input: $0.120Output: $0.500per 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3 8B

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

Context
131K
Input
text
Output
text
Input: $0.117Output: $0.455per 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3 14B

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

Context
131K
Input
text
Output
text
Input: $0.120Output: $0.240per 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3 32B

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

Context
131K
Input
text
Output
text
Input: $0.080Output: $0.280per 1M tokens
View model details
qwen logoqwen

Qwen: Qwen3 235B A22B

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

Context
131K
Input
text
Output
text
Input: $0.455Output: $1.82per 1M tokens
View model details
openai logoopenai

OpenAI: o4 Mini High

OpenAI o4-mini-high is the same model as o4-mini with reasoningeffort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

Context
200K
Input
image · text · file
Output
text
Input: $1.10Output: $4.40per 1M tokens
View model details