MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
631
providers
86
last synced
Sep 29, 2026
631 models
openai logoopenai

OpenAI: GPT-4.1 Mini

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

Context
1.0M
Input
image · text · file
Output
text
Input: $0.4Output: $1.6per 1M tokens
View model details →
openai logoopenai

OpenAI: GPT-4.1 Mini (batch)

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

Context
1.0M
Input
image · text · file
Output
text
Input: $0.2Output: $0.8per 1M tokens
View model details →
openai logoopenai

OpenAI: GPT-4.1 Nano

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

Context
1.0M
Input
image · text · file
Output
text
Input: $0.1Output: $0.4per 1M tokens
View model details →
openai logoopenai

OpenAI: GPT-4.1 Nano (batch)

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

Context
1.0M
Input
image · text · file
Output
text
Input: $0.05Output: $0.2per 1M tokens
View model details →
meta-llama logometa-llama

Meta: Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Context
1.0M
Input
text · image
Output
text
Input: $0.1875Output: $0.6525per 1M tokens
View model details →
meta-llama logometa-llama

Meta: Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

Context
1.3M
Input
text · image
Output
text
Input: $0.1Output: $0.3per 1M tokens
View model details →
deepseek logodeepseek

DeepSeek: DeepSeek V3 0324

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the DeepSeek V3 model and performs really well...

Context
164K
Input
text
Output
text
Input: $0.29Output: $1.14per 1M tokens
View model details →
openai logoopenai

OpenAI: o1-pro

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

Context
200K
Input
text · image · file
Output
text
Input: $150Output: $600per 1M tokens
View model details →
mistralai logomistralai

Mistral: Mistral Small 3.1 24B

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

Context
128K
Input
text · image
Output
text
Input: $0.351Output: $0.555per 1M tokens
View model details →
google logogoogle

Google: Gemma 3 4B

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

Context
131K
Input
text · image
Output
text
Input: $0.05Output: $0.1per 1M tokens
View model details →
google logogoogle

Google: Gemma 3 12B

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

Context
131K
Input
text · image
Output
text
Input: $0.05Output: $0.15per 1M tokens
View model details →
cohere logocohere

Cohere: Command A

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...

Context
256K
Input
text
Output
text
Input: $2.5Output: $10per 1M tokens
View model details →
rekaai logorekaai

Reka Flash 3

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...

Context
66K
Input
text
Output
text
Input: $0.1Output: $0.2per 1M tokens
View model details →
google logogoogle

Google: Gemma 3 27B

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

Context
131K
Input
text · image
Output
text
Input: $0.08Output: $0.45per 1M tokens
View model details →
thedrummer logothedrummer

TheDrummer: Skyfall 36B V2

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

Context
33K
Input
text
Output
text
Input: $0.55Output: $0.8per 1M tokens
View model details →
perplexity logoperplexity

Perplexity: Sonar Reasoning Pro

Note: Sonar Pro pricing includes Perplexity search pricing. See details here Sonar Reasoning Pro is a premier reasoning model powered by DeepSeek R1 with Chain of Thought (CoT). Designed for...

Context
128K
Input
text · image
Output
text
Input: $2Output: $8per 1M tokens
View model details →
perplexity logoperplexity

Perplexity: Sonar Pro

Note: Sonar Pro pricing includes Perplexity search pricing. See details here For enterprises seeking more advanced capabilities, the Sonar Pro API can handle in-depth, multi-step queries with added extensibility, like...

Context
200K
Input
text · image
Output
text
Input: $3Output: $15per 1M tokens
View model details →
perplexity logoperplexity

Perplexity: Sonar Deep Research

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across complex topics. It autonomously searches, reads, and evaluates sources, refining its approach as it gathers...

Context
128K
Input
text
Output
text
Input: $2Output: $8per 1M tokens
View model details →
mistralai logomistralai

Mistral: Saba

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

Context
33K
Input
text · file
Output
text
Input: $0.2Output: $0.6per 1M tokens
View model details →
openai logoopenai

OpenAI: o3 Mini High

OpenAI o3-mini-high is the same model as o3-mini with reasoningeffort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

Context
200K
Input
text · file
Output
text
Input: $1.1Output: $4.4per 1M tokens
View model details →
aion-labs logoaion-labs

AionLabs: Aion-RP 1.0 (8B)

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

Context
33K
Input
text
Output
text
Input: $0.8Output: $1.6per 1M tokens
View model details →
qwen logoqwen

Qwen: Qwen2.5 VL 72B Instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

Context
128K
Input
text · image
Output
text
Input: $0.8Output: $1per 1M tokens
View model details →
qwen logoqwen

Qwen: Qwen-Plus

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

Context
1M
Input
text
Output
text
Input: $0.26Output: $0.78per 1M tokens
View model details →
openai logoopenai

OpenAI: o3 Mini

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the reasoningeffort parameter, which can be set to...

Context
200K
Input
text · file
Output
text
Input: $1.1Output: $4.4per 1M tokens
View model details →