MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
631
providers
86
last synced
Sep 29, 2026
631 models
openai logoopenai

OpenAI: o3 Mini (batch)

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the reasoningeffort parameter, which can be set to...

Context
200K
Input
text · file
Output
text
Input: $0.55Output: $2.2per 1M tokens
View model details →
mistralai logomistralai

Mistral: Mistral Small 3

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...

Context
33K
Input
text
Output
text
Input: $0.05Output: $0.08per 1M tokens
View model details →
perplexity logoperplexity

Perplexity: Sonar

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features...

Context
127K
Input
text · image
Output
text
Input: $1Output: $1per 1M tokens
View model details →
deepseek logodeepseek

DeepSeek: R1

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

Context
64K
Input
text
Output
text
Input: $0.7Output: $2.5per 1M tokens
View model details →
minimax logominimax

MiniMax: MiniMax-01

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

Context
1.0M
Input
text · image
Output
text
Input: $0.2Output: $1.1per 1M tokens
View model details →
microsoft logomicrosoft

Microsoft: Phi 4

Microsoft Research Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...

Context
16K
Input
text
Output
text
Input: $0.07Output: $0.14per 1M tokens
View model details →
deepseek logodeepseek

DeepSeek: DeepSeek V3

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

Context
164K
Input
text
Output
text
Input: $0.2574Output: $1.0287per 1M tokens
View model details →
sao10k logosao10k

Sao10K: Llama 3.3 Euryale 70B

Euryale L3.3 70B is a model focused on creative roleplay from Sao10k. It is the successor of Euryale L3 70B v2.2.

Context
131K
Input
text
Output
text
Input: $0.65Output: $0.75per 1M tokens
View model details →
openai logoopenai

OpenAI: o1

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

Context
200K
Input
text · image · file
Output
text
Input: $15Output: $60per 1M tokens
View model details →
cohere logocohere

Cohere: Command R7B (12-2024)

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning...

Context
128K
Input
text
Output
text
Input: $0.0375Output: $0.15per 1M tokens
View model details →
meta-llama logometa-llama

Meta: Llama 3.3 70B Instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

Context
131K
Input
text
Output
text
Input: $0.1Output: $0.32per 1M tokens
View model details →
amazon logoamazon

Amazon: Nova Lite 1.0

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

Context
300K
Input
text · image
Output
text
Input: $0.06Output: $0.24per 1M tokens
View model details →
amazon logoamazon

Amazon: Nova Micro 1.0

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

Context
128K
Input
text
Output
text
Input: $0.035Output: $0.14per 1M tokens
View model details →
amazon logoamazon

Amazon: Nova Pro 1.0

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

Context
300K
Input
text · image
Output
text
Input: $0.8Output: $3.2per 1M tokens
View model details →
openai logoopenai

OpenAI: GPT-4o (2024-11-20)

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

Context
128K
Input
text · image · file
Output
text
Input: $2.5Output: $10per 1M tokens
View model details →
mistralai logomistralai

Mistral Large 2407

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement here....

Context
131K
Input
text · file
Output
text
Input: $2Output: $6per 1M tokens
View model details →
qwen logoqwen

Qwen2.5 Coder 32B Instruct

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in code generation, code reasoning...

Context
33K
Input
text
Output
text
Input: $0.66Output: $1per 1M tokens
View model details →
thedrummer logothedrummer

TheDrummer: UnslopNemo 12B

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

Context
1.0M
Input
text
Output
text
Input: $0.4Output: $0.4per 1M tokens
View model details →
anthracite-org logoanthracite-org

Magnum v4 72B

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of Qwen2.5 72B.

Context
33K
Input
text
Output
text
Input: $2.5Output: $5per 1M tokens
View model details →
qwen logoqwen

Qwen: Qwen2.5 7B Instruct

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

Context
33K
Input
text
Output
text
Input: $0.1Output: $0.2per 1M tokens
View model details →
meta-llama logometa-llama

Meta: Llama 3.2 1B Instruct

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

Context
60K
Input
text
Output
text
Input: $0.027Output: $0.201per 1M tokens
View model details →
meta-llama logometa-llama

Meta: Llama 3.2 3B Instruct

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

Context
131K
Input
text
Output
text
Input: $0.05Output: $0.33per 1M tokens
View model details →
qwen logoqwen

Qwen2.5 72B Instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

Context
33K
Input
text
Output
text
Input: $0.36Output: $0.4per 1M tokens
View model details →
cohere logocohere

Cohere: Command R (08-2024)

command-r-08-2024 is an update of the Command R with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

Context
128K
Input
text
Output
text
Input: $0.15Output: $0.6per 1M tokens
View model details →