MODEL RADAR · OPENROUTER

Compare modelos de IA com os mesmos critérios.

Todos os modelos atualmente listados pela OpenRouter, com dados consistentes sobre modalidades, contexto, preços, configuração e provedores.

ESCOPO DO CATÁLOGO
TODOS PÚBLICOS
modelos
534
provedores
75
última sincronização
31 de ago. de 2026
534 modelos
minimax logominimax

MiniMax: MiniMax M2.7 (free)

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

Contexto
197K
Entrada
text
Saída
text
Entrada: GrátisSaída: Grátispor milhão de tokens
Ver detalhes do modelo
openai logoopenai

OpenAI: GPT-5.4 Nano

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

Contexto
400K
Entrada
file · image · text
Saída
text
Entrada: $0.200Saída: $1.25por milhão de tokens
Ver detalhes do modelo
openai logoopenai

OpenAI: GPT-5.4 Mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

Contexto
400K
Entrada
file · image · text
Saída
text
Entrada: $0.750Saída: $4.50por milhão de tokens
Ver detalhes do modelo
mistralai logomistralai

Mistral: Mistral Small 4

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Contexto
262K
Entrada
text · image
Saída
text
Entrada: $0.150Saída: $0.600por milhão de tokens
Ver detalhes do modelo
mistralai logomistralai

Mistral: Mistral Small 4 (batch)

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Contexto
262K
Entrada
text · image
Saída
text
Entrada: $0.150Saída: $0.600por milhão de tokens
Ver detalhes do modelo
perplexity logoperplexity

Perplexity: Embed V1 4B

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval...

Contexto
32K
Entrada
text
Saída
embeddings
Entrada: $0.030Saída: Grátispor milhão de tokens
Ver detalhes do modelo
perplexity logoperplexity

Perplexity: Embed V1 0.6B

pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency...

Contexto
32K
Entrada
text
Saída
embeddings
Entrada: $0.0040Saída: Grátispor milhão de tokens
Ver detalhes do modelo
nvidia logonvidia

NVIDIA: Nemotron 3 Super

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Contexto
1M
Entrada
text
Saída
text
Entrada: $0.085Saída: $0.400por milhão de tokens
Ver detalhes do modelo
nvidia logonvidia

NVIDIA: Nemotron 3 Super (free)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Contexto
262K
Entrada
text
Saída
text
Entrada: GrátisSaída: Grátispor milhão de tokens
Ver detalhes do modelo
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Lite

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.250Saída: $2.00por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-9B

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.100Saída: $0.150por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-9B (batch)

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.170Saída: $0.250por milhão de tokens
Ver detalhes do modelo
openai logoopenai

OpenAI: GPT-5.4 Pro

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

Contexto
1.1M
Entrada
text · image · file
Saída
text
Entrada: $30.00Saída: $180.00por milhão de tokens
Ver detalhes do modelo
openai logoopenai

OpenAI: GPT-5.4

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

Contexto
1.1M
Entrada
text · image · file
Saída
text
Entrada: $2.50Saída: $15.00por milhão de tokens
Ver detalhes do modelo
inception logoinception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

Contexto
128K
Entrada
text
Saída
text
Entrada: $0.250Saída: $0.750por milhão de tokens
Ver detalhes do modelo
google logogoogle

Google: Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Contexto
1.0M
Entrada
text · image · video · file · audio
Saída
text
Entrada: $0.250Saída: $1.50por milhão de tokens
Ver detalhes do modelo
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Mini

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.100Saída: $0.400por milhão de tokens
Ver detalhes do modelo
google logogoogle

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

Contexto
66K
Entrada
image · text
Saída
image · text
Entrada: $0.500Saída: $3.00por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.250Saída: $1.25por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-27B

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.195Saída: $1.56por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-122B-A10B

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Contexto
262K
Entrada
text · image · video
Saída
text
Entrada: $0.290Saída: $2.40por milhão de tokens
Ver detalhes do modelo
qwen logoqwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Contexto
1M
Entrada
text · image · video
Saída
text
Entrada: $0.065Saída: $0.260por milhão de tokens
Ver detalhes do modelo
google logogoogle

Google: Gemini 3.1 Pro Preview Custom Tools

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

Contexto
1.0M
Entrada
text · audio · image · video · file
Saída
text
Entrada: $2.00Saída: $12.00por milhão de tokens
Ver detalhes do modelo
nvidia logonvidia

NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

Contexto
131K
Entrada
text · image
Saída
embeddings
Entrada: GrátisSaída: Grátispor milhão de tokens
Ver detalhes do modelo