MODEL RADAR · OPENROUTER
Compara modelos de IA con los mismos criterios.
Todos los modelos que OpenRouter publica actualmente, con datos coherentes sobre modalidades, contexto, precios, configuración y proveedores.
- ALCANCE DEL CATÁLOGO
- TODOS PÚBLICOS
- modelos
- 534
- proveedores
- 75
- última sincronización
- 31 ago 2026
MiniMax: MiniMax M2.7 (free)
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
- Contexto
- 197K
- Entrada
- text
- Salida
- text
OpenAI: GPT-5.4 Nano
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Contexto
- 400K
- Entrada
- file · image · text
- Salida
- text
OpenAI: GPT-5.4 Mini
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Contexto
- 400K
- Entrada
- file · image · text
- Salida
- text
Mistral: Mistral Small 4
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
- Contexto
- 262K
- Entrada
- text · image
- Salida
- text
Mistral: Mistral Small 4 (batch)
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
- Contexto
- 262K
- Entrada
- text · image
- Salida
- text
Perplexity: Embed V1 4B
pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval...
- Contexto
- 32K
- Entrada
- text
- Salida
- embeddings
Perplexity: Embed V1 0.6B
pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency...
- Contexto
- 32K
- Entrada
- text
- Salida
- embeddings
NVIDIA: Nemotron 3 Super
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
- Contexto
- 1M
- Entrada
- text
- Salida
- text
NVIDIA: Nemotron 3 Super (free)
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
- Contexto
- 262K
- Entrada
- text
- Salida
- text
ByteDance Seed: Seed-2.0-Lite
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Qwen: Qwen3.5-9B
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Qwen: Qwen3.5-9B (batch)
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
OpenAI: GPT-5.4 Pro
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
- Contexto
- 1.1M
- Entrada
- text · image · file
- Salida
- text
OpenAI: GPT-5.4
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Contexto
- 1.1M
- Entrada
- text · image · file
- Salida
- text
Inception: Mercury 2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
- Contexto
- 128K
- Entrada
- text
- Salida
- text
Google: Gemini 3.1 Flash Lite Preview
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
- Contexto
- 1.0M
- Entrada
- text · image · video · file · audio
- Salida
- text
ByteDance Seed: Seed-2.0-Mini
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...
- Contexto
- 66K
- Entrada
- image · text
- Salida
- image · text
Qwen: Qwen3.5-35B-A3B
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Qwen: Qwen3.5-27B
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Qwen: Qwen3.5-122B-A10B
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
- Contexto
- 262K
- Entrada
- text · image · video
- Salida
- text
Qwen: Qwen3.5-Flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
- Contexto
- 1M
- Entrada
- text · image · video
- Salida
- text
Google: Gemini 3.1 Pro Preview Custom Tools
Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...
- Contexto
- 1.0M
- Entrada
- text · audio · image · video · file
- Salida
- text
NVIDIA: Llama Nemotron Embed VL 1B V2 (free)
The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...
- Contexto
- 131K
- Entrada
- text · image
- Salida
- embeddings