MODEL RADAR · OPENROUTER

Confronta i modelli AI con gli stessi criteri.

Tutti i modelli attualmente elencati da OpenRouter, con dati coerenti su modalità, contesto, prezzi, configurazione e provider.

AMPIEZZA CATALOGO
TUTTI PUBBLICI
modelli
534
provider
75
ultimo aggiornamento
31 ago 2026
534 modelli
minimax logominimax

MiniMax: MiniMax M2.7 (free)

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

Contesto
197K
Input
text
Output
text
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-5.4 Nano

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

Contesto
400K
Input
file · image · text
Output
text
Input: $0.200Output: $1.25per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-5.4 Mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

Contesto
400K
Input
file · image · text
Output
text
Input: $0.750Output: $4.50per milione di token
Vedi i dettagli del modello
mistralai logomistralai

Mistral: Mistral Small 4

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Contesto
262K
Input
text · image
Output
text
Input: $0.150Output: $0.600per milione di token
Vedi i dettagli del modello
mistralai logomistralai

Mistral: Mistral Small 4 (batch)

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Contesto
262K
Input
text · image
Output
text
Input: $0.150Output: $0.600per milione di token
Vedi i dettagli del modello
perplexity logoperplexity

Perplexity: Embed V1 4B

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval...

Contesto
32K
Input
text
Output
embeddings
Input: $0.030Output: Gratuitoper milione di token
Vedi i dettagli del modello
perplexity logoperplexity

Perplexity: Embed V1 0.6B

pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency...

Contesto
32K
Input
text
Output
embeddings
Input: $0.0040Output: Gratuitoper milione di token
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Nemotron 3 Super

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Contesto
1M
Input
text
Output
text
Input: $0.085Output: $0.400per milione di token
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Nemotron 3 Super (free)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Contesto
262K
Input
text
Output
text
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Lite

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.250Output: $2.00per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-9B

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.100Output: $0.150per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-9B (batch)

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.170Output: $0.250per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-5.4 Pro

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

Contesto
1.1M
Input
text · image · file
Output
text
Input: $30.00Output: $180.00per milione di token
Vedi i dettagli del modello
openai logoopenai

OpenAI: GPT-5.4

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

Contesto
1.1M
Input
text · image · file
Output
text
Input: $2.50Output: $15.00per milione di token
Vedi i dettagli del modello
inception logoinception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

Contesto
128K
Input
text
Output
text
Input: $0.250Output: $0.750per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Contesto
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.250Output: $1.50per milione di token
Vedi i dettagli del modello
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Mini

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.100Output: $0.400per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

Contesto
66K
Input
image · text
Output
image · text
Input: $0.500Output: $3.00per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.250Output: $1.25per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-27B

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.195Output: $1.56per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-122B-A10B

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Contesto
262K
Input
text · image · video
Output
text
Input: $0.290Output: $2.40per milione di token
Vedi i dettagli del modello
qwen logoqwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Contesto
1M
Input
text · image · video
Output
text
Input: $0.065Output: $0.260per milione di token
Vedi i dettagli del modello
google logogoogle

Google: Gemini 3.1 Pro Preview Custom Tools

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

Contesto
1.0M
Input
text · audio · image · video · file
Output
text
Input: $2.00Output: $12.00per milione di token
Vedi i dettagli del modello
nvidia logonvidia

NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

Contesto
131K
Input
text · image
Output
embeddings
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello