MODEL RADAR · OPENROUTER

KI-Modelle auf einen Blick vergleichen.

Alle derzeit bei OpenRouter gelisteten Modelle mit einheitlichen Angaben zu Modalitäten, Kontext, Preisen, Konfiguration und Anbietern.

KATALOGUMFANG
ALLE ÖFFENTLICHEN
Modelle
534
Anbieter
75
zuletzt synchronisiert
31.08.2026
534 Modelle
minimax logominimax

MiniMax: MiniMax M2.7 (free)

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

Kontext
197K
Eingabe
text
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5.4 Nano

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

Kontext
400K
Eingabe
file · image · text
Ausgabe
text
Eingabe: $0.200Ausgabe: $1.25pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5.4 Mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

Kontext
400K
Eingabe
file · image · text
Ausgabe
text
Eingabe: $0.750Ausgabe: $4.50pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Small 4

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Kontext
262K
Eingabe
text · image
Ausgabe
text
Eingabe: $0.150Ausgabe: $0.600pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Small 4 (batch)

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

Kontext
262K
Eingabe
text · image
Ausgabe
text
Eingabe: $0.150Ausgabe: $0.600pro 1 Mio. Token
Modelldetails ansehen
perplexity logoperplexity

Perplexity: Embed V1 4B

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval...

Kontext
32K
Eingabe
text
Ausgabe
embeddings
Eingabe: $0.030Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
perplexity logoperplexity

Perplexity: Embed V1 0.6B

pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency...

Kontext
32K
Eingabe
text
Ausgabe
embeddings
Eingabe: $0.0040Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3 Super

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Kontext
1M
Eingabe
text
Ausgabe
text
Eingabe: $0.085Ausgabe: $0.400pro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3 Super (free)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Lite

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.250Ausgabe: $2.00pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-9B

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.100Ausgabe: $0.150pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-9B (batch)

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.170Ausgabe: $0.250pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5.4 Pro

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

Kontext
1.1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $30.00Ausgabe: $180.00pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5.4

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

Kontext
1.1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $2.50Ausgabe: $15.00pro 1 Mio. Token
Modelldetails ansehen
inception logoinception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

Kontext
128K
Eingabe
text
Ausgabe
text
Eingabe: $0.250Ausgabe: $0.750pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.250Ausgabe: $1.50pro 1 Mio. Token
Modelldetails ansehen
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Mini

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.100Ausgabe: $0.400pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

Kontext
66K
Eingabe
image · text
Ausgabe
image · text
Eingabe: $0.500Ausgabe: $3.00pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.250Ausgabe: $1.25pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-27B

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.195Ausgabe: $1.56pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-122B-A10B

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.290Ausgabe: $2.40pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.065Ausgabe: $0.260pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 3.1 Pro Preview Custom Tools

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

Kontext
1.0M
Eingabe
text · audio · image · video · file
Ausgabe
text
Eingabe: $2.00Ausgabe: $12.00pro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

Kontext
131K
Eingabe
text · image
Ausgabe
embeddings
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen