MODEL RADAR · OPENROUTER

KI-Modelle auf einen Blick vergleichen.

Alle derzeit bei OpenRouter gelisteten Modelle mit einheitlichen Angaben zu Modalitäten, Kontext, Preisen, Konfiguration und Anbietern.

KATALOGUMFANG
ALLE ÖFFENTLICHEN
Modelle
534
Anbieter
75
zuletzt synchronisiert
31.08.2026
534 Modelle
mistralai logomistralai

Mistral: Mistral Medium 3.1 (batch)

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

Kontext
131K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $0.400Ausgabe: $2.00pro 1 Mio. Token
Modelldetails ansehen
z-ai logoz-ai

Z.ai: GLM 4.5V

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

Kontext
66K
Eingabe
text · image
Ausgabe
text
Eingabe: $0.600Ausgabe: $1.80pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

Kontext
400K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.25Ausgabe: $10.00pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5 Mini

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

Kontext
400K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $0.250Ausgabe: $2.00pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-5 Nano

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

Kontext
400K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $0.050Ausgabe: $0.400pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: gpt-oss-120b

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.037Ausgabe: $0.170pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: gpt-oss-120b (batch)

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.150Ausgabe: $0.600pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: gpt-oss-20b

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.030Ausgabe: $0.130pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: gpt-oss-20b (batch)

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.050Ausgabe: $0.200pro 1 Mio. Token
Modelldetails ansehen
anthropic logoanthropic

Anthropic: Claude Opus 4.1

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

Kontext
200K
Eingabe
image · text · file
Ausgabe
text
Eingabe: $15.00Ausgabe: $75.00pro 1 Mio. Token
Modelldetails ansehen
anthropic logoanthropic

Anthropic: Claude Opus 4.1 (batch)

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

Kontext
200K
Eingabe
image · text · file
Ausgabe
text
Eingabe: $7.50Ausgabe: $37.50pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Codestral 2508

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. Blog Post

Kontext
256K
Eingabe
text · file
Ausgabe
text
Eingabe: $0.300Ausgabe: $0.900pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Codestral 2508 (batch)

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. Blog Post

Kontext
256K
Eingabe
text · file
Ausgabe
text
Eingabe: $0.300Ausgabe: $0.900pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 Coder 30B A3B Instruct

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.070Ausgabe: $0.280pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 30B A3B Instruct 2507

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.048Ausgabe: $0.193pro 1 Mio. Token
Modelldetails ansehen
z-ai logoz-ai

Z.ai: GLM 4.5 Air

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.130Ausgabe: $0.850pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 235B A22B Thinking 2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.230Ausgabe: $2.30pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 Coder 480B A35B

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.300Ausgabe: $1.00pro 1 Mio. Token
Modelldetails ansehen
bytedance logobytedance

ByteDance: UI-TARS 7B

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...

Kontext
128K
Eingabe
image · text
Ausgabe
text
Eingabe: $0.100Ausgabe: $0.200pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 2.5 Flash Lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Kontext
1.0M
Eingabe
text · image · file · audio · video
Ausgabe
text
Eingabe: $0.100Ausgabe: $0.400pro 1 Mio. Token
Modelldetails ansehen
google logogoogle

Google: Gemini 2.5 Flash Lite (batch)

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Kontext
1.0M
Eingabe
text · image · file · audio · video
Ausgabe
text
Eingabe: $0.050Ausgabe: $0.200pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.087Ausgabe: $0.350pro 1 Mio. Token
Modelldetails ansehen
moonshotai logomoonshotai

MoonshotAI: Kimi K2 0711

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.570Ausgabe: $2.30pro 1 Mio. Token
Modelldetails ansehen
cognitivecomputations logocognitivecomputations

Venice: Uncensored

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

Kontext
128K
Eingabe
text
Ausgabe
text
Eingabe: $0.200Ausgabe: $0.900pro 1 Mio. Token
Modelldetails ansehen