MODEL RADAR · OPENROUTER
KI-Modelle auf einen Blick vergleichen.
Alle derzeit bei OpenRouter gelisteten Modelle mit einheitlichen Angaben zu Modalitäten, Kontext, Preisen, Konfiguration und Anbietern.
- KATALOGUMFANG
- ALLE ÖFFENTLICHEN
- Modelle
- 534
- Anbieter
- 75
- zuletzt synchronisiert
- 31.08.2026
Tencent: Hunyuan A13B Instruct
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
Morph: Morph V3 Large
Morph's high-accuracy apply model for complex code edits. 4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initialcode}</code>...
- Kontext
- 262K
- Eingabe
- text
- Ausgabe
- text
Morph: Morph V3 Fast
Morph's fastest apply model for code edits. 10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initialcode}</code> <update>{editsnippet}</update>...
- Kontext
- 82K
- Eingabe
- text
- Ausgabe
- text
Baidu: ERNIE 4.5 VL 424B A47B
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
- Kontext
- 123K
- Eingabe
- image · text
- Ausgabe
- text
Mistral: Mistral Small 3.2 24B
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
- Kontext
- 131K
- Eingabe
- image · text
- Ausgabe
- text
MiniMax: MiniMax M1
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
- Kontext
- 1M
- Eingabe
- text
- Ausgabe
- text
Google: Gemini 2.5 Flash
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
- Kontext
- 1.0M
- Eingabe
- file · image · text · audio · video
- Ausgabe
- text
Google: Gemini 2.5 Flash (batch)
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
- Kontext
- 1.0M
- Eingabe
- file · image · text · audio · video
- Ausgabe
- text
Google: Gemini 2.5 Pro
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
- Kontext
- 1.0M
- Eingabe
- text · image · file · audio · video
- Ausgabe
- text
Google: Gemini 2.5 Pro (batch)
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
- Kontext
- 1.0M
- Eingabe
- text · image · file · audio · video
- Ausgabe
- text
OpenAI: o3 Pro
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
- Kontext
- 200K
- Eingabe
- text · file · image
- Ausgabe
- text
Google: Gemini 2.5 Pro Preview 06-05
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
- Kontext
- 1.0M
- Eingabe
- file · image · text · audio
- Ausgabe
- text
DeepSeek: R1 0528
May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
- Kontext
- 164K
- Eingabe
- text
- Ausgabe
- text
Anthropic: Claude Opus 4
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
- Kontext
- 200K
- Eingabe
- image · text · file
- Ausgabe
- text
Anthropic: Claude Sonnet 4
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
- Kontext
- 1M
- Eingabe
- image · text · file
- Ausgabe
- text
Mistral: Mistral Medium 3
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...
- Kontext
- 131K
- Eingabe
- text · image · file
- Ausgabe
- text
Google: Gemini 2.5 Pro Preview 05-06
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
- Kontext
- 1.0M
- Eingabe
- text · image · file · audio · video
- Ausgabe
- text
Meta: Llama Guard 4 12B
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
- Kontext
- 164K
- Eingabe
- image · text
- Ausgabe
- text
Qwen: Qwen3 30B A3B
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
Qwen: Qwen3 8B
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
Qwen: Qwen3 14B
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
Qwen: Qwen3 32B
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
Qwen: Qwen3 235B A22B
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...
- Kontext
- 131K
- Eingabe
- text
- Ausgabe
- text
OpenAI: o4 Mini High
OpenAI o4-mini-high is the same model as o4-mini with reasoningeffort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
- Kontext
- 200K
- Eingabe
- image · text · file
- Ausgabe
- text