MODEL RADAR · OPENROUTER

KI-Modelle auf einen Blick vergleichen.

Alle derzeit bei OpenRouter gelisteten Modelle mit einheitlichen Angaben zu Modalitäten, Kontext, Preisen, Konfiguration und Anbietern.

KATALOGUMFANG
ALLE ÖFFENTLICHEN
Modelle
631
Anbieter
86
zuletzt synchronisiert
29.09.2026
631 Modelle
kwaivgi logokwaivgi

Kling: Video O1

Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...

Kontext
Nicht angegeben
Eingabe
text · image
Ausgabe
video
Videoausgabe: $0.112 pro Sekunde
Modelldetails ansehen →
minimax logominimax

MiniMax: Hailuo 2.3

Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...

Kontext
Nicht angegeben
Eingabe
text · image
Ausgabe
video
Videoausgabe: $0.0817 pro Sekunde
Modelldetails ansehen →
moonshotai logomoonshotai

MoonshotAI: Kimi K2.6

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

Kontext
262K
Eingabe
text · image
Ausgabe
text
Eingabe: $0.65Ausgabe: $3.41pro 1 Mio. Token
Modelldetails ansehen →
mistralai logomistralai

Mistral: Voxtral Mini TTS

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Kontext
4K
Eingabe
text
Ausgabe
speech
Zeichen: $16 pro Million Zeichen
Modelldetails ansehen →
google logogoogle

Google: Gemini Embedding 2 Preview

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...

Kontext
8K
Eingabe
text · image · file · audio · video
Ausgabe
embeddings
Eingabe: $0.2pro 1 Mio. Token
Modelldetails ansehen →
anthropic logoanthropic

Anthropic: Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Kontext
1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $5Ausgabe: $25pro 1 Mio. Token
Modelldetails ansehen →
anthropic logoanthropic

Anthropic: Claude Opus 4.7 (batch)

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Kontext
1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $2.5Ausgabe: $12.5pro 1 Mio. Token
Modelldetails ansehen →
alibaba logoalibaba

Alibaba: Wan 2.7

Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...

Kontext
Nicht angegeben
Eingabe
text · image
Ausgabe
video
Videoausgabe: $0.1 pro Sekunde
Modelldetails ansehen →
bytedance logobytedance

ByteDance: Seedance 2.0

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

Kontext
Nicht angegeben
Eingabe
text · image · video · audio
Ausgabe
video
Video (with audio): $0.1512 pro SekundeVideo (no audio): $0.1512 pro Sekunde
Modelldetails ansehen →
bytedance logobytedance

ByteDance: Seedance 2.0 Fast

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

Kontext
Nicht angegeben
Eingabe
text · image · video · audio
Ausgabe
video
Video (with audio): $0.0907 pro SekundeVideo (no audio): $0.0907 pro Sekunde
Modelldetails ansehen →
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Kontext
205K
Eingabe
text
Ausgabe
text
Eingabe: $1.4Ausgabe: $4.4pro 1 Mio. Token
Modelldetails ansehen →
cohere logocohere

Cohere: Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Kontext
33K
Eingabe
text
Ausgabe
rerank
Sucheinheiten: $0.0025 pro Suche
Modelldetails ansehen →
cohere logocohere

Cohere: Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Kontext
33K
Eingabe
text
Ausgabe
rerank
Sucheinheiten: $0.002 pro Suche
Modelldetails ansehen →
cohere logocohere

Cohere: Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

Kontext
4K
Eingabe
text
Ausgabe
rerank
Sucheinheiten: $0.001 pro Suche
Modelldetails ansehen →
google logogoogle

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Kontext
262K
Eingabe
image · text · video
Ausgabe
text
Eingabe: $0.0765Ausgabe: $0.255pro 1 Mio. Token
Modelldetails ansehen →
google logogoogle

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Kontext
262K
Eingabe
image · text · video
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen →
google logogoogle

Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Kontext
262K
Eingabe
image · text · video
Ausgabe
text
Eingabe: $0.09Ausgabe: $0.34pro 1 Mio. Token
Modelldetails ansehen →
google logogoogle

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Kontext
262K
Eingabe
image · text · video
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen →
qwen logoqwen

Qwen: Qwen3.6 Plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.325Ausgabe: $1.95pro 1 Mio. Token
Modelldetails ansehen →
z-ai logoz-ai

Z.ai: GLM 5V Turbo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Kontext
203K
Eingabe
image · text · video
Ausgabe
text
Eingabe: $1.2Ausgabe: $4pro 1 Mio. Token
Modelldetails ansehen →
arcee-ai logoarcee-ai

Arcee AI: Trinity Large Thinking

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $0.25Ausgabe: $0.8pro 1 Mio. Token
Modelldetails ansehen →
x-ai logox-ai

SpaceXAI: Grok 4.20 Multi-Agent

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

Kontext
2M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.25Ausgabe: $2.5pro 1 Mio. Token
Modelldetails ansehen →
x-ai logox-ai

SpaceXAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Kontext
2M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.25Ausgabe: $2.5pro 1 Mio. Token
Modelldetails ansehen →
google logogoogle

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Kontext
1.0M
Eingabe
text · image
Ausgabe
text · audio
Musikerzeugung: $0.08 pro Lied
Modelldetails ansehen →