MODEL RADAR · OPENROUTER

KI-Modelle auf einen Blick vergleichen.

Alle derzeit bei OpenRouter gelisteten Modelle mit einheitlichen Angaben zu Modalitäten, Kontext, Preisen, Konfiguration und Anbietern.

KATALOGUMFANG
ALLE ÖFFENTLICHEN
Modelle
534
Anbieter
75
zuletzt synchronisiert
31.08.2026
534 Modelle
openai logoopenai

OpenAI: GPT-4o Mini Transcribe

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...

Kontext
128K
Eingabe
audio
Ausgabe
transcription
Eingabe: $1.25Ausgabe: $5.00pro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: Whisper Large V3

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Eingabe: $7.50Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: Whisper Large V3 Turbo

Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Eingabe: $3.33Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
x-ai logox-ai

SpaceXAI: Grok 4.3

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

Kontext
1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.25Ausgabe: $2.50pro 1 Mio. Token
Modelldetails ansehen
ibm-granite logoibm-granite

IBM: Granite 4.1 8B

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

Kontext
131K
Eingabe
text
Ausgabe
text
Eingabe: $0.050Ausgabe: $0.100pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Kontext
262K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.50Ausgabe: $7.50pro 1 Mio. Token
Modelldetails ansehen
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Kontext
262K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
kwaivgi logokwaivgi

Kling: Video v3.0 Pro

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

Kontext
Nicht angegeben
Eingabe
text · image
Ausgabe
video
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
kwaivgi logokwaivgi

Kling: Video v3.0 Standard

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

Kontext
Nicht angegeben
Eingabe
text · image
Ausgabe
video
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
nvidia logonvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Kontext
256K
Eingabe
text · audio · image · video
Ausgabe
text
Eingabe: KostenlosAusgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: Whisper 1

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...

Kontext
Nicht angegeben
Eingabe
audio
Ausgabe
transcription
Eingabe: $6000.00Ausgabe: Kostenlospro 1 Mio. Token
Modelldetails ansehen
openai logoopenai

OpenAI: GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

Kontext
128K
Eingabe
audio
Ausgabe
transcription
Eingabe: $2.50Ausgabe: $10.00pro 1 Mio. Token
Modelldetails ansehen
~anthropic logo~anthropic

Anthropic Claude Haiku Latest

This model always redirects to the latest model in the Anthropic Claude Haiku family.

Kontext
200K
Eingabe
text · image · file
Ausgabe
text
Eingabe: $1.00Ausgabe: $5.00pro 1 Mio. Token
Modelldetails ansehen
~openai logo~openai

OpenAI GPT Mini Latest

This model always redirects to the latest model in the OpenAI GPT Mini family.

Kontext
400K
Eingabe
file · image · text
Ausgabe
text
Eingabe: $0.750Ausgabe: $4.50pro 1 Mio. Token
Modelldetails ansehen
~google logo~google

Google Gemini Pro Latest

This model always redirects to the latest model in the Google Gemini Pro family.

Kontext
1.0M
Eingabe
audio · file · image · text · video
Ausgabe
text
Eingabe: $2.00Ausgabe: $12.00pro 1 Mio. Token
Modelldetails ansehen
~moonshotai logo~moonshotai

MoonshotAI Kimi Latest

This model always redirects to the latest model in the MoonshotAI Kimi family.

Kontext
1.0M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $2.55Ausgabe: $12.75pro 1 Mio. Token
Modelldetails ansehen
~google logo~google

Google Gemini Flash Latest

This model always redirects to the latest model in the Google Gemini Flash family.

Kontext
1.0M
Eingabe
text · image · video · file · audio
Ausgabe
text
Eingabe: $0.750Ausgabe: $3.75pro 1 Mio. Token
Modelldetails ansehen
~anthropic logo~anthropic

Anthropic Claude Sonnet Latest

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

Kontext
1M
Eingabe
text · image · file
Ausgabe
text
Eingabe: $2.00Ausgabe: $10.00pro 1 Mio. Token
Modelldetails ansehen
~openai logo~openai

OpenAI GPT Latest

This model always redirects to the latest model in the OpenAI GPT family.

Kontext
1.1M
Eingabe
file · image · text
Ausgabe
text
Eingabe: $2.00Ausgabe: $10.00pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.5 Plus 2026-04-20

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.300Ausgabe: $1.80pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.6 Flash

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

Kontext
1M
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.188Ausgabe: $1.13pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.6 35B A3B

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.100Ausgabe: $0.900pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.6 Max Preview

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

Kontext
262K
Eingabe
text
Ausgabe
text
Eingabe: $1.03Ausgabe: $6.16pro 1 Mio. Token
Modelldetails ansehen
qwen logoqwen

Qwen: Qwen3.6 27B

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

Kontext
262K
Eingabe
text · image · video
Ausgabe
text
Eingabe: $0.600Ausgabe: $3.60pro 1 Mio. Token
Modelldetails ansehen