MODEL RADAR · OPENROUTER

The AI model catalog, made comparable.

Search every model currently listed by OpenRouter. Compare modalities, context windows, pricing, configuration, and provider facts from one consistent snapshot.

CATALOG SCOPE
ALL PUBLIC
models
631
providers
86
last synced
Sep 29, 2026
631 models
recraft logorecraft

Recraft: Recraft V4 Pro

Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios, double the resolution of...

Context
66K
Input
text · image
Output
image
Image output: $0.25 per image
View model details →
recraft logorecraft

Recraft: Recraft V4

Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,...

Context
66K
Input
text · image
Output
image
Image output: $0.04 per image
View model details →
recraft logorecraft

Recraft: Recraft V3

Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. Supports the following imageconfig parameters:...

Context
66K
Input
text · image
Output
image
Image output: $0.04 per image
View model details →
google logogoogle

Google: Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.25Output: $1.5per 1M tokens
View model details →
google logogoogle

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.125Output: $0.75per 1M tokens
View model details →
openai logoopenai

OpenAI: GPT Chat Latest

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

Context
400K
Input
text · image · file
Output
text
Input: $5Output: $30per 1M tokens
View model details →
google logogoogle

Google: Chirp 3

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...

Context
Not provided
Input
audio
Output
transcription
Audio seconds: $0.000267 per second
View model details →
openai logoopenai

OpenAI: GPT-4o Mini Transcribe

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...

Context
128K
Input
audio
Output
transcription
Input: $1.25 per 1M tokensOutput: $5 per 1M tokens
View model details →
openai logoopenai

OpenAI: Whisper Large V3

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...

Context
Not provided
Input
audio
Output
transcription
Audio seconds: $0.000008 per second
View model details →
openai logoopenai

OpenAI: Whisper Large V3 Turbo

Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...

Context
Not provided
Input
audio
Output
transcription
Audio seconds: $0.000003 per second
View model details →
x-ai logox-ai

SpaceXAI: Grok 4.3

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

Context
1M
Input
text · image · file
Output
text
Input: $1.25Output: $2.5per 1M tokens
View model details →
x-ai logox-ai

SpaceXAI: Grok 4.3 (batch)

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

Context
1M
Input
text · image · file
Output
text
Input: $1Output: $2per 1M tokens
View model details →
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Context
262K
Input
text · image · file
Output
text
Input: $1.5Output: $7.5per 1M tokens
View model details →
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

Context
262K
Input
text · image · file
Output
text
Input: $0.75Output: $3.75per 1M tokens
View model details →
kwaivgi logokwaivgi

Kling: Video v3.0 Pro

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

Context
Not provided
Input
text · image
Output
video
Video (with audio): $0.168 per secondVideo (no audio): $0.112 per second
View model details →
kwaivgi logokwaivgi

Kling: Video v3.0 Standard

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

Context
Not provided
Input
text · image
Output
video
Video (with audio): $0.126 per secondVideo (no audio): $0.084 per second
View model details →
nvidia logonvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Context
256K
Input
text · audio · image · video
Output
text
Input: FreeOutput: Freeper 1M tokens
View model details →
openai logoopenai

OpenAI: Whisper 1

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...

Context
Not provided
Input
audio
Output
transcription
Audio seconds: $0.0001 per second
View model details →
openai logoopenai

OpenAI: GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

Context
128K
Input
audio
Output
transcription
Input: $2.5 per 1M tokensOutput: $10 per 1M tokens
View model details →
~anthropic logo~anthropic

Anthropic: Claude Haiku Latest

This model always redirects to the latest model in the Claude Haiku family.

Context
200K
Input
text · image · file
Output
text
Input: $1Output: $5per 1M tokens
View model details →
~openai logo~openai

OpenAI: GPT Mini Latest

This model always redirects to the latest model in the GPT Mini family.

Context
400K
Input
file · image · text
Output
text
Input: $0.75Output: $4.5per 1M tokens
View model details →
~google logo~google

Google: Gemini Pro Latest

This model always redirects to the latest model in the Gemini Pro family.

Context
1.0M
Input
audio · file · image · text · video
Output
text
Input: $2Output: $12per 1M tokens
View model details →
~moonshotai logo~moonshotai

MoonshotAI: Kimi Latest

This model always redirects to the latest model in the Kimi family.

Context
1.0M
Input
text · image · video
Output
text
Input: $0.8582Output: $11.1062per 1M tokens
View model details →
~google logo~google

Google: Gemini Flash Latest

This model always redirects to the latest model in the Gemini Flash family.

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.75Output: $3.75per 1M tokens
View model details →