モデルレーダー · OPENROUTER

AIモデルを、同じ基準で比較。

OpenRouterで現在公開されているすべてのモデルを検索し、モダリティ、コンテキスト、料金、設定、プロバイダー情報を比較できます。

カタログ範囲
公開モデルすべて
モデル
631
プロバイダー
86
最終同期
2026/09/29
631 モデル
recraft logorecraft

Recraft: Recraft V4 Pro

Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios, double the resolution of...

コンテキスト
66K
入力
text · image
出力
image
画像出力: $0.25 画像1枚あたり
モデル詳細を見る →
recraft logorecraft

Recraft: Recraft V4

Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,...

コンテキスト
66K
入力
text · image
出力
image
画像出力: $0.04 画像1枚あたり
モデル詳細を見る →
recraft logorecraft

Recraft: Recraft V3

Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. Supports the following imageconfig parameters:...

コンテキスト
66K
入力
text · image
出力
image
画像出力: $0.04 画像1枚あたり
モデル詳細を見る →
google logogoogle

Google: Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

コンテキスト
1.0M
入力
text · image · video · file · audio
出力
text
入力: $0.25出力: $1.5100万トークンあたり
モデル詳細を見る →
google logogoogle

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

コンテキスト
1.0M
入力
text · image · video · file · audio
出力
text
入力: $0.125出力: $0.75100万トークンあたり
モデル詳細を見る →
openai logoopenai

OpenAI: GPT Chat Latest

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

コンテキスト
400K
入力
text · image · file
出力
text
入力: $5出力: $30100万トークンあたり
モデル詳細を見る →
google logogoogle

Google: Chirp 3

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...

コンテキスト
情報なし
入力
audio
出力
transcription
音声の長さ: $0.000267 1秒あたり
モデル詳細を見る →
openai logoopenai

OpenAI: GPT-4o Mini Transcribe

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...

コンテキスト
128K
入力
audio
出力
transcription
入力: $1.25 100万トークンあたり出力: $5 100万トークンあたり
モデル詳細を見る →
openai logoopenai

OpenAI: Whisper Large V3

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...

コンテキスト
情報なし
入力
audio
出力
transcription
音声の長さ: $0.000008 1秒あたり
モデル詳細を見る →
openai logoopenai

OpenAI: Whisper Large V3 Turbo

Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...

コンテキスト
情報なし
入力
audio
出力
transcription
音声の長さ: $0.000003 1秒あたり
モデル詳細を見る →
x-ai logox-ai

SpaceXAI: Grok 4.3

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $1.25出力: $2.5100万トークンあたり
モデル詳細を見る →
x-ai logox-ai

SpaceXAI: Grok 4.3 (batch)

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $1出力: $2100万トークンあたり
モデル詳細を見る →
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

コンテキスト
262K
入力
text · image · file
出力
text
入力: $1.5出力: $7.5100万トークンあたり
モデル詳細を見る →
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

コンテキスト
262K
入力
text · image · file
出力
text
入力: $0.75出力: $3.75100万トークンあたり
モデル詳細を見る →
kwaivgi logokwaivgi

Kling: Video v3.0 Pro

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

コンテキスト
情報なし
入力
text · image
出力
video
Video (with audio): $0.168 1秒あたりVideo (no audio): $0.112 1秒あたり
モデル詳細を見る →
kwaivgi logokwaivgi

Kling: Video v3.0 Standard

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

コンテキスト
情報なし
入力
text · image
出力
video
Video (with audio): $0.126 1秒あたりVideo (no audio): $0.084 1秒あたり
モデル詳細を見る →
nvidia logonvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

コンテキスト
256K
入力
text · audio · image · video
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る →
openai logoopenai

OpenAI: Whisper 1

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...

コンテキスト
情報なし
入力
audio
出力
transcription
音声の長さ: $0.0001 1秒あたり
モデル詳細を見る →
openai logoopenai

OpenAI: GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

コンテキスト
128K
入力
audio
出力
transcription
入力: $2.5 100万トークンあたり出力: $10 100万トークンあたり
モデル詳細を見る →
~anthropic logo~anthropic

Anthropic: Claude Haiku Latest

This model always redirects to the latest model in the Claude Haiku family.

コンテキスト
200K
入力
text · image · file
出力
text
入力: $1出力: $5100万トークンあたり
モデル詳細を見る →
~openai logo~openai

OpenAI: GPT Mini Latest

This model always redirects to the latest model in the GPT Mini family.

コンテキスト
400K
入力
file · image · text
出力
text
入力: $0.75出力: $4.5100万トークンあたり
モデル詳細を見る →
~google logo~google

Google: Gemini Pro Latest

This model always redirects to the latest model in the Gemini Pro family.

コンテキスト
1.0M
入力
audio · file · image · text · video
出力
text
入力: $2出力: $12100万トークンあたり
モデル詳細を見る →
~moonshotai logo~moonshotai

MoonshotAI: Kimi Latest

This model always redirects to the latest model in the Kimi family.

コンテキスト
1.0M
入力
text · image · video
出力
text
入力: $0.8582出力: $11.1062100万トークンあたり
モデル詳細を見る →
~google logo~google

Google: Gemini Flash Latest

This model always redirects to the latest model in the Gemini Flash family.

コンテキスト
1.0M
入力
text · image · video · file · audio
出力
text
入力: $0.75出力: $3.75100万トークンあたり
モデル詳細を見る →