モデルレーダー · OPENROUTER

AIモデルを、同じ基準で比較。

OpenRouterで現在公開されているすべてのモデルを検索し、モダリティ、コンテキスト、料金、設定、プロバイダー情報を比較できます。

カタログ範囲
公開モデルすべて
モデル
534
プロバイダー
75
最終同期
2026/08/31
534 モデル
openai logoopenai

OpenAI: GPT-4o Mini Transcribe

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...

コンテキスト
128K
入力
audio
出力
transcription
入力: $1.25出力: $5.00100万トークンあたり
モデル詳細を見る
openai logoopenai

OpenAI: Whisper Large V3

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $7.50出力: 無料100万トークンあたり
モデル詳細を見る
openai logoopenai

OpenAI: Whisper Large V3 Turbo

Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $3.33出力: 無料100万トークンあたり
モデル詳細を見る
x-ai logox-ai

SpaceXAI: Grok 4.3

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $1.25出力: $2.50100万トークンあたり
モデル詳細を見る
ibm-granite logoibm-granite

IBM: Granite 4.1 8B

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

コンテキスト
131K
入力
text
出力
text
入力: $0.050出力: $0.100100万トークンあたり
モデル詳細を見る
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

コンテキスト
262K
入力
text · image · file
出力
text
入力: $1.50出力: $7.50100万トークンあたり
モデル詳細を見る
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

コンテキスト
262K
入力
text · image · file
出力
text
入力: $0.750出力: $3.75100万トークンあたり
モデル詳細を見る
kwaivgi logokwaivgi

Kling: Video v3.0 Pro

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

コンテキスト
情報なし
入力
text · image
出力
video
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
kwaivgi logokwaivgi

Kling: Video v3.0 Standard

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

コンテキスト
情報なし
入力
text · image
出力
video
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

コンテキスト
256K
入力
text · audio · image · video
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
openai logoopenai

OpenAI: Whisper 1

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $6000.00出力: 無料100万トークンあたり
モデル詳細を見る
openai logoopenai

OpenAI: GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

コンテキスト
128K
入力
audio
出力
transcription
入力: $2.50出力: $10.00100万トークンあたり
モデル詳細を見る
~anthropic logo~anthropic

Anthropic Claude Haiku Latest

This model always redirects to the latest model in the Anthropic Claude Haiku family.

コンテキスト
200K
入力
text · image · file
出力
text
入力: $1.00出力: $5.00100万トークンあたり
モデル詳細を見る
~openai logo~openai

OpenAI GPT Mini Latest

This model always redirects to the latest model in the OpenAI GPT Mini family.

コンテキスト
400K
入力
file · image · text
出力
text
入力: $0.750出力: $4.50100万トークンあたり
モデル詳細を見る
~google logo~google

Google Gemini Pro Latest

This model always redirects to the latest model in the Google Gemini Pro family.

コンテキスト
1.0M
入力
audio · file · image · text · video
出力
text
入力: $2.00出力: $12.00100万トークンあたり
モデル詳細を見る
~moonshotai logo~moonshotai

MoonshotAI Kimi Latest

This model always redirects to the latest model in the MoonshotAI Kimi family.

コンテキスト
1.0M
入力
text · image · video
出力
text
入力: $2.55出力: $12.75100万トークンあたり
モデル詳細を見る
~google logo~google

Google Gemini Flash Latest

This model always redirects to the latest model in the Google Gemini Flash family.

コンテキスト
1.0M
入力
text · image · video · file · audio
出力
text
入力: $0.750出力: $3.75100万トークンあたり
モデル詳細を見る
~anthropic logo~anthropic

Anthropic Claude Sonnet Latest

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

コンテキスト
1M
入力
text · image · file
出力
text
入力: $2.00出力: $10.00100万トークンあたり
モデル詳細を見る
~openai logo~openai

OpenAI GPT Latest

This model always redirects to the latest model in the OpenAI GPT family.

コンテキスト
1.1M
入力
file · image · text
出力
text
入力: $2.00出力: $10.00100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.5 Plus 2026-04-20

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

コンテキスト
1M
入力
text · image · video
出力
text
入力: $0.300出力: $1.80100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.6 Flash

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

コンテキスト
1M
入力
text · image · video
出力
text
入力: $0.188出力: $1.13100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.6 35B A3B

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

コンテキスト
262K
入力
text · image · video
出力
text
入力: $0.100出力: $0.900100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.6 Max Preview

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

コンテキスト
262K
入力
text
出力
text
入力: $1.03出力: $6.16100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.6 27B

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

コンテキスト
262K
入力
text · image · video
出力
text
入力: $0.600出力: $3.60100万トークンあたり
モデル詳細を見る