モデルレーダー · OPENROUTER
AIモデルを、同じ基準で比較。
OpenRouterで現在公開されているすべてのモデルを検索し、モダリティ、コンテキスト、料金、設定、プロバイダー情報を比較できます。
- カタログ範囲
- 公開モデルすべて
- モデル
- 534
- プロバイダー
- 75
- 最終同期
- 2026/08/31
OpenAI: GPT-4o Mini Transcribe
GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...
- コンテキスト
- 128K
- 入力
- audio
- 出力
- transcription
OpenAI: Whisper Large V3
Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...
- コンテキスト
- 情報なし
- 入力
- audio
- 出力
- transcription
OpenAI: Whisper Large V3 Turbo
Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...
- コンテキスト
- 情報なし
- 入力
- audio
- 出力
- transcription
SpaceXAI: Grok 4.3
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
- コンテキスト
- 1M
- 入力
- text · image · file
- 出力
- text
IBM: Granite 4.1 8B
Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...
- コンテキスト
- 131K
- 入力
- text
- 出力
- text
Mistral: Mistral Medium 3.5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- コンテキスト
- 262K
- 入力
- text · image · file
- 出力
- text
Mistral: Mistral Medium 3.5 (batch)
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- コンテキスト
- 262K
- 入力
- text · image · file
- 出力
- text
Kling: Video v3.0 Pro
Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...
- コンテキスト
- 情報なし
- 入力
- text · image
- 出力
- video
Kling: Video v3.0 Standard
Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...
- コンテキスト
- 情報なし
- 入力
- text · image
- 出力
- video
NVIDIA: Nemotron 3 Nano Omni (free)
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
- コンテキスト
- 256K
- 入力
- text · audio · image · video
- 出力
- text
OpenAI: Whisper 1
Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...
- コンテキスト
- 情報なし
- 入力
- audio
- 出力
- transcription
OpenAI: GPT-4o Transcribe
GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.
- コンテキスト
- 128K
- 入力
- audio
- 出力
- transcription
Anthropic Claude Haiku Latest
This model always redirects to the latest model in the Anthropic Claude Haiku family.
- コンテキスト
- 200K
- 入力
- text · image · file
- 出力
- text
OpenAI GPT Mini Latest
This model always redirects to the latest model in the OpenAI GPT Mini family.
- コンテキスト
- 400K
- 入力
- file · image · text
- 出力
- text
Google Gemini Pro Latest
This model always redirects to the latest model in the Google Gemini Pro family.
- コンテキスト
- 1.0M
- 入力
- audio · file · image · text · video
- 出力
- text
MoonshotAI Kimi Latest
This model always redirects to the latest model in the MoonshotAI Kimi family.
- コンテキスト
- 1.0M
- 入力
- text · image · video
- 出力
- text
Google Gemini Flash Latest
This model always redirects to the latest model in the Google Gemini Flash family.
- コンテキスト
- 1.0M
- 入力
- text · image · video · file · audio
- 出力
- text
Anthropic Claude Sonnet Latest
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
- コンテキスト
- 1M
- 入力
- text · image · file
- 出力
- text
OpenAI GPT Latest
This model always redirects to the latest model in the OpenAI GPT family.
- コンテキスト
- 1.1M
- 入力
- file · image · text
- 出力
- text
Qwen: Qwen3.5 Plus 2026-04-20
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
- コンテキスト
- 1M
- 入力
- text · image · video
- 出力
- text
Qwen: Qwen3.6 Flash
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
- コンテキスト
- 1M
- 入力
- text · image · video
- 出力
- text
Qwen: Qwen3.6 35B A3B
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
- コンテキスト
- 262K
- 入力
- text · image · video
- 出力
- text
Qwen: Qwen3.6 Max Preview
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
- コンテキスト
- 262K
- 入力
- text
- 出力
- text
Qwen: Qwen3.6 27B
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
- コンテキスト
- 262K
- 入力
- text · image · video
- 出力
- text