モデルレーダー · OPENROUTER
AIモデルを、同じ基準で比較。
OpenRouterで現在公開されているすべてのモデルを検索し、モダリティ、コンテキスト、料金、設定、プロバイダー情報を比較できます。
- カタログ範囲
- 公開モデルすべて
- モデル
- 534
- プロバイダー
- 75
- 最終同期
- 2026/08/31
Meta: Muse Glimmer 30B (batch)
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
- コンテキスト
- 131K
- 入力
- text · image
- 出力
- text
ByteDance: Seedance 2.5
Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...
- コンテキスト
- 情報なし
- 入力
- text · image · video · audio
- 出力
- video
OpenAI: GPT Transcribe
GPT Transcribe is a high-accuracy speech-to-text model from OpenAI. It is suited for recorded audio, streamed file transcription, and committed Realtime turns, with free-form context, keyword hints, and multiple language...
- コンテキスト
- 情報なし
- 入力
- audio
- 出力
- transcription
Meta: Muse Spark 1.2
Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...
- コンテキスト
- 1.0M
- 入力
- text · image · video · file · audio
- 出力
- text
Qwen: Qwen Image 3
Qwen Image 3 is a unified image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with a richer world...
- コンテキスト
- 66K
- 入力
- text · image
- 出力
- image
Qwen: Qwen Image 3 Pro
Qwen Image 3 Pro is an image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with richer world knowledge...
- コンテキスト
- 66K
- 入力
- text · image
- 出力
- image
Black Forest Labs: FLUX.3 Video
FLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled...
- コンテキスト
- 情報なし
- 入力
- text · image · video
- 出力
- video
Qwen: Qwen3.8 Max
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...
- コンテキスト
- 1M
- 入力
- text · image · video
- 出力
- text
DeepSeek V4 Flash Latest
This model always redirects to the latest model in the DeepSeek V4 Flash family.
- コンテキスト
- 1.3M
- 入力
- text
- 出力
- text
DeepSeek: DeepSeek V4 Flash 0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
- コンテキスト
- 1.3M
- 入力
- text
- 出力
- text
DeepSeek: DeepSeek V4 Flash 0731 (batch)
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
- コンテキスト
- 1.0M
- 入力
- text
- 出力
- text
Thinking Machines: Inkling Small
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
- コンテキスト
- 1.0M
- 入力
- text · image · audio
- 出力
- text
Thinking Machines: Inkling Small (batch)
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
- コンテキスト
- 524K
- 入力
- text · image · audio
- 出力
- text
Thinking Machines: Inkling Small (free)
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
- コンテキスト
- 1.0M
- 入力
- text · image · audio
- 出力
- text
MiniMax: H3
MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...
- コンテキスト
- 情報なし
- 入力
- text · image · video · audio
- 出力
- video
Fish Audio: Transcribe 1
Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.
- コンテキスト
- 情報なし
- 入力
- audio
- 出力
- transcription
Fish Audio: S1
S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...
- コンテキスト
- 情報なし
- 入力
- text
- 出力
- speech
Fish Audio: S2 Pro
S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.
- コンテキスト
- 情報なし
- 入力
- text
- 出力
- speech
Fish Audio: S2.1 Pro Free (free)
S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...
- コンテキスト
- 情報なし
- 入力
- text
- 出力
- speech
Fish Audio: S2.1 Pro
S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...
- コンテキスト
- 情報なし
- 入力
- text
- 出力
- speech
Runway: Aleph 2.0
Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....
- コンテキスト
- 情報なし
- 入力
- text · image · video
- 出力
- video
Runway: Gen-4.5
Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....
- コンテキスト
- 情報なし
- 入力
- text · image
- 出力
- video
Qwen: Qwen3.7 Flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
- コンテキスト
- 1M
- 入力
- text · image · video
- 出力
- text
VoyageAI by MongoDB: rerank-2.5-lite
rerank-2.5-lite is a reranker optimized for both latency and quality, delivering a 7.16% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5...
- コンテキスト
- 32K
- 入力
- text
- 出力
- rerank