モデルレーダー · OPENROUTER

AIモデルを、同じ基準で比較。

OpenRouterで現在公開されているすべてのモデルを検索し、モダリティ、コンテキスト、料金、設定、プロバイダー情報を比較できます。

カタログ範囲
公開モデルすべて
モデル
534
プロバイダー
75
最終同期
2026/08/31
534 モデル
anthropic logoanthropic

Anthropic: Claude Fable 5 (batch)

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $5.00出力: $25.00100万トークンあたり
モデル詳細を見る
nex-agi logonex-agi

Nex AGI: Nex-N2-Pro

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

コンテキスト
262K
入力
text · image
出力
text
入力: $0.250出力: $1.00100万トークンあたり
モデル詳細を見る
sourceful logosourceful

Sourceful: Riverflow V2.5 Pro

Riverflow V2.5 Pro is the most powerful variant of Sourceful's Riverflow 2.5 lineup, best for top-tier control and quality-sensitive outputs. The Riverflow 2.5 series is a unified text-to-image and image-to-image...

コンテキスト
33K
入力
text · image
出力
image
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
sourceful logosourceful

Sourceful: Riverflow V2.5 Fast

Riverflow V2.5 Fast is the speed-optimized variant of Sourceful's Riverflow 2.5 lineup, best for production deployments and latency-critical workflows. The Riverflow 2.5 series is a unified text-to-image and image-to-image family...

コンテキスト
33K
入力
text · image
出力
image
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3.5 Content Safety (free)

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

コンテキスト
128K
入力
text · image
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Ultra

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

コンテキスト
262K
入力
text
出力
text
入力: $0.500出力: $2.20100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Ultra (batch)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

コンテキスト
512K
入力
text
出力
text
入力: $0.600出力: $3.60100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Nemotron 3 Ultra (free)

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

コンテキスト
1M
入力
text
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.7 Plus

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

コンテキスト
1M
入力
text · image
出力
text
入力: $0.320出力: $1.28100万トークンあたり
モデル詳細を見る
microsoft logomicrosoft

Microsoft: MAI-Voice-2

MAI-Voice-2 is an expressive text-to-speech model from Microsoft. It is suited for conversational assistants, media narration, accessibility, education, and other long-form voice applications. It supports 15 languages across 18 locales,...

コンテキスト
情報なし
入力
text
出力
speech
入力: $22.00出力: 無料100万トークンあたり
モデル詳細を見る
microsoft logomicrosoft

Microsoft: MAI-Transcribe 1.5

MAI-Transcribe 1.5 is a multilingual speech-to-text model from Microsoft AI. It is suited for captions, call transcription, subtitling, accessibility, and other voice-enabled applications, with reliable transcription across 43 languages, diverse...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $360000.00出力: 無料100万トークンあたり
モデル詳細を見る
microsoft logomicrosoft

Microsoft: MAI-Image-2.5

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

コンテキスト
4K
入力
text · image
出力
image
入力: $5.00出力: 無料100万トークンあたり
モデル詳細を見る
minimax logominimax

MiniMax: MiniMax M3

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

コンテキスト
1.0M
入力
text · image · video
出力
text
入力: $0.300出力: $1.20100万トークンあたり
モデル詳細を見る
minimax logominimax

MiniMax: MiniMax M3 (batch)

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

コンテキスト
524K
入力
text · image · video
出力
text
入力: $0.300出力: $1.20100万トークンあたり
モデル詳細を見る
minimax logominimax

MiniMax: MiniMax M3 (free)

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

コンテキスト
1.0M
入力
text · image · video
出力
text
入力: 無料出力: 無料100万トークンあたり
モデル詳細を見る
stepfun logostepfun

StepFun: Step 3.7 Flash

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

コンテキスト
262K
入力
text · image · video
出力
text
入力: $0.200出力: $1.15100万トークンあたり
モデル詳細を見る
anthropic logoanthropic

Anthropic: Claude Opus 4.8 (Fast)

Fast-mode variant of Opus 4.8 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

コンテキスト
1M
入力
text · image · file
出力
text
入力: $10.00出力: $50.00100万トークンあたり
モデル詳細を見る
anthropic logoanthropic

Anthropic: Claude Opus 4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $5.00出力: $25.00100万トークンあたり
モデル詳細を見る
anthropic logoanthropic

Anthropic: Claude Opus 4.8 (batch)

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

コンテキスト
1M
入力
text · image · file
出力
text
入力: $2.50出力: $12.50100万トークンあたり
モデル詳細を見る
nvidia logonvidia

NVIDIA: Parakeet TDT 0.6B v3

Parakeet TDT 0.6B v3 is NVIDIA's 600M-parameter multilingual speech-to-text model built on the FastConformer-TDT architecture. Trained on the Granary dataset (670,000+ hours of audio), it supports automatic language detection across...

コンテキスト
情報なし
入力
audio
出力
transcription
入力: $1500.00出力: 無料100万トークンあたり
モデル詳細を見る
qwen logoqwen

Qwen: Qwen3.7 Max

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

コンテキスト
1M
入力
text
出力
text
入力: $1.48出力: $4.42100万トークンあたり
モデル詳細を見る
x-ai logox-ai

SpaceXAI: Grok Build 0.1

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

コンテキスト
256K
入力
text · image · file
出力
text
入力: $1.00出力: $2.00100万トークンあたり
モデル詳細を見る
google logogoogle

Google: Gemini Embedding 2

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports...

コンテキスト
8K
入力
text · image · file · audio · video
出力
embeddings
入力: $0.200出力: 無料100万トークンあたり
モデル詳細を見る
google logogoogle

Google: Gemini Embedding 2 (batch)

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports...

コンテキスト
8K
入力
text · image · file · audio · video
出力
embeddings
入力: $0.100出力: 無料100万トークンあたり
モデル詳細を見る