模型雷达 · OPENROUTER
把 AI 模型放在一起比较。
收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。
- 目录范围
- 全部公开
- 个模型
- 534
- 个提供方
- 75
- 最近同步
- 2026年8月31日
OpenAI: GPT-4o Mini Transcribe
GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...
- 上下文
- 128K
- 输入
- audio
- 输出
- transcription
OpenAI: Whisper Large V3
Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...
- 上下文
- 暂未提供
- 输入
- audio
- 输出
- transcription
OpenAI: Whisper Large V3 Turbo
Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...
- 上下文
- 暂未提供
- 输入
- audio
- 输出
- transcription
SpaceXAI: Grok 4.3
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
- 上下文
- 1M
- 输入
- text · image · file
- 输出
- text
IBM: Granite 4.1 8B
Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...
- 上下文
- 131K
- 输入
- text
- 输出
- text
Mistral: Mistral Medium 3.5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- 上下文
- 262K
- 输入
- text · image · file
- 输出
- text
Mistral: Mistral Medium 3.5 (batch)
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- 上下文
- 262K
- 输入
- text · image · file
- 输出
- text
Kling: Video v3.0 Pro
Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
Kling: Video v3.0 Standard
Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
NVIDIA: Nemotron 3 Nano Omni (free)
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
- 上下文
- 256K
- 输入
- text · audio · image · video
- 输出
- text
OpenAI: Whisper 1
Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...
- 上下文
- 暂未提供
- 输入
- audio
- 输出
- transcription
OpenAI: GPT-4o Transcribe
GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.
- 上下文
- 128K
- 输入
- audio
- 输出
- transcription
Anthropic Claude Haiku Latest
This model always redirects to the latest model in the Anthropic Claude Haiku family.
- 上下文
- 200K
- 输入
- text · image · file
- 输出
- text
OpenAI GPT Mini Latest
This model always redirects to the latest model in the OpenAI GPT Mini family.
- 上下文
- 400K
- 输入
- file · image · text
- 输出
- text
Google Gemini Pro Latest
This model always redirects to the latest model in the Google Gemini Pro family.
- 上下文
- 1.0M
- 输入
- audio · file · image · text · video
- 输出
- text
MoonshotAI Kimi Latest
This model always redirects to the latest model in the MoonshotAI Kimi family.
- 上下文
- 1.0M
- 输入
- text · image · video
- 输出
- text
Google Gemini Flash Latest
This model always redirects to the latest model in the Google Gemini Flash family.
- 上下文
- 1.0M
- 输入
- text · image · video · file · audio
- 输出
- text
Anthropic Claude Sonnet Latest
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
- 上下文
- 1M
- 输入
- text · image · file
- 输出
- text
OpenAI GPT Latest
This model always redirects to the latest model in the OpenAI GPT family.
- 上下文
- 1.1M
- 输入
- file · image · text
- 输出
- text
Qwen: Qwen3.5 Plus 2026-04-20
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
- 上下文
- 1M
- 输入
- text · image · video
- 输出
- text
Qwen: Qwen3.6 Flash
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
- 上下文
- 1M
- 输入
- text · image · video
- 输出
- text
Qwen: Qwen3.6 35B A3B
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
- 上下文
- 262K
- 输入
- text · image · video
- 输出
- text
Qwen: Qwen3.6 Max Preview
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
- 上下文
- 262K
- 输入
- text
- 输出
- text
Qwen: Qwen3.6 27B
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
- 上下文
- 262K
- 输入
- text · image · video
- 输出
- text