模型雷达 · OPENROUTER

把 AI 模型放在一起比较。

收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。

目录范围
全部公开
个模型
631
个提供方
86
最近同步
2026年9月29日
631 个模型
recraft logorecraft

Recraft: Recraft V4 Pro

Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios, double the resolution of...

上下文
66K
输入
text · image
输出
image
图片输出: $0.25 每张图片
查看模型详情 →
recraft logorecraft

Recraft: Recraft V4

Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,...

上下文
66K
输入
text · image
输出
image
图片输出: $0.04 每张图片
查看模型详情 →
recraft logorecraft

Recraft: Recraft V3

Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. Supports the following imageconfig parameters:...

上下文
66K
输入
text · image
输出
image
图片输出: $0.04 每张图片
查看模型详情 →
google logogoogle

Google: Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.25输出: $1.5每百万 Token
查看模型详情 →
google logogoogle

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.125输出: $0.75每百万 Token
查看模型详情 →
openai logoopenai

OpenAI: GPT Chat Latest

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

上下文
400K
输入
text · image · file
输出
text
输入: $5输出: $30每百万 Token
查看模型详情 →
google logogoogle

Google: Chirp 3

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...

上下文
暂未提供
输入
audio
输出
transcription
音频时长: $0.000267 每秒
查看模型详情 →
openai logoopenai

OpenAI: GPT-4o Mini Transcribe

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...

上下文
128K
输入
audio
输出
transcription
输入: $1.25 每百万 Token输出: $5 每百万 Token
查看模型详情 →
openai logoopenai

OpenAI: Whisper Large V3

Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...

上下文
暂未提供
输入
audio
输出
transcription
音频时长: $0.000008 每秒
查看模型详情 →
openai logoopenai

OpenAI: Whisper Large V3 Turbo

Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...

上下文
暂未提供
输入
audio
输出
transcription
音频时长: $0.000003 每秒
查看模型详情 →
x-ai logox-ai

SpaceXAI: Grok 4.3

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

上下文
1M
输入
text · image · file
输出
text
输入: $1.25输出: $2.5每百万 Token
查看模型详情 →
x-ai logox-ai

SpaceXAI: Grok 4.3 (batch)

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

上下文
1M
输入
text · image · file
输出
text
输入: $1输出: $2每百万 Token
查看模型详情 →
mistralai logomistralai

Mistral: Mistral Medium 3.5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

上下文
262K
输入
text · image · file
输出
text
输入: $1.5输出: $7.5每百万 Token
查看模型详情 →
mistralai logomistralai

Mistral: Mistral Medium 3.5 (batch)

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

上下文
262K
输入
text · image · file
输出
text
输入: $0.75输出: $3.75每百万 Token
查看模型详情 →
kwaivgi logokwaivgi

Kling: Video v3.0 Pro

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...

上下文
暂未提供
输入
text · image
输出
video
Video (with audio): $0.168 每秒Video (no audio): $0.112 每秒
查看模型详情 →
kwaivgi logokwaivgi

Kling: Video v3.0 Standard

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...

上下文
暂未提供
输入
text · image
输出
video
Video (with audio): $0.126 每秒Video (no audio): $0.084 每秒
查看模型详情 →
nvidia logonvidia

NVIDIA: Nemotron 3 Nano Omni (free)

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

上下文
256K
输入
text · audio · image · video
输出
text
输入: 免费输出: 免费每百万 Token
查看模型详情 →
openai logoopenai

OpenAI: Whisper 1

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...

上下文
暂未提供
输入
audio
输出
transcription
音频时长: $0.0001 每秒
查看模型详情 →
openai logoopenai

OpenAI: GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

上下文
128K
输入
audio
输出
transcription
输入: $2.5 每百万 Token输出: $10 每百万 Token
查看模型详情 →
~anthropic logo~anthropic

Anthropic: Claude Haiku Latest

This model always redirects to the latest model in the Claude Haiku family.

上下文
200K
输入
text · image · file
输出
text
输入: $1输出: $5每百万 Token
查看模型详情 →
~openai logo~openai

OpenAI: GPT Mini Latest

This model always redirects to the latest model in the GPT Mini family.

上下文
400K
输入
file · image · text
输出
text
输入: $0.75输出: $4.5每百万 Token
查看模型详情 →
~google logo~google

Google: Gemini Pro Latest

This model always redirects to the latest model in the Gemini Pro family.

上下文
1.0M
输入
audio · file · image · text · video
输出
text
输入: $2输出: $12每百万 Token
查看模型详情 →
~moonshotai logo~moonshotai

MoonshotAI: Kimi Latest

This model always redirects to the latest model in the Kimi family.

上下文
1.0M
输入
text · image · video
输出
text
输入: $0.8582输出: $11.1062每百万 Token
查看模型详情 →
~google logo~google

Google: Gemini Flash Latest

This model always redirects to the latest model in the Gemini Flash family.

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.75输出: $3.75每百万 Token
查看模型详情 →