模型雷达 · OPENROUTER

把 AI 模型放在一起比较。

收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。

目录范围
全部公开
个模型
534
个提供方
75
最近同步
2026年8月31日
534 个模型
google logogoogle

Google: Gemini 3.5 Flash

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $1.50输出: $9.00每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.5 Flash (batch)

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.750输出: $4.50每百万 Token
查看模型详情
x-ai logox-ai

SpaceXAI: Grok Imagine Video

Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...

上下文
暂未提供
输入
text · image
输出
video
输入: 免费输出: 免费每百万 Token
查看模型详情
x-ai logox-ai

SpaceXAI: Grok Imagine Image Quality

Grok Imagine Image Quality is SpaceXAI's fast, high-fidelity image generation and editing model. It accepts text prompts and optional reference images, producing photorealistic outputs at 1K or 2K across a...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

上下文
暂未提供
输入
audio
输出
transcription
输入: $3000.00输出: 免费每百万 Token
查看模型详情
x-ai logox-ai

SpaceXAI: Grok Voice TTS 1.0

Grok Voice TTS 1.0 is a text-to-speech model from SpaceXAI. It converts text into spoken audio across 20+ languages with automatic language detection, and offers five built-in voices (Eve, Ara,...

上下文
15K
输入
text
输出
speech
输入: $15.00输出: 免费每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3 ASR Flash

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —...

上下文
暂未提供
输入
audio
输出
transcription
输入: $35.00输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1 Pro Vector

Recraft V4.1 Pro Vector is the vector (SVG) variant of Recraft V4.1 Pro, tuned for high aesthetics. It supports text and image inputs and produces higher-resolution SVG image output across...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1 Vector

Recraft V4.1 Vector is the vector (SVG) variant of Recraft V4.1, tuned for high aesthetics. It supports text and image inputs and produces SVG image output across multiple aspect ratios,...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1 Utility Pro

Recraft V4.1 Utility Pro is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios — double...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1 Utility

Recraft V4.1 Utility is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with typical generation...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1 Pro

Recraft V4.1 Pro is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4.1

Recraft V4.1 is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4 Pro Vector

Recraft V4 Pro Vector is the vector (SVG) variant of Recraft V4 Pro. It supports text and image inputs and produces vector image output across multiple aspect ratios at the...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4 Vector

Recraft V4 Vector is the vector (SVG) variant of Recraft V4. It supports text and image inputs and produces vector image output across multiple aspect ratios. Compared to the raster...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
anthropic logoanthropic

Anthropic: Claude Opus 4.7 (Fast)

Fast-mode variant of Opus 4.7 - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

上下文
1M
输入
text · image · file
输出
text
输入: $30.00输出: $150.00每百万 Token
查看模型详情
perceptron logoperceptron

Perceptron: Perceptron Mk1

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning. It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

上下文
33K
输入
text · image · video
输出
text
输入: $0.150输出: $1.50每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4 Pro

Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios, double the resolution of...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V4

Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
recraft logorecraft

Recraft: Recraft V3

Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. Supports the following imageconfig parameters:...

上下文
66K
输入
text · image
输出
image
输入: 免费输出: 免费每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.250输出: $1.50每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.1 Flash Lite (batch)

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.125输出: $0.750每百万 Token
查看模型详情
openai logoopenai

OpenAI: GPT Chat Latest

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

上下文
400K
输入
text · image · file
输出
text
输入: $5.00输出: $30.00每百万 Token
查看模型详情
google logogoogle

Google: Chirp 3

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...

上下文
暂未提供
输入
audio
输出
transcription
输入: $16000.00输出: 免费每百万 Token
查看模型详情