模型雷达 · OPENROUTER

把 AI 模型放在一起比较。

收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。

目录范围
全部公开
个模型
534
个提供方
75
最近同步
2026年8月31日
534 个模型
minimax logominimax

MiniMax: MiniMax M2.7 (free)

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

上下文
197K
输入
text
输出
text
输入: 免费输出: 免费每百万 Token
查看模型详情
openai logoopenai

OpenAI: GPT-5.4 Nano

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

上下文
400K
输入
file · image · text
输出
text
输入: $0.200输出: $1.25每百万 Token
查看模型详情
openai logoopenai

OpenAI: GPT-5.4 Mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

上下文
400K
输入
file · image · text
输出
text
输入: $0.750输出: $4.50每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Mistral Small 4

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

上下文
262K
输入
text · image
输出
text
输入: $0.150输出: $0.600每百万 Token
查看模型详情
mistralai logomistralai

Mistral: Mistral Small 4 (batch)

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

上下文
262K
输入
text · image
输出
text
输入: $0.150输出: $0.600每百万 Token
查看模型详情
perplexity logoperplexity

Perplexity: Embed V1 4B

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval...

上下文
32K
输入
text
输出
embeddings
输入: $0.030输出: 免费每百万 Token
查看模型详情
perplexity logoperplexity

Perplexity: Embed V1 0.6B

pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency...

上下文
32K
输入
text
输出
embeddings
输入: $0.0040输出: 免费每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Nemotron 3 Super

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

上下文
1M
输入
text
输出
text
输入: $0.085输出: $0.400每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Nemotron 3 Super (free)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

上下文
262K
输入
text
输出
text
输入: 免费输出: 免费每百万 Token
查看模型详情
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Lite

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

上下文
262K
输入
text · image · video
输出
text
输入: $0.250输出: $2.00每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-9B

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

上下文
262K
输入
text · image · video
输出
text
输入: $0.100输出: $0.150每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-9B (batch)

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

上下文
262K
输入
text · image · video
输出
text
输入: $0.170输出: $0.250每百万 Token
查看模型详情
openai logoopenai

OpenAI: GPT-5.4 Pro

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

上下文
1.1M
输入
text · image · file
输出
text
输入: $30.00输出: $180.00每百万 Token
查看模型详情
openai logoopenai

OpenAI: GPT-5.4

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

上下文
1.1M
输入
text · image · file
输出
text
输入: $2.50输出: $15.00每百万 Token
查看模型详情
inception logoinception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

上下文
128K
输入
text
输出
text
输入: $0.250输出: $0.750每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

上下文
1.0M
输入
text · image · video · file · audio
输出
text
输入: $0.250输出: $1.50每百万 Token
查看模型详情
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Mini

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

上下文
262K
输入
text · image · video
输出
text
输入: $0.100输出: $0.400每百万 Token
查看模型详情
google logogoogle

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

上下文
66K
输入
image · text
输出
image · text
输入: $0.500输出: $3.00每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

上下文
262K
输入
text · image · video
输出
text
输入: $0.250输出: $1.25每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-27B

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

上下文
262K
输入
text · image · video
输出
text
输入: $0.195输出: $1.56每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-122B-A10B

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

上下文
262K
输入
text · image · video
输出
text
输入: $0.290输出: $2.40每百万 Token
查看模型详情
qwen logoqwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

上下文
1M
输入
text · image · video
输出
text
输入: $0.065输出: $0.260每百万 Token
查看模型详情
google logogoogle

Google: Gemini 3.1 Pro Preview Custom Tools

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

上下文
1.0M
输入
text · audio · image · video · file
输出
text
输入: $2.00输出: $12.00每百万 Token
查看模型详情
nvidia logonvidia

NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

上下文
131K
输入
text · image
输出
embeddings
输入: 免费输出: 免费每百万 Token
查看模型详情