模型雷达 · OPENROUTER
把 AI 模型放在一起比较。
收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。
- 目录范围
- 全部公开
- 个模型
- 534
- 个提供方
- 75
- 最近同步
- 2026年8月31日
ByteDance: Seedance 2.0
Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...
- 上下文
- 暂未提供
- 输入
- text · image · video · audio
- 输出
- video
ByteDance: Seedance 2.0 Fast
Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...
- 上下文
- 暂未提供
- 输入
- text · image · video · audio
- 输出
- video
Z.ai: GLM 5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- 上下文
- 205K
- 输入
- text
- 输出
- text
Cohere: Rerank 4 Pro
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- 上下文
- 33K
- 输入
- text
- 输出
- rerank
Cohere: Rerank 4 Fast
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- 上下文
- 33K
- 输入
- text
- 输出
- rerank
Cohere: Rerank v3.5
Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...
- 上下文
- 4K
- 输入
- text
- 输出
- rerank
Google: Gemma 4 26B A4B
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- 上下文
- 262K
- 输入
- image · text · video
- 输出
- text
Google: Gemma 4 26B A4B (free)
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- 上下文
- 262K
- 输入
- image · text · video
- 输出
- text
Google: Gemma 4 31B
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- 上下文
- 262K
- 输入
- image · text · video
- 输出
- text
Google: Gemma 4 31B (batch)
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- 上下文
- 262K
- 输入
- image · text · video
- 输出
- text
Google: Gemma 4 31B (free)
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- 上下文
- 262K
- 输入
- image · text · video
- 输出
- text
Qwen: Qwen3.6 Plus
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
- 上下文
- 1M
- 输入
- text · image · video
- 输出
- text
Arcee AI: Trinity Large Thinking
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
- 上下文
- 262K
- 输入
- text
- 输出
- text
SpaceXAI: Grok 4.20 Multi-Agent
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
- 上下文
- 2M
- 输入
- text · image · file
- 输出
- text
SpaceXAI: Grok 4.20
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
- 上下文
- 2M
- 输入
- text · image · file
- 输出
- text
Google: Lyria 3 Pro Preview
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
- 上下文
- 1.0M
- 输入
- text · image
- 输出
- text · audio
Google: Lyria 3 Clip Preview
30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...
- 上下文
- 1.0M
- 输入
- text · image
- 输出
- text · audio
Alibaba: Wan 2.6
Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
Kwaipilot: KAT-Coder-Pro V2
KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...
- 上下文
- 262K
- 输入
- text
- 输出
- text
ByteDance: Seedance 1.5 Pro
ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
OpenAI: Sora 2 Pro
OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
Google: Veo 3.1
Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- video
Reka Edge
Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...
- 上下文
- 16K
- 输入
- image · text · video
- 输出
- text
MiniMax: MiniMax M2.7
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
- 上下文
- 205K
- 输入
- text
- 输出
- text