模型雷达 · OPENROUTER
把 AI 模型放在一起比较。
收录 OpenRouter 当前公开的全部模型,统一呈现模态、上下文窗口、价格、配置与提供方信息,并持续增量更新。
- 目录范围
- 全部公开
- 个模型
- 631
- 个提供方
- 86
- 最近同步
- 2026年9月29日
Anthropic: Claude Sonnet 5.5
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...
- 上下文
- 1M
- 输入
- text · image · file
- 输出
- text
Anthropic: Claude Sonnet 5.5 (batch)
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...
- 上下文
- 1M
- 输入
- text · image · file
- 输出
- text
Upstage: Solar Decide
Solar Decide is Upstage's structured decision model, served as a System One endpoint on Solar Mini 4. Send a state along with typed questions, and it returns a choice, a...
- 上下文
- 524K
- 输入
- text
- 输出
- decisions
Respan: Span-01
Span-01 is a behavior scoring model from Respan. It reads a conversation span and returns, for each plain-language behavior you define, the probability that the behavior is present. It is...
- 上下文
- 暂未提供
- 输入
- text
- 输出
- decisions
Respan: Span-01 Lite
Span-01 Lite is the free, lighter tier of Span-01, a behavior scoring model from Respan. It returns, for each plain-language behavior you define, the probability that the behavior is present...
- 上下文
- 暂未提供
- 输入
- text
- 输出
- decisions
Respan: Span-01 Lite (free)
Span-01 Lite is the free, lighter tier of Span-01, a behavior scoring model from Respan. It returns, for each plain-language behavior you define, the probability that the behavior is present...
- 上下文
- 暂未提供
- 输入
- text
- 输出
- decisions
ByteDance Seed: Seed Audio 1.0
Seed Audio 1.0 is ByteDance Seed's non-streaming audio generation model. It produces speech and other audio from a natural-language text prompt that can describe the desired voice, tone, and sound...
- 上下文
- 暂未提供
- 输入
- text
- 输出
- speech
TypeSafe: Jev Router
Jev Router picks the best model and reasoning effort for each request, balancing quality, speed, and cost. It runs on Jev, TypeSafe's first System One model, and adapts as your...
- 上下文
- 1M
- 输入
- audio · file · image · text · video
- 输出
- text
Jared Palmer: Kev 4B
Kev 4B is a small open-weight decision model from Jared Palmer, built as a LoRA adapter and pointer head on Qwen3.5-4B-Base and served over the same /v1/systemone contract as TypeSafe's...
- 上下文
- 8K
- 输入
- text
- 输出
- decisions
Perceptron: Perceptron Mk1.5
Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,...
- 上下文
- 37K
- 输入
- text · image · video · audio
- 输出
- text
Google: Gemini 3.5 Transcribe
Gemini 3.5 Transcribe is a speech-to-text model from Google. It is suited for synchronous transcription that needs word-level timestamps or speaker diarization, with support for up to eight speakers. Audio...
- 上下文
- 98K
- 输入
- audio
- 输出
- transcription
Fish Audio: Transcribe 1 Pro
Transcribe 1 Pro is a speech-to-text model from Fish Audio tuned for interviews, meetings, and podcasts. It labels speakers with inline speaker markers, preserves emotion and vocal-event cues such as...
- 上下文
- 暂未提供
- 输入
- audio
- 输出
- transcription
Fireworks: Ember-1
Ember-1 is a specialized reasoning model from Fireworks Research, built on Kimi K3. It is designed to make every token go further: it produces shorter reasoning traces, using roughly 40%...
- 上下文
- 1.0M
- 输入
- text · image
- 输出
- text
inclusionAI: Ming Image 0.1 Design Layer
Ming Image 0.1 Design Layer is an image-to-image model from inclusionAI that decomposes a flattened design image into separate RGBA layers, such as a background layer and foreground elements, and...
- 上下文
- 暂未提供
- 输入
- text · image
- 输出
- image
Google: Gemini 3.8 Flash Lite TTS
Gemini 3.8 Flash Lite TTS is a text-to-speech model from Google and the fast, high-throughput member of the 3.8 TTS family alongside Gemini 3.8 Flash TTS. It is suited for...
- 上下文
- 8K
- 输入
- text
- 输出
- speech
Google: Gemini 3.8 Flash TTS
Gemini 3.8 Flash TTS is a text-to-speech model from Google and the successor to Gemini 3.1 Flash TTS Preview. It is the creative tier of the 3.8 TTS family, suited...
- 上下文
- 8K
- 输入
- text
- 输出
- speech
Z.ai: GLM 5.3 Prime
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...
- 上下文
- 1M
- 输入
- text
- 输出
- text
Qwen: Qwen3.8 Max Prime
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
- 上下文
- 1M
- 输入
- text · image · video
- 输出
- text
Recraft: Recraft V4.1 Flash
Recraft V4.1 Flash is a text-to-image model from Recraft, the speed and cost tier of the V4.1 family. It generates 1K raster images in about 1.5 seconds end to end,...
- 上下文
- 66K
- 输入
- text
- 输出
- image
Space Bunny Alpha
Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space...
- 上下文
- 1M
- 输入
- text · image · video
- 输出
- text
AionLabs: Aion 3.5 Mini
Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...
- 上下文
- 262K
- 输入
- text
- 输出
- text
AionLabs: Aion 3.5
Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each...
- 上下文
- 262K
- 输入
- text
- 输出
- text
Upstage: Solar Mini 4
Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window. It is built for agentic use cases where response...
- 上下文
- 524K
- 输入
- text
- 输出
- text
Cohere: Command A+
Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured...
- 上下文
- 192K
- 输入
- text · image
- 输出
- text