MODEL RADAR · OPENROUTER
AI 모델을 같은 기준으로 비교하세요.
OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.
- 카탈로그 범위
- 전체 공개 모델
- 개 모델
- 631
- 개 제공업체
- 86
- 마지막 동기화
- 2026. 9. 29.
HeyGen: Avatar IV
HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal...
- 컨텍스트
- 정보 없음
- 입력
- text · image · audio
- 출력
- video
Meta: Muse Spark 1.2 Contributor
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file
- 출력
- text
DeepSeek: DeepSeek V4 Flash Vision Exp
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
- 컨텍스트
- 1.0M
- 입력
- text · image
- 출력
- text
Tencent: Hy-MT2-1.8B
Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...
- 컨텍스트
- 8K
- 입력
- text
- 출력
- text
Tencent: Hy-MT2-30B-A3B
Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...
- 컨텍스트
- 8K
- 입력
- text
- 출력
- text
Black Forest Labs: FLUX Video Upscale
FLUX Video Upscale is a video upscaling model from Black Forest Labs. It enlarges a single source video by 1.5× to 3× while preserving its duration, with an optional prompt...
- 컨텍스트
- 정보 없음
- 입력
- text · video
- 출력
- video
Z.ai: GLM Latest
This model always redirects to the latest GLM model from Z.ai.
- 컨텍스트
- 1.3M
- 입력
- text
- 출력
- text
Tencent: Hy-MT2-7B
Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.
- 컨텍스트
- 8K
- 입력
- text
- 출력
- text
Z.ai: GLM 5.3
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
- 컨텍스트
- 1.3M
- 입력
- text
- 출력
- text
Z.ai: GLM 5.3 (batch)
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
- 컨텍스트
- 1.0M
- 입력
- text
- 출력
- text
LiquidAI: LFM2.5-Embedding-350M (free)
LFM2.5-Embedding-350M is a text embedding model from Liquid AI. It produces 1,024-dimensional embeddings for retrieval and semantic search. Successful OpenRouter requests and embeddings may be retained and used to train...
- 컨텍스트
- 1K
- 입력
- text
- 출력
- embeddings
Qwen: Qwen3.8 27B
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
- 컨텍스트
- 1M
- 입력
- text · image · video
- 출력
- text
Qwen: Qwen3.8 27B (free)
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...
- 컨텍스트
- 262K
- 입력
- text · image · video
- 출력
- text
Dots Studio: Dots3-Note Preview (free)
Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...
- 컨텍스트
- 512K
- 입력
- text · image
- 출력
- text
NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B
Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
Mistral: Voxtral Small 24B 2507 STT
Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
Mistral: Voxtral Mini 3B 2507
Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
ByteDance Seed: Seedream 5.0 Lite
Seedream 5.0 Lite is an image generation model from ByteDance Seed. It is suited for professional visual creation that benefits from web-connected retrieval, complex-prompt comprehension, visual references, and broad knowledge...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- image
Google: Gemini 3.7 Flash
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
Google: Gemini 3.7 Flash (batch)
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
VoyageAI by MongoDB: voyage-code-4
voyage-code-4 is a code embedding model from Voyage AI, a MongoDB company. It is designed for coding agents and code retrieval, with Matryoshka embeddings at 2048, 1024, 512, and 256...
- 컨텍스트
- 32K
- 입력
- text
- 출력
- embeddings
Qwen3 Reranker 8B
Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...
- 컨텍스트
- 41K
- 입력
- text
- 출력
- rerank
Qwen: Qwen3 ASR 1.7B
Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
Qwen: Qwen3 ASR 0.6B
Qwen3 ASR 0.6B is a compact automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription