MODEL RADAR · OPENROUTER
AI 모델을 같은 기준으로 비교하세요.
OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.
- 카탈로그 범위
- 전체 공개 모델
- 개 모델
- 631
- 개 제공업체
- 86
- 마지막 동기화
- 2026. 9. 29.
Kling: Video O1
Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
MiniMax: Hailuo 2.3
Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
MoonshotAI: Kimi K2.6
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
- 컨텍스트
- 262K
- 입력
- text · image
- 출력
- text
Mistral: Voxtral Mini TTS
Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.
- 컨텍스트
- 4K
- 입력
- text
- 출력
- speech
Google: Gemini Embedding 2 Preview
Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...
- 컨텍스트
- 8K
- 입력
- text · image · file · audio · video
- 출력
- embeddings
Anthropic: Claude Opus 4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- 컨텍스트
- 1M
- 입력
- text · image · file
- 출력
- text
Anthropic: Claude Opus 4.7 (batch)
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- 컨텍스트
- 1M
- 입력
- text · image · file
- 출력
- text
Alibaba: Wan 2.7
Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
ByteDance: Seedance 2.0
Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...
- 컨텍스트
- 정보 없음
- 입력
- text · image · video · audio
- 출력
- video
ByteDance: Seedance 2.0 Fast
Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...
- 컨텍스트
- 정보 없음
- 입력
- text · image · video · audio
- 출력
- video
Z.ai: GLM 5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- 컨텍스트
- 205K
- 입력
- text
- 출력
- text
Cohere: Rerank 4 Pro
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- 컨텍스트
- 33K
- 입력
- text
- 출력
- rerank
Cohere: Rerank 4 Fast
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- 컨텍스트
- 33K
- 입력
- text
- 출력
- rerank
Cohere: Rerank v3.5
Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...
- 컨텍스트
- 4K
- 입력
- text
- 출력
- rerank
Google: Gemma 4 26B A4B
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- 컨텍스트
- 262K
- 입력
- image · text · video
- 출력
- text
Google: Gemma 4 26B A4B (free)
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- 컨텍스트
- 262K
- 입력
- image · text · video
- 출력
- text
Google: Gemma 4 31B
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- 컨텍스트
- 262K
- 입력
- image · text · video
- 출력
- text
Google: Gemma 4 31B (free)
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- 컨텍스트
- 262K
- 입력
- image · text · video
- 출력
- text
Qwen: Qwen3.6 Plus
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
- 컨텍스트
- 1M
- 입력
- text · image · video
- 출력
- text
Z.ai: GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
- 컨텍스트
- 203K
- 입력
- image · text · video
- 출력
- text
Arcee AI: Trinity Large Thinking
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
- 컨텍스트
- 262K
- 입력
- text
- 출력
- text
SpaceXAI: Grok 4.20 Multi-Agent
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
- 컨텍스트
- 2M
- 입력
- text · image · file
- 출력
- text
SpaceXAI: Grok 4.20
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
- 컨텍스트
- 2M
- 입력
- text · image · file
- 출력
- text
Google: Lyria 3 Pro Preview
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
- 컨텍스트
- 1.0M
- 입력
- text · image
- 출력
- text · audio