MODEL RADAR · OPENROUTER

AI 모델을 같은 기준으로 비교하세요.

OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.

카탈로그 범위
전체 공개 모델
개 모델
631
개 제공업체
86
마지막 동기화
2026. 9. 29.
631 개 모델
inception logoinception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

컨텍스트
128K
입력
text
출력
text
입력: $0.25출력: $0.75백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.1 Flash Lite Preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.25출력: $1.5백만 토큰당
모델 상세 보기 →
bytedance-seed logobytedance-seed

ByteDance Seed: Seed-2.0-Mini

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.1출력: $0.4백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

컨텍스트
66K
입력
image · text
출력
image · text
입력: $0.5 백만 토큰당출력: $3 백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5-35B-A3B

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.1625출력: $1.3백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5-27B

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.195출력: $1.56백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5-122B-A10B

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.26출력: $2.08백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5-Flash

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

컨텍스트
1M
입력
text · image · video
출력
text
입력: $0.065출력: $0.26백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.1 Pro Preview Custom Tools

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

컨텍스트
1.0M
입력
text · audio · image · video · file
출력
text
입력: $2출력: $12백만 토큰당
모델 상세 보기 →
nvidia logonvidia

NVIDIA: Llama Nemotron Embed VL 1B V2 (free)

The Llama Nemotron Embed VL 1B V2 embedding model is optimized for multimodal question-answering retrieval. The model can embed 'documents' in the form of image, text, or image and text...

컨텍스트
131K
입력
text · image
출력
embeddings
입력: 무료백만 토큰당
모델 상세 보기 →
openai logoopenai

OpenAI: GPT-5.3-Codex

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...

컨텍스트
400K
입력
text · image · file
출력
text
입력: $1.75출력: $14백만 토큰당
모델 상세 보기 →
aion-labs logoaion-labs

AionLabs: Aion-2.0

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

컨텍스트
131K
입력
text
출력
text
입력: $0.8출력: $1.6백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.1 Pro Preview

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

컨텍스트
1.0M
입력
audio · file · image · text · video
출력
text
입력: $2출력: $12백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.1 Pro Preview (batch)

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

컨텍스트
1.0M
입력
audio · file · image · text · video
출력
text
입력: $1출력: $6백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Sonnet 4.6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $3출력: $15백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Sonnet 4.6 (batch)

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $1.5출력: $7.5백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5 Plus 2026-02-15

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...

컨텍스트
1M
입력
text · image · video
출력
text
입력: $0.26출력: $1.56백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.5 397B A17B

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.55출력: $3.5백만 토큰당
모델 상세 보기 →
minimax logominimax

MiniMax: MiniMax M2.5

MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...

컨텍스트
205K
입력
text
출력
text
입력: $0.27출력: $1.08백만 토큰당
모델 상세 보기 →
z-ai logoz-ai

Z.ai: GLM 5

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

컨텍스트
205K
입력
text
출력
text
입력: $0.6출력: $1.92백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3 Max Thinking

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it...

컨텍스트
262K
입력
text
출력
text
입력: $0.78출력: $3.9백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 4.6

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $5출력: $25백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 4.6 (batch)

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $2.5출력: $12.5백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3 Coder Next

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

컨텍스트
262K
입력
text
출력
text
입력: $0.12출력: $0.8백만 토큰당
모델 상세 보기 →