MODEL RADAR · OPENROUTER
AI 모델을 같은 기준으로 비교하세요.
OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.
- 카탈로그 범위
- 전체 공개 모델
- 개 모델
- 534
- 개 제공업체
- 75
- 마지막 동기화
- 2026. 8. 31.
OpenAI: GPT-4o Mini Transcribe
GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that...
- 컨텍스트
- 128K
- 입력
- audio
- 출력
- transcription
OpenAI: Whisper Large V3
Whisper Large V3 is OpenAI's open-source automatic speech recognition model offering both audio transcription and translation. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm,...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
OpenAI: Whisper Large V3 Turbo
Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for speed and cost efficiency. It supports transcription across 99+ languages with a...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
SpaceXAI: Grok 4.3
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
- 컨텍스트
- 1M
- 입력
- text · image · file
- 출력
- text
IBM: Granite 4.1 8B
Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...
- 컨텍스트
- 131K
- 입력
- text
- 출력
- text
Mistral: Mistral Medium 3.5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- 컨텍스트
- 262K
- 입력
- text · image · file
- 출력
- text
Mistral: Mistral Medium 3.5 (batch)
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- 컨텍스트
- 262K
- 입력
- text · image · file
- 출력
- text
Kling: Video v3.0 Pro
Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
Kling: Video v3.0 Standard
Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
NVIDIA: Nemotron 3 Nano Omni (free)
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
- 컨텍스트
- 256K
- 입력
- text · audio · image · video
- 출력
- text
OpenAI: Whisper 1
Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
OpenAI: GPT-4o Transcribe
GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.
- 컨텍스트
- 128K
- 입력
- audio
- 출력
- transcription
Anthropic Claude Haiku Latest
This model always redirects to the latest model in the Anthropic Claude Haiku family.
- 컨텍스트
- 200K
- 입력
- text · image · file
- 출력
- text
OpenAI GPT Mini Latest
This model always redirects to the latest model in the OpenAI GPT Mini family.
- 컨텍스트
- 400K
- 입력
- file · image · text
- 출력
- text
Google Gemini Pro Latest
This model always redirects to the latest model in the Google Gemini Pro family.
- 컨텍스트
- 1.0M
- 입력
- audio · file · image · text · video
- 출력
- text
MoonshotAI Kimi Latest
This model always redirects to the latest model in the MoonshotAI Kimi family.
- 컨텍스트
- 1.0M
- 입력
- text · image · video
- 출력
- text
Google Gemini Flash Latest
This model always redirects to the latest model in the Google Gemini Flash family.
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
Anthropic Claude Sonnet Latest
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
- 컨텍스트
- 1M
- 입력
- text · image · file
- 출력
- text
OpenAI GPT Latest
This model always redirects to the latest model in the OpenAI GPT family.
- 컨텍스트
- 1.1M
- 입력
- file · image · text
- 출력
- text
Qwen: Qwen3.5 Plus 2026-04-20
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
- 컨텍스트
- 1M
- 입력
- text · image · video
- 출력
- text
Qwen: Qwen3.6 Flash
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
- 컨텍스트
- 1M
- 입력
- text · image · video
- 출력
- text
Qwen: Qwen3.6 35B A3B
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
- 컨텍스트
- 262K
- 입력
- text · image · video
- 출력
- text
Qwen: Qwen3.6 Max Preview
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
- 컨텍스트
- 262K
- 입력
- text
- 출력
- text
Qwen: Qwen3.6 27B
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...
- 컨텍스트
- 262K
- 입력
- text · image · video
- 출력
- text