MODEL RADAR · OPENROUTER
AI 모델을 같은 기준으로 비교하세요.
OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.
- 카탈로그 범위
- 전체 공개 모델
- 개 모델
- 534
- 개 제공업체
- 75
- 마지막 동기화
- 2026. 8. 31.
Google: Gemini 3.5 Flash
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
Google: Gemini 3.5 Flash (batch)
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
SpaceXAI: Grok Imagine Video
Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...
- 컨텍스트
- 정보 없음
- 입력
- text · image
- 출력
- video
SpaceXAI: Grok Imagine Image Quality
Grok Imagine Image Quality is SpaceXAI's fast, high-fidelity image generation and editing model. It accepts text prompts and optional reference images, producing photorealistic outputs at 1K or 2K across a...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Mistral: Voxtral Mini Transcribe
Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
SpaceXAI: Grok Voice TTS 1.0
Grok Voice TTS 1.0 is a text-to-speech model from SpaceXAI. It converts text into spoken audio across 20+ languages with automatic language detection, and offers five built-in voices (Eve, Ara,...
- 컨텍스트
- 15K
- 입력
- text
- 출력
- speech
Qwen: Qwen3 ASR Flash
Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription
Recraft: Recraft V4.1 Pro Vector
Recraft V4.1 Pro Vector is the vector (SVG) variant of Recraft V4.1 Pro, tuned for high aesthetics. It supports text and image inputs and produces higher-resolution SVG image output across...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4.1 Vector
Recraft V4.1 Vector is the vector (SVG) variant of Recraft V4.1, tuned for high aesthetics. It supports text and image inputs and produces SVG image output across multiple aspect ratios,...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4.1 Utility Pro
Recraft V4.1 Utility Pro is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios — double...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4.1 Utility
Recraft V4.1 Utility is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with typical generation...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4.1 Pro
Recraft V4.1 Pro is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4.1
Recraft V4.1 is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4 Pro Vector
Recraft V4 Pro Vector is the vector (SVG) variant of Recraft V4 Pro. It supports text and image inputs and produces vector image output across multiple aspect ratios at the...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4 Vector
Recraft V4 Vector is the vector (SVG) variant of Recraft V4. It supports text and image inputs and produces vector image output across multiple aspect ratios. Compared to the raster...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Anthropic: Claude Opus 4.7 (Fast)
Fast-mode variant of Opus 4.7 - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
- 컨텍스트
- 1M
- 입력
- text · image · file
- 출력
- text
Perceptron: Perceptron Mk1
Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning. It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...
- 컨텍스트
- 33K
- 입력
- text · image · video
- 출력
- text
Recraft: Recraft V4 Pro
Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios, double the resolution of...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V4
Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Recraft: Recraft V3
Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios. Supports the following imageconfig parameters:...
- 컨텍스트
- 66K
- 입력
- text · image
- 출력
- image
Google: Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
Google: Gemini 3.1 Flash Lite (batch)
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- 컨텍스트
- 1.0M
- 입력
- text · image · video · file · audio
- 출력
- text
OpenAI: GPT Chat Latest
GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...
- 컨텍스트
- 400K
- 입력
- text · image · file
- 출력
- text
Google: Chirp 3
Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and...
- 컨텍스트
- 정보 없음
- 입력
- audio
- 출력
- transcription