MODEL RADAR · OPENROUTER

AI 모델을 같은 기준으로 비교하세요.

OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.

카탈로그 범위
전체 공개 모델
개 모델
631
개 제공업체
86
마지막 동기화
2026. 9. 29.
631 개 모델
stepfun logostepfun

StepFun: Step 3.7 Flash

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

컨텍스트
262K
입력
text · image · video
출력
text
입력: $0.2출력: $1.15백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $5출력: $25백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 4.8 (batch)

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $2.5출력: $12.5백만 토큰당
모델 상세 보기 →
nvidia logonvidia

NVIDIA: Parakeet TDT 0.6B v3

Parakeet TDT 0.6B v3 is NVIDIA's 600M-parameter multilingual speech-to-text model built on the FastConformer-TDT architecture. Trained on the Granary dataset (670,000+ hours of audio), it supports automatic language detection across...

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000025 초당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.7 Max

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

컨텍스트
1M
입력
text
출력
text
입력: $1.475출력: $4.425백만 토큰당
모델 상세 보기 →
x-ai logox-ai

SpaceXAI: Grok Build 0.1

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

컨텍스트
256K
입력
text · image · file
출력
text
입력: $1출력: $2백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini Embedding 2

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports...

컨텍스트
8K
입력
text · image · file · audio · video
출력
embeddings
입력: $0.2백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini Embedding 2 (batch)

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports...

컨텍스트
8K
입력
text · image · file · audio · video
출력
embeddings
입력: $0.1백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.5 Flash

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $1.5출력: $9백만 토큰당
모델 상세 보기 →
google logogoogle

Google: Gemini 3.5 Flash (batch)

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

컨텍스트
1.0M
입력
text · image · video · file · audio
출력
text
입력: $0.75출력: $4.5백만 토큰당
모델 상세 보기 →
x-ai logox-ai

SpaceXAI: Grok Imagine Video

Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -...

컨텍스트
정보 없음
입력
text · image
출력
video
이미지 입력: $0.002 이미지당동영상 출력: $0.05 초당
모델 상세 보기 →
x-ai logox-ai

SpaceXAI: Grok Imagine Image Quality

Grok Imagine Image Quality is SpaceXAI's fast, high-fidelity image generation and editing model. It accepts text prompts and optional reference images, producing photorealistic outputs at 1K or 2K across a...

컨텍스트
66K
입력
text · image
출력
image
이미지 입력: $0.01 이미지당이미지 출력: $0.05 이미지당
모델 상세 보기 →
mistralai logomistralai

Mistral: Voxtral Mini Transcribe

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,...

컨텍스트
16K
입력
audio
출력
transcription
오디오 길이: $0.00005 초당
모델 상세 보기 →
x-ai logox-ai

SpaceXAI: Grok Voice TTS 1.0

Grok Voice TTS 1.0 is a text-to-speech model from SpaceXAI. It converts text into spoken audio across 20+ languages with automatic language detection, and offers five built-in voices (Eve, Ara,...

컨텍스트
15K
입력
text
출력
speech
문자: $15 백만 문자당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3 ASR Flash

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —...

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000035 초당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1 Pro Vector

Recraft V4.1 Pro Vector is the vector (SVG) variant of Recraft V4.1 Pro, tuned for high aesthetics. It supports text and image inputs and produces higher-resolution SVG image output across...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.3 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1 Vector

Recraft V4.1 Vector is the vector (SVG) variant of Recraft V4.1, tuned for high aesthetics. It supports text and image inputs and produces SVG image output across multiple aspect ratios,...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.08 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1 Utility Pro

Recraft V4.1 Utility Pro is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios — double...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.21 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1 Utility

Recraft V4.1 Utility is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with typical generation...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.035 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1 Pro

Recraft V4.1 Pro is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 2K resolution across multiple aspect ratios...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.21 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4.1

Recraft V4.1 is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at 1K resolution across multiple aspect ratios, with...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.035 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4 Pro Vector

Recraft V4 Pro Vector is the vector (SVG) variant of Recraft V4 Pro. It supports text and image inputs and produces vector image output across multiple aspect ratios at the...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.3 이미지당
모델 상세 보기 →
recraft logorecraft

Recraft: Recraft V4 Vector

Recraft V4 Vector is the vector (SVG) variant of Recraft V4. It supports text and image inputs and produces vector image output across multiple aspect ratios. Compared to the raster...

컨텍스트
66K
입력
text · image
출력
image
이미지 출력: $0.08 이미지당
모델 상세 보기 →
perceptron logoperceptron

Perceptron: Perceptron Mk1

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning. It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

컨텍스트
33K
입력
text · image · video
출력
text
입력: $0.15출력: $1.5백만 토큰당
모델 상세 보기 →