MODEL RADAR · OPENROUTER

AI 모델을 같은 기준으로 비교하세요.

OpenRouter에 현재 공개된 모든 모델의 모달리티, 컨텍스트, 가격, 설정, 제공업체 정보를 한곳에서 비교할 수 있습니다.

카탈로그 범위
전체 공개 모델
개 모델
631
개 제공업체
86
마지막 동기화
2026. 9. 29.
631 개 모델
thinkingmachines logothinkingmachines

Thinking Machines: Inkling Small (free)

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

컨텍스트
1.0M
입력
text · image · audio
출력
text
입력: 무료출력: 무료백만 토큰당
모델 상세 보기 →
minimax logominimax

MiniMax: H3

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...

컨텍스트
정보 없음
입력
text · image · video · audio
출력
video
동영상 출력: $0.13 초당Reference Image (first 5 free): $0.04 이미지당
모델 상세 보기 →
fish-audio logofish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.0001 초당
모델 상세 보기 →
fish-audio logofish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

컨텍스트
정보 없음
입력
text
출력
speech
UTF-8 바이트: $15 백만 UTF-8 바이트당
모델 상세 보기 →
fish-audio logofish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

컨텍스트
정보 없음
입력
text
출력
speech
UTF-8 바이트: $15 백만 UTF-8 바이트당
모델 상세 보기 →
fish-audio logofish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

컨텍스트
정보 없음
입력
text
출력
speech
입력: 무료 백만 토큰당출력: 무료 백만 토큰당
모델 상세 보기 →
fish-audio logofish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

컨텍스트
정보 없음
입력
text
출력
speech
UTF-8 바이트: $15 백만 UTF-8 바이트당
모델 상세 보기 →
runway logorunway

Runway: Aleph 2.0

Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....

컨텍스트
정보 없음
입력
text · image · video
출력
video
동영상 출력: $0.28 초당
모델 상세 보기 →
runway logorunway

Runway: Gen-4.5

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....

컨텍스트
정보 없음
입력
text · image
출력
video
동영상 출력: $0.12 초당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

컨텍스트
1M
입력
text · image · video
출력
text
입력: $0.03출력: $0.13백만 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: rerank-2.5-lite

rerank-2.5-lite is a reranker optimized for both latency and quality, delivering a 7.16% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5...

컨텍스트
32K
입력
text
출력
rerank
입력: $0.02 백만 입력 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: rerank-2.5

rerank-2.5 is a cutting-edge reranker optimized for quality, delivering a 7.94% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5 by 12.70%...

컨텍스트
32K
입력
text
출력
rerank
입력: $0.05 백만 입력 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: voyage-multimodal-3.5

voyage-multimodal-3.5 is a state-of-the-art multimodal embedding model capable of vectorizing not only text, images, and video individually, but also content that interleaves all three modalities. It delivers excellent performance for...

컨텍스트
32K
입력
text · image
출력
embeddings
입력: $0.12백만 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: voyage-4-lite

voyage-4-lite is a lightweight, general-purpose embedding model optimized for low latency and cost. Enabled by Matryoshka learning and quantization-aware training, voyage-4-lite supports embeddings in 2048, 1024, 512, and 256 dimensions,...

컨텍스트
32K
입력
text
출력
embeddings
입력: $0.02백만 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: voyage-4

voyage-4 is a general-purpose (including multilingual) embedding model optimized for retrieval/search and AI applications. voyage-4 supports embeddings in 2048, 1024, 512, and 256 dimensions, with multiple quantization options. Learn more...

컨텍스트
32K
입력
text
출력
embeddings
입력: $0.06백만 토큰당
모델 상세 보기 →
voyageai logovoyageai

VoyageAI by MongoDB: voyage-4-large

voyage-4-large is a state-of-the-art general-purpose and multilingual embedding optimized for retrieval quality. Enabled by Matryoshka learning and quantization-aware training, voyage-4-large supports embeddings in 2048, 1024, 512, and 256 dimensions, with...

컨텍스트
32K
입력
text
출력
embeddings
입력: $0.12백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $5출력: $25백만 토큰당
모델 상세 보기 →
anthropic logoanthropic

Anthropic: Claude Opus 5 (batch)

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

컨텍스트
1M
입력
text · image · file
출력
text
입력: $2.5출력: $12.5백만 토큰당
모델 상세 보기 →
microsoft logomicrosoft

Microsoft AI: MAI-Image-2.5 Pro

Microsoft AI's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

컨텍스트
4K
입력
text · image
출력
image
입력: $5 백만 토큰당이미지 출력: $108 백만 토큰당
모델 상세 보기 →
microsoft logomicrosoft

Microsoft AI: MAI-Voice-2-Flash

MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft AI for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15...

컨텍스트
정보 없음
입력
text
출력
speech
문자: $15 백만 문자당
모델 상세 보기 →
inclusionai logoinclusionai

inclusionAI: Ling 3.0 Flash

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...

컨텍스트
262K
입력
text
출력
text
입력: $0.021출력: $0.063백만 토큰당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen-Audio-3.0-TTS Flash

Qwen-Audio-3.0-TTS Flash is Alibaba's fast, cost-efficient text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

컨텍스트
정보 없음
입력
text
출력
speech
문자: $15 백만 문자당
모델 상세 보기 →
qwen logoqwen

Qwen: Qwen-Audio-3.0-TTS Plus

Qwen-Audio-3.0-TTS Plus is Alibaba's higher-quality text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

컨텍스트
정보 없음
입력
text
출력
speech
문자: $20 백만 문자당
모델 상세 보기 →
x-ai logox-ai

SpaceXAI: Grok STT 1.0

Grok STT is SpaceXAI's speech-to-text model, available via the REST /v1/stt endpoint. It supports transcription with word-level timestamps, optional speaker diarization, and multichannel audio.

컨텍스트
정보 없음
입력
audio
출력
transcription
오디오 길이: $0.000028 초당
모델 상세 보기 →