MODEL RADAR · OPENROUTER

Porównuj modele AI według tych samych kryteriów.

Wszystkie modele obecnie dostępne w OpenRouter wraz ze spójnymi danymi o modalnościach, kontekście, cenach, konfiguracji i dostawcach.

ZAKRES KATALOGU
WSZYSTKIE PUBLICZNE
modele
631
dostawców
86
ostatnia synchronizacja
29 wrz 2026
631 modele
heygen logoheygen

HeyGen: Avatar IV

HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal...

Kontekst
Brak danych
Wejście
text · image · audio
Wyjście
video
Wideo wyjściowe: $0.05 za sekundę
Zobacz szczegóły modelu →
meta logometa

Meta: Muse Spark 1.2 Contributor

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

Kontekst
1.0M
Wejście
text · image · video · file
Wyjście
text
Wejście: $0.1Wyjście: $0.2za milion tokenów
Zobacz szczegóły modelu →
deepseek logodeepseek

DeepSeek: DeepSeek V4 Flash Vision Exp

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

Kontekst
1.0M
Wejście
text · image
Wyjście
text
Wejście: $0.2156Wyjście: $0.6468za milion tokenów
Zobacz szczegóły modelu →
tencent logotencent

Tencent: Hy-MT2-1.8B

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

Kontekst
8K
Wejście
text
Wyjście
text
Wejście: $0.044Wyjście: $0.177za milion tokenów
Zobacz szczegóły modelu →
tencent logotencent

Tencent: Hy-MT2-30B-A3B

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...

Kontekst
8K
Wejście
text
Wyjście
text
Wejście: $0.074Wyjście: $0.295za milion tokenów
Zobacz szczegóły modelu →
black-forest-labs logoblack-forest-labs

Black Forest Labs: FLUX Video Upscale

FLUX Video Upscale is a video upscaling model from Black Forest Labs. It enlarges a single source video by 1.5× to 3× while preserving its duration, with an optional prompt...

Kontekst
Brak danych
Wejście
text · video
Wyjście
video
Video Upscale Output: $0.075 za megapikselosekundę
Zobacz szczegóły modelu →
~z-ai logo~z-ai

Z.ai: GLM Latest

This model always redirects to the latest GLM model from Z.ai.

Kontekst
1.3M
Wejście
text
Wyjście
text
Wejście: $0.1949Wyjście: $4.4za milion tokenów
Zobacz szczegóły modelu →
tencent logotencent

Tencent: Hy-MT2-7B

Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

Kontekst
8K
Wejście
text
Wyjście
text
Wejście: $0.074Wyjście: $0.295za milion tokenów
Zobacz szczegóły modelu →
z-ai logoz-ai

Z.ai: GLM 5.3

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Kontekst
1.3M
Wejście
text
Wyjście
text
Wejście: $1.4Wyjście: $4.4za milion tokenów
Zobacz szczegóły modelu →
z-ai logoz-ai

Z.ai: GLM 5.3 (batch)

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Kontekst
1.0M
Wejście
text
Wyjście
text
Wejście: $0.45Wyjście: $2za milion tokenów
Zobacz szczegóły modelu →
liquid logoliquid

LiquidAI: LFM2.5-Embedding-350M (free)

LFM2.5-Embedding-350M is a text embedding model from Liquid AI. It produces 1,024-dimensional embeddings for retrieval and semantic search. Successful OpenRouter requests and embeddings may be retained and used to train...

Kontekst
1K
Wejście
text
Wyjście
embeddings
Wejście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu →
qwen logoqwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Kontekst
1M
Wejście
text · image · video
Wyjście
text
Wejście: $0.0507Wyjście: $4.4za milion tokenów
Zobacz szczegóły modelu →
qwen logoqwen

Qwen: Qwen3.8 27B (free)

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Kontekst
262K
Wejście
text · image · video
Wyjście
text
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu →
dots-studio logodots-studio

Dots Studio: Dots3-Note Preview (free)

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

Kontekst
512K
Wejście
text · image
Wyjście
text
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu →
nvidia logonvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.000003 za sekundę
Zobacz szczegóły modelu →
mistralai logomistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.00005 za sekundę
Zobacz szczegóły modelu →
mistralai logomistralai

Mistral: Voxtral Mini 3B 2507

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.000017 za sekundę
Zobacz szczegóły modelu →
bytedance-seed logobytedance-seed

ByteDance Seed: Seedream 5.0 Lite

Seedream 5.0 Lite is an image generation model from ByteDance Seed. It is suited for professional visual creation that benefits from web-connected retrieval, complex-prompt comprehension, visual references, and broad knowledge...

Kontekst
Brak danych
Wejście
text · image
Wyjście
image
Obraz wejściowy: Bezpłatnie za obrazObraz wyjściowy: $0.035 za obraz
Zobacz szczegóły modelu →
google logogoogle

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontekst
1.0M
Wejście
text · image · video · file · audio
Wyjście
text
Wejście: $0.75Wyjście: $3.75za milion tokenów
Zobacz szczegóły modelu →
google logogoogle

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Kontekst
1.0M
Wejście
text · image · video · file · audio
Wyjście
text
Wejście: $0.375Wyjście: $1.875za milion tokenów
Zobacz szczegóły modelu →
voyageai logovoyageai

VoyageAI by MongoDB: voyage-code-4

voyage-code-4 is a code embedding model from Voyage AI, a MongoDB company. It is designed for coding agents and code retrieval, with Matryoshka embeddings at 2048, 1024, 512, and 256...

Kontekst
32K
Wejście
text
Wyjście
embeddings
Wejście: $0.12za milion tokenów
Zobacz szczegóły modelu →
qwen logoqwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

Kontekst
41K
Wejście
text
Wyjście
rerank
Wejście: $0.2 za milion tokenów wejściowych
Zobacz szczegóły modelu →
qwen logoqwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.000008 za sekundę
Zobacz szczegóły modelu →
qwen logoqwen

Qwen: Qwen3 ASR 0.6B

Qwen3 ASR 0.6B is a compact automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline...

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Czas audio: $0.000003 za sekundę
Zobacz szczegóły modelu →