MODEL RADAR · OPENROUTER

Porównuj modele AI według tych samych kryteriów.

Wszystkie modele obecnie dostępne w OpenRouter wraz ze spójnymi danymi o modalnościach, kontekście, cenach, konfiguracji i dostawcach.

ZAKRES KATALOGU
WSZYSTKIE PUBLICZNE
modele
534
dostawców
75
ostatnia synchronizacja
31 sie 2026
534 modele
meta logometa

Meta: Muse Glimmer 30B (batch)

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

Kontekst
131K
Wejście
text · image
Wyjście
text
Wejście: $0.350Wyjście: $1.50za milion tokenów
Zobacz szczegóły modelu
bytedance logobytedance

ByteDance: Seedance 2.5

Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...

Kontekst
Brak danych
Wejście
text · image · video · audio
Wyjście
video
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
openai logoopenai

OpenAI: GPT Transcribe

GPT Transcribe is a high-accuracy speech-to-text model from OpenAI. It is suited for recorded audio, streamed file transcription, and committed Realtime turns, with free-form context, keyword hints, and multiple language...

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Wejście: $4500.00Wyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
meta logometa

Meta: Muse Spark 1.2

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

Kontekst
1.0M
Wejście
text · image · video · file · audio
Wyjście
text
Wejście: $1.25Wyjście: $4.25za milion tokenów
Zobacz szczegóły modelu
qwen logoqwen

Qwen: Qwen Image 3

Qwen Image 3 is a unified image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with a richer world...

Kontekst
66K
Wejście
text · image
Wyjście
image
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
qwen logoqwen

Qwen: Qwen Image 3 Pro

Qwen Image 3 Pro is an image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with richer world knowledge...

Kontekst
66K
Wejście
text · image
Wyjście
image
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
black-forest-labs logoblack-forest-labs

Black Forest Labs: FLUX.3 Video

FLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled...

Kontekst
Brak danych
Wejście
text · image · video
Wyjście
video
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
qwen logoqwen

Qwen: Qwen3.8 Max

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...

Kontekst
1M
Wejście
text · image · video
Wyjście
text
Wejście: $2.00Wyjście: $6.00za milion tokenów
Zobacz szczegóły modelu
~deepseek logo~deepseek

DeepSeek V4 Flash Latest

This model always redirects to the latest model in the DeepSeek V4 Flash family.

Kontekst
1.3M
Wejście
text
Wyjście
text
Wejście: $0.030Wyjście: $0.160za milion tokenów
Zobacz szczegóły modelu
deepseek logodeepseek

DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

Kontekst
1.3M
Wejście
text
Wyjście
text
Wejście: $0.065Wyjście: $0.180za milion tokenów
Zobacz szczegóły modelu
deepseek logodeepseek

DeepSeek: DeepSeek V4 Flash 0731 (batch)

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

Kontekst
1.0M
Wejście
text
Wyjście
text
Wejście: $0.140Wyjście: $0.280za milion tokenów
Zobacz szczegóły modelu
thinkingmachines logothinkingmachines

Thinking Machines: Inkling Small

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Kontekst
1.0M
Wejście
text · image · audio
Wyjście
text
Wejście: $0.450Wyjście: $1.20za milion tokenów
Zobacz szczegóły modelu
thinkingmachines logothinkingmachines

Thinking Machines: Inkling Small (batch)

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Kontekst
524K
Wejście
text · image · audio
Wyjście
text
Wejście: $0.500Wyjście: $1.20za milion tokenów
Zobacz szczegóły modelu
thinkingmachines logothinkingmachines

Thinking Machines: Inkling Small (free)

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Kontekst
1.0M
Wejście
text · image · audio
Wyjście
text
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
minimax logominimax

MiniMax: H3

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and...

Kontekst
Brak danych
Wejście
text · image · video · audio
Wyjście
video
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
fish-audio logofish-audio

Fish Audio: Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Kontekst
Brak danych
Wejście
audio
Wyjście
transcription
Wejście: $100.00Wyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
fish-audio logofish-audio

Fish Audio: S1

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported...

Kontekst
Brak danych
Wejście
text
Wyjście
speech
Wejście: $15.00Wyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
fish-audio logofish-audio

Fish Audio: S2 Pro

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

Kontekst
Brak danych
Wejście
text
Wyjście
speech
Wejście: $15.00Wyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
fish-audio logofish-audio

Fish Audio: S2.1 Pro Free (free)

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

Kontekst
Brak danych
Wejście
text
Wyjście
speech
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
fish-audio logofish-audio

Fish Audio: S2.1 Pro

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and...

Kontekst
Brak danych
Wejście
text
Wyjście
speech
Wejście: $15.00Wyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
runway logorunway

Runway: Aleph 2.0

Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change....

Kontekst
Brak danych
Wejście
text · image · video
Wyjście
video
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
runway logorunway

Runway: Gen-4.5

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence....

Kontekst
Brak danych
Wejście
text · image
Wyjście
video
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu
qwen logoqwen

Qwen: Qwen3.7 Flash

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

Kontekst
1M
Wejście
text · image · video
Wyjście
text
Wejście: $0.030Wyjście: $0.130za milion tokenów
Zobacz szczegóły modelu
voyageai logovoyageai

VoyageAI by MongoDB: rerank-2.5-lite

rerank-2.5-lite is a reranker optimized for both latency and quality, delivering a 7.16% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5...

Kontekst
32K
Wejście
text
Wyjście
rerank
Wejście: BezpłatnieWyjście: Bezpłatnieza milion tokenów
Zobacz szczegóły modelu