MODEL RADAR · OPENROUTER

Comparer les modèles d’IA sur une même base.

Tous les modèles actuellement référencés par OpenRouter, avec des informations cohérentes sur les modalités, le contexte, les tarifs, la configuration et les fournisseurs.

PORTÉE DU CATALOGUE
TOUS PUBLICS
modèles
631
fournisseurs
86
dernière synchronisation
29 sept. 2026
631 modèles
kwaivgi logokwaivgi

Kling: Video O1

Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...

Contexte
Non indiqué
Entrée
text · image
Sortie
video
Vidéo en sortie: $0.112 par seconde
Voir la fiche du modèle →
minimax logominimax

MiniMax: Hailuo 2.3

Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...

Contexte
Non indiqué
Entrée
text · image
Sortie
video
Vidéo en sortie: $0.0817 par seconde
Voir la fiche du modèle →
moonshotai logomoonshotai

MoonshotAI: Kimi K2.6

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

Contexte
262K
Entrée
text · image
Sortie
text
Entrée: $0.65Sortie: $3.41par million de tokens
Voir la fiche du modèle →
mistralai logomistralai

Mistral: Voxtral Mini TTS

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Contexte
4K
Entrée
text
Sortie
speech
Caractères: $16 par million de caractères
Voir la fiche du modèle →
google logogoogle

Google: Gemini Embedding 2 Preview

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...

Contexte
8K
Entrée
text · image · file · audio · video
Sortie
embeddings
Entrée: $0.2par million de tokens
Voir la fiche du modèle →
anthropic logoanthropic

Anthropic: Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Contexte
1M
Entrée
text · image · file
Sortie
text
Entrée: $5Sortie: $25par million de tokens
Voir la fiche du modèle →
anthropic logoanthropic

Anthropic: Claude Opus 4.7 (batch)

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Contexte
1M
Entrée
text · image · file
Sortie
text
Entrée: $2.5Sortie: $12.5par million de tokens
Voir la fiche du modèle →
alibaba logoalibaba

Alibaba: Wan 2.7

Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...

Contexte
Non indiqué
Entrée
text · image
Sortie
video
Vidéo en sortie: $0.1 par seconde
Voir la fiche du modèle →
bytedance logobytedance

ByteDance: Seedance 2.0

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

Contexte
Non indiqué
Entrée
text · image · video · audio
Sortie
video
Video (with audio): $0.1512 par secondeVideo (no audio): $0.1512 par seconde
Voir la fiche du modèle →
bytedance logobytedance

ByteDance: Seedance 2.0 Fast

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

Contexte
Non indiqué
Entrée
text · image · video · audio
Sortie
video
Video (with audio): $0.0907 par secondeVideo (no audio): $0.0907 par seconde
Voir la fiche du modèle →
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Contexte
205K
Entrée
text
Sortie
text
Entrée: $1.4Sortie: $4.4par million de tokens
Voir la fiche du modèle →
cohere logocohere

Cohere: Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Contexte
33K
Entrée
text
Sortie
rerank
Unités de recherche: $0.0025 par recherche
Voir la fiche du modèle →
cohere logocohere

Cohere: Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Contexte
33K
Entrée
text
Sortie
rerank
Unités de recherche: $0.002 par recherche
Voir la fiche du modèle →
cohere logocohere

Cohere: Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

Contexte
4K
Entrée
text
Sortie
rerank
Unités de recherche: $0.001 par recherche
Voir la fiche du modèle →
google logogoogle

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Contexte
262K
Entrée
image · text · video
Sortie
text
Entrée: $0.0765Sortie: $0.255par million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Contexte
262K
Entrée
image · text · video
Sortie
text
Entrée: GratuitSortie: Gratuitpar million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Contexte
262K
Entrée
image · text · video
Sortie
text
Entrée: $0.09Sortie: $0.34par million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Contexte
262K
Entrée
image · text · video
Sortie
text
Entrée: GratuitSortie: Gratuitpar million de tokens
Voir la fiche du modèle →
qwen logoqwen

Qwen: Qwen3.6 Plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

Contexte
1M
Entrée
text · image · video
Sortie
text
Entrée: $0.325Sortie: $1.95par million de tokens
Voir la fiche du modèle →
z-ai logoz-ai

Z.ai: GLM 5V Turbo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Contexte
203K
Entrée
image · text · video
Sortie
text
Entrée: $1.2Sortie: $4par million de tokens
Voir la fiche du modèle →
arcee-ai logoarcee-ai

Arcee AI: Trinity Large Thinking

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

Contexte
262K
Entrée
text
Sortie
text
Entrée: $0.25Sortie: $0.8par million de tokens
Voir la fiche du modèle →
x-ai logox-ai

SpaceXAI: Grok 4.20 Multi-Agent

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

Contexte
2M
Entrée
text · image · file
Sortie
text
Entrée: $1.25Sortie: $2.5par million de tokens
Voir la fiche du modèle →
x-ai logox-ai

SpaceXAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Contexte
2M
Entrée
text · image · file
Sortie
text
Entrée: $1.25Sortie: $2.5par million de tokens
Voir la fiche du modèle →
google logogoogle

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Contexte
1.0M
Entrée
text · image
Sortie
text · audio
Génération musicale: $0.08 par morceau
Voir la fiche du modèle →