MODEL RADAR · OPENROUTER
Comparer les modèles d’IA sur une même base.
Tous les modèles actuellement référencés par OpenRouter, avec des informations cohérentes sur les modalités, le contexte, les tarifs, la configuration et les fournisseurs.
- PORTÉE DU CATALOGUE
- TOUS PUBLICS
- modèles
- 631
- fournisseurs
- 86
- dernière synchronisation
- 29 sept. 2026
Kling: Video O1
Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...
- Contexte
- Non indiqué
- Entrée
- text · image
- Sortie
- video
MiniMax: Hailuo 2.3
Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...
- Contexte
- Non indiqué
- Entrée
- text · image
- Sortie
- video
MoonshotAI: Kimi K2.6
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
- Contexte
- 262K
- Entrée
- text · image
- Sortie
- text
Mistral: Voxtral Mini TTS
Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.
- Contexte
- 4K
- Entrée
- text
- Sortie
- speech
Google: Gemini Embedding 2 Preview
Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...
- Contexte
- 8K
- Entrée
- text · image · file · audio · video
- Sortie
- embeddings
Anthropic: Claude Opus 4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- Contexte
- 1M
- Entrée
- text · image · file
- Sortie
- text
Anthropic: Claude Opus 4.7 (batch)
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
- Contexte
- 1M
- Entrée
- text · image · file
- Sortie
- text
Alibaba: Wan 2.7
Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...
- Contexte
- Non indiqué
- Entrée
- text · image
- Sortie
- video
ByteDance: Seedance 2.0
Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...
- Contexte
- Non indiqué
- Entrée
- text · image · video · audio
- Sortie
- video
ByteDance: Seedance 2.0 Fast
Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...
- Contexte
- Non indiqué
- Entrée
- text · image · video · audio
- Sortie
- video
Z.ai: GLM 5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
- Contexte
- 205K
- Entrée
- text
- Sortie
- text
Cohere: Rerank 4 Pro
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- Contexte
- 33K
- Entrée
- text
- Sortie
- rerank
Cohere: Rerank 4 Fast
Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...
- Contexte
- 33K
- Entrée
- text
- Sortie
- rerank
Cohere: Rerank v3.5
Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...
- Contexte
- 4K
- Entrée
- text
- Sortie
- rerank
Google: Gemma 4 26B A4B
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- Contexte
- 262K
- Entrée
- image · text · video
- Sortie
- text
Google: Gemma 4 26B A4B (free)
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- Contexte
- 262K
- Entrée
- image · text · video
- Sortie
- text
Google: Gemma 4 31B
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- Contexte
- 262K
- Entrée
- image · text · video
- Sortie
- text
Google: Gemma 4 31B (free)
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- Contexte
- 262K
- Entrée
- image · text · video
- Sortie
- text
Qwen: Qwen3.6 Plus
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
- Contexte
- 1M
- Entrée
- text · image · video
- Sortie
- text
Z.ai: GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
- Contexte
- 203K
- Entrée
- image · text · video
- Sortie
- text
Arcee AI: Trinity Large Thinking
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
- Contexte
- 262K
- Entrée
- text
- Sortie
- text
SpaceXAI: Grok 4.20 Multi-Agent
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
- Contexte
- 2M
- Entrée
- text · image · file
- Sortie
- text
SpaceXAI: Grok 4.20
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
- Contexte
- 2M
- Entrée
- text · image · file
- Sortie
- text
Google: Lyria 3 Pro Preview
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
- Contexte
- 1.0M
- Entrée
- text · image
- Sortie
- text · audio