MODEL RADAR · OPENROUTER

Confronta i modelli AI con gli stessi criteri.

Tutti i modelli attualmente elencati da OpenRouter, con dati coerenti su modalità, contesto, prezzi, configurazione e provider.

AMPIEZZA CATALOGO
TUTTI PUBBLICI
modelli
631
provider
86
ultimo aggiornamento
29 set 2026
631 modelli
kwaivgi logokwaivgi

Kling: Video O1

Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content...

Contesto
Non indicato
Input
text · image
Output
video
Output video: $0.112 al secondo
Vedi i dettagli del modello →
minimax logominimax

MiniMax: Hailuo 2.3

Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is...

Contesto
Non indicato
Input
text · image
Output
video
Output video: $0.0817 al secondo
Vedi i dettagli del modello →
moonshotai logomoonshotai

MoonshotAI: Kimi K2.6

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

Contesto
262K
Input
text · image
Output
text
Input: $0.65Output: $3.41per milione di token
Vedi i dettagli del modello →
mistralai logomistralai

Mistral: Voxtral Mini TTS

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Contesto
4K
Input
text
Output
speech
Caratteri: $16 per milione di caratteri
Vedi i dettagli del modello →
google logogoogle

Google: Gemini Embedding 2 Preview

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It...

Contesto
8K
Input
text · image · file · audio · video
Output
embeddings
Input: $0.2per milione di token
Vedi i dettagli del modello →
anthropic logoanthropic

Anthropic: Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Contesto
1M
Input
text · image · file
Output
text
Input: $5Output: $25per milione di token
Vedi i dettagli del modello →
anthropic logoanthropic

Anthropic: Claude Opus 4.7 (batch)

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

Contesto
1M
Input
text · image · file
Output
text
Input: $2.5Output: $12.5per milione di token
Vedi i dettagli del modello →
alibaba logoalibaba

Alibaba: Wan 2.7

Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content...

Contesto
Non indicato
Input
text · image
Output
video
Output video: $0.1 al secondo
Vedi i dettagli del modello →
bytedance logobytedance

ByteDance: Seedance 2.0

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,...

Contesto
Non indicato
Input
text · image · video · audio
Output
video
Video (with audio): $0.1512 al secondoVideo (no audio): $0.1512 al secondo
Vedi i dettagli del modello →
bytedance logobytedance

ByteDance: Seedance 2.0 Fast

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost...

Contesto
Non indicato
Input
text · image · video · audio
Output
video
Video (with audio): $0.0907 al secondoVideo (no audio): $0.0907 al secondo
Vedi i dettagli del modello →
z-ai logoz-ai

Z.ai: GLM 5.1

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

Contesto
205K
Input
text
Output
text
Input: $1.4Output: $4.4per milione di token
Vedi i dettagli del modello →
cohere logocohere

Cohere: Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Contesto
33K
Input
text
Output
rerank
Unità di ricerca: $0.0025 per ricerca
Vedi i dettagli del modello →
cohere logocohere

Cohere: Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

Contesto
33K
Input
text
Output
rerank
Unità di ricerca: $0.002 per ricerca
Vedi i dettagli del modello →
cohere logocohere

Cohere: Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

Contesto
4K
Input
text
Output
rerank
Unità di ricerca: $0.001 per ricerca
Vedi i dettagli del modello →
google logogoogle

Google: Gemma 4 26B A4B

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Contesto
262K
Input
image · text · video
Output
text
Input: $0.0765Output: $0.255per milione di token
Vedi i dettagli del modello →
google logogoogle

Google: Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Contesto
262K
Input
image · text · video
Output
text
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello →
google logogoogle

Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Contesto
262K
Input
image · text · video
Output
text
Input: $0.09Output: $0.34per milione di token
Vedi i dettagli del modello →
google logogoogle

Google: Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Contesto
262K
Input
image · text · video
Output
text
Input: GratuitoOutput: Gratuitoper milione di token
Vedi i dettagli del modello →
qwen logoqwen

Qwen: Qwen3.6 Plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

Contesto
1M
Input
text · image · video
Output
text
Input: $0.325Output: $1.95per milione di token
Vedi i dettagli del modello →
z-ai logoz-ai

Z.ai: GLM 5V Turbo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Contesto
203K
Input
image · text · video
Output
text
Input: $1.2Output: $4per milione di token
Vedi i dettagli del modello →
arcee-ai logoarcee-ai

Arcee AI: Trinity Large Thinking

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

Contesto
262K
Input
text
Output
text
Input: $0.25Output: $0.8per milione di token
Vedi i dettagli del modello →
x-ai logox-ai

SpaceXAI: Grok 4.20 Multi-Agent

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

Contesto
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.5per milione di token
Vedi i dettagli del modello →
x-ai logox-ai

SpaceXAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

Contesto
2M
Input
text · image · file
Output
text
Input: $1.25Output: $2.5per milione di token
Vedi i dettagli del modello →
google logogoogle

Google: Lyria 3 Pro Preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Contesto
1.0M
Input
text · image
Output
text · audio
Generazione musicale: $0.08 per brano
Vedi i dettagli del modello →