MODEL RADAR · OPENROUTER

Novos modelos de IA, sem ruído.

Uma visão contínua dos modelos lançados nos últimos 60 dias.

ROLLING WINDOW
60 DAYS
models
119
provedores
36
última sincronização
28 de ago. de 2026
119 models
tencent

Tencent: Hy4 preview

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

Context
1.0M
Input
text
Output
text
Input: $0.834Output: $2.50per 1M tokens
Ver no OpenRouter
alibaba

Alibaba: Wan 3.0 Prime

Wan 3.0 Prime is a fast-mode variant of Wan 3.0 from Alibaba. It supports text-to-video and first-frame image-to-video generation.

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
inclusionai

Ling 3.0 Flash Fin (free)

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

Context
262K
Input
text
Output
text
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Context
1M
Input
text · image · video
Output
text
Input: $0.150Output: $0.470per 1M tokens
Ver no OpenRouter
meta

Meta: Muse Image

Muse Image is an agentic image generation model from Meta that generates and edits images from text and reference images. Unlike single-pass image models, it reasons before it renders, breaking...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
recraft

Recraft: Recraft V4 Styles Pro

Recraft V4 Styles Pro is a style-consistent image generation model from Recraft. Every request requires at least one style reference image and generates a new image that reproduces the reference's...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
recraft

Recraft: Recraft V4 Styles Vector

Recraft V4 Styles Vector is a style-consistent image generation model from Recraft. Every request requires at least one style reference image and generates a new image that reproduces the reference's...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
recraft

Recraft: Recraft V4 Styles Pro Vector

Recraft V4 Styles Pro Vector is a style-consistent image generation model from Recraft. Every request requires at least one style reference image and generates a new image that reproduces the...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
recraft

Recraft: Recraft V4 Styles

Recraft V4 Styles is a style-consistent image generation model from Recraft. Every request requires at least one style reference image and generates a new image that reproduces the reference's rendering...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
alibaba

Alibaba: Wan 3.0

Wan 3.0 is a video generation model from Alibaba for text-to-video, image-to-video, and reference-guided video generation. It produces 480p, 720p, or 1080p video with durations from 2 to 30 seconds.

Context
Not provided
Input
text · image
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
heygen

HeyGen: Avatar IV

HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal...

Context
Not provided
Input
text · image · audio
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
meta

Meta: Muse Spark 1.2 Contributor

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.100Output: $0.200per 1M tokens
Ver no OpenRouter
deepseek

DeepSeek: DeepSeek V4 Flash Vision Exp

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

Context
1.0M
Input
text · image
Output
text
Input: $0.440Output: $1.32per 1M tokens
Ver no OpenRouter
tencent

Tencent: Hy-MT2-1.8B

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

Context
8K
Input
text
Output
text
Input: $0.044Output: $0.177per 1M tokens
Ver no OpenRouter
tencent

Tencent: Hy-MT2-30B-A3B

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...

Context
8K
Input
text
Output
text
Input: $0.074Output: $0.295per 1M tokens
Ver no OpenRouter
black-forest-labs

Black Forest Labs: FLUX Video Upscale

FLUX Video Upscale is a video upscaling model from Black Forest Labs. It enlarges a single source video by 1.5× to 3× while preserving its duration, with an optional prompt...

Context
Not provided
Input
text · video
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
tencent

Tencent: Hy-MT2-7B

Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

Context
8K
Input
text
Output
text
Input: $0.074Output: $0.295per 1M tokens
Ver no OpenRouter
liquid

LiquidAI: LFM2.5-Embedding-350M (free)

LFM2.5-Embedding-350M is a text embedding model from Liquid AI. It produces 1,024-dimensional embeddings for retrieval and semantic search. Successful OpenRouter requests and embeddings may be retained and used to train...

Context
1K
Input
text
Output
embeddings
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen3.8 27B

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

Context
1M
Input
text · image · video
Output
text
Input: $0.425Output: $2.55per 1M tokens
Ver no OpenRouter
nvidia

NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6B

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice...

Context
Not provided
Input
audio
Output
transcription
Input: $3.33Output: Freeper 1M tokens
Ver no OpenRouter
mistralai

Mistral: Voxtral Small 24B 2507 STT

Voxtral Small 24B 2507 STT is a speech transcription model from Mistral AI. It is suited for transcription, translation, and audio understanding workloads that benefit from its larger model capacity.

Context
Not provided
Input
audio
Output
transcription
Input: $50.00Output: Freeper 1M tokens
Ver no OpenRouter
mistralai

Mistral: Voxtral Mini 3B 2507

Voxtral Mini 3B 2507 is a speech and audio understanding model from Mistral AI. It is suited for transcription, translation, and compact audio processing workloads.

Context
Not provided
Input
audio
Output
transcription
Input: $16.67Output: Freeper 1M tokens
Ver no OpenRouter
bytedance-seed

ByteDance Seed: Seedream 5.0 Lite

Seedream 5.0 Lite is an image generation model from ByteDance Seed. It is suited for professional visual creation that benefits from web-connected retrieval, complex-prompt comprehension, visual references, and broad knowledge...

Context
Not provided
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
google

Google: Gemini 3.7 Flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.375Output: $1.88per 1M tokens
Ver no OpenRouter
google

Google: Gemini 3.7 Flash (batch)

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $0.188Output: $0.938per 1M tokens
Ver no OpenRouter
voyageai

VoyageAI by MongoDB: voyage-code-4

voyage-code-4 is a code embedding model from Voyage AI, a MongoDB company. It is designed for coding agents and code retrieval, with Matryoshka embeddings at 2048, 1024, 512, and 256...

Context
32K
Input
text
Output
embeddings
Input: $0.120Output: Freeper 1M tokens
Ver no OpenRouter
qwen

Qwen3 Reranker 8B

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG...

Context
41K
Input
text
Output
rerank
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen3 ASR 1.7B

Qwen3 ASR 1.7B is an automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline inference...

Context
Not provided
Input
audio
Output
transcription
Input: $7.50Output: Freeper 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen3 ASR 0.6B

Qwen3 ASR 0.6B is a compact automatic speech recognition model from Qwen. It supports multilingual language identification and transcription across 30 languages and 22 Chinese dialects, with streaming and offline...

Context
Not provided
Input
audio
Output
transcription
Input: $3.33Output: Freeper 1M tokens
Ver no OpenRouter
bytedance-seed

ByteDance Seed: Seedream 5.0 Pro

Seedream 5.0 Pro is an image generation and editing model from ByteDance Seed. It is suited for commercial visual-production workflows that require precise editing control, lifelike scenes, and natural rendering.

Context
Not provided
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
deepgram

Deepgram: Flux TTS (free)

Flux TTS is a text-to-speech model from Deepgram. It is suited for natural, expressive English speech synthesis across Deepgram's Flux voice catalog.

Context
Not provided
Input
text
Output
speech
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
bytedance

ByteDance: Seedance 2.0 Mini

Seedance 2.0 Mini is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video with image, video, and audio inputs. It...

Context
Not provided
Input
text · image · video · audio
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
bytedance-seed

ByteDance Seed: Seed 2.1 Turbo

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

Context
262K
Input
text · image · video
Output
text
Input: $0.500Output: $2.50per 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Context
1.0M
Input
text
Output
text
Input: $2.00Output: $6.00per 1M tokens
Ver no OpenRouter
bytedance-seed

ByteDance Seed: Seed-2.0-Code

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

Context
262K
Input
text · image · video
Output
text
Input: $0.500Output: $3.00per 1M tokens
Ver no OpenRouter
deepseek

DeepSeek: DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

Context
1.0M
Input
text
Output
text
Input: $1.32Output: $3.96per 1M tokens
Ver no OpenRouter
x-ai

SpaceXAI: Grok 4.6

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Context
500K
Input
text · image · file
Output
text
Input: $2.00Output: $6.00per 1M tokens
Ver no OpenRouter
x-ai

xAI: Grok Imagine Image 2.0

Grok Imagine Image 2.0 is an image generation and editing model from xAI. It is suited for creating images from text prompts and editing images from references, with low and...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
liquid

LiquidAI: LFM2.5-2.6B (free)

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

Context
66K
Input
text
Output
text
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
nvidia

NVIDIA: Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Context
262K
Input
text
Output
text
Input: $0.100Output: $0.250per 1M tokens
Ver no OpenRouter
nvidia

NVIDIA: Nemotron 3.5 Lightning (free)

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Context
1M
Input
text
Output
text
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
sakana

Sakana: Sakana Namazu

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

Context
262K
Input
text · image · file
Output
text
Input: $0.950Output: $4.00per 1M tokens
Ver no OpenRouter
upstage

Upstage: Solar Pro 4

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

Context
524K
Input
text
Output
text
Input: $0.030Output: $0.120per 1M tokens
Ver no OpenRouter
meta

Meta: Muse Glimmer 30B

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

Context
131K
Input
text · image
Output
text
Input: $0.350Output: $1.50per 1M tokens
Ver no OpenRouter
bytedance

ByteDance: Seedance 2.5

Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up...

Context
Not provided
Input
text · image · video · audio
Output
video
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter
openai

OpenAI: GPT Transcribe

GPT Transcribe is a high-accuracy speech-to-text model from OpenAI. It is suited for recorded audio, streamed file transcription, and committed Realtime turns, with free-form context, keyword hints, and multiple language...

Context
Not provided
Input
audio
Output
transcription
Input: $4500.00Output: Freeper 1M tokens
Ver no OpenRouter
meta

Meta: Muse Spark 1.2

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

Context
1.0M
Input
text · image · video · file · audio
Output
text
Input: $1.25Output: $4.25per 1M tokens
Ver no OpenRouter
qwen

Qwen: Qwen Image 3

Qwen Image 3 is a unified image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with a richer world...

Context
66K
Input
text · image
Output
image
Input: FreeOutput: Freeper 1M tokens
Ver no OpenRouter