inception logo
inception

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). $0.25 per million input tokens, $0.75 per million output tokens. 128,000 token context window, maximum output of 50,000 tokens. Includes independent benchmarks from Artificial Analysis.

Overview

Model specifications

Context
128,000 tokens
Maximum output
50,000 tokens
Architecture
text->text
Tokenizer
Other
Knowledge cutoff
Not provided
Moderated
No
OPENROUTER

Complete pricing

Synchronized OpenRouter rates. Token prices are shown per one million tokens.

Input
$0.25
/M tokens
Output
$0.75
/M tokens
Cache read
$0.025
/M tokens
API

Model configuration

Default parameters

temperature
0.75

Reasoning

Moderated
No
Default parameters
medium
Capabilities
high, medium, low, none
API

Quick start

Call this exact model ID through OpenRouter’s OpenAI-compatible API.

curl https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"inception/mercury-2","messages":[{"role":"user","content":"Hello!"}]}'
Capabilities

InputOutput

Input
text
Output
text
Reasoning

Yes · high · medium · low · none

API

Supported API parameters

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools
Available providers

1 Provider

Live providers on OpenRouter

Provider availability, latency, throughput and routing can change continuously. Open the source page for current operational data.

OpenRouter

Inception

unknown
Context
128K
Maximum output
50K
Input
$0.250
Output
$0.750
Cache read
$0.025
Cache write
Not provided