inception logo
inception

Inception: Mercury 2.5 Preview

Descrizione della fonte (inglese)

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Panoramica

Specifiche del modello

Contesto
260.000 tokens
Output massimo
65.536 tokens
Architettura
text->text
Tokenizer
Other
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$0.04
per milione di token
Output
$0.15
per milione di token
Lettura cache
$0.004
per milione di token
API

Configurazione del modello

Ragionamento

Ragionamento obbligatorio
No
Parametri predefiniti
medium
Capacità e modalità
high, medium, low, none
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"inception/mercury-2.5-preview","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
text
Output
text
Ragionamento

· high · medium · low · none

API

Parametri API supportati

include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools
Provider disponibili

1 Provider

Verificato: 7 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

Inception

Non indicato
Contesto
260K
Output massimo
66K
Input
$0.04
Output
$0.15
Lettura cache
$0.004
Scrittura cache
Non indicato
inception

1 modelli

Tutti i modelli