thinkingmachines logo
thinkingmachines

Thinking Machines: Inkling Small

Descrizione della fonte (inglese)

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Panoramica

Specifiche del modello

Contesto
1.048.576 tokens
Output massimo
262.144 tokens
Architettura
text+image+audio->text
Tokenizer
Other
Limite di conoscenza
Non indicato
Moderato
No
OPENROUTER

Prezzi completi

Tariffe sincronizzate da OpenRouter, per milione di token.

Input
$0.45
per milione di token
Output
$1.2
per milione di token
Lettura cache
$0.1
per milione di token
API

Configurazione del modello

Ragionamento

Ragionamento obbligatorio
No
Parametri predefiniti
high
Capacità e modalità
max, high, medium, low, minimal, none
API

Avvio rapido

Imposta OPENROUTER_API_KEY localmente. Python richiede requests; JavaScript viene eseguito in Node.js. Conserva la chiave sul server.

Documentazione API

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"thinkingmachines/inkling-small","messages":[{"role":"user","content":"Hello!"}]}'
Capacità e modalità

InputOutput

Input
textimageaudio
Output
text
Ragionamento

· max · high · medium · low · minimal · none

API

Parametri API supportati

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyseedstoptemperaturetool_choicetoolstop_ktop_p
Provider disponibili

2 provider

Verificato: 23 settembre 2026

Provider live su OpenRouter

Disponibilità, latenza, throughput e routing cambiano continuamente. Consulta la fonte per i dati correnti.

OpenRouter

DeepInfra

fp8
Contesto
524K
Output massimo
262K
Input
$0.45
Output
$1.2
Lettura cache
$0.1
Scrittura cache
Non indicato

BaseTen

fp8
Contesto
1.0M
Output massimo
33K
Input
$0.5
Output
$1.2
Lettura cache
$0.1
Scrittura cache
Non indicato
thinkingmachines

3 modelli

Tutti i modelli