thinkingmachines logo
thinkingmachines

Thinking Machines: Inkling Small

来源介绍(英文)

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

模型概览

模型规格

上下文
1,048,576 tokens
最大输出
262,144 tokens
架构
text+image+audio->text
分词器
Other
知识截止时间
暂未提供
内容审核
OPENROUTER

完整价格

同步自 OpenRouter。Token 价格均按每百万 Token 展示。

输入
$0.45
每百万 Token
输出
$1.2
每百万 Token
缓存读取
$0.1
每百万 Token
API

模型配置

推理能力

是否必须推理
默认参数
high
能力与模态
max, high, medium, low, minimal, none
API

快速调用

在本地设置 OPENROUTER_API_KEY。Python 需安装 requests;JavaScript 在 Node.js 中运行。密钥应保存在服务端。

API 文档

curl --fail-with-body https://openrouter.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $OPENROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"thinkingmachines/inkling-small","messages":[{"role":"user","content":"Hello!"}]}'
能力与模态

输入输出

输入
textimageaudio
输出
text
推理能力

· max · high · medium · low · minimal · none

API

支持的 API 参数

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyseedstoptemperaturetool_choicetoolstop_ktop_p
可用提供方

2 个提供方

核验日期: 2026年9月23日

在 OpenRouter 查看实时 Provider

Provider 可用性、延迟、吞吐量和路由会持续变化,请前往来源页面查看实时运行数据。

OpenRouter

DeepInfra

fp8
上下文
524K
最大输出
262K
输入
$0.45
输出
$1.2
缓存读取
$0.1
缓存写入
暂未提供

BaseTen

fp8
上下文
1.0M
最大输出
33K
输入
$0.5
输出
$1.2
缓存读取
$0.1
缓存写入
暂未提供
thinkingmachines

3 个模型

全部模型