inclusionai logo
inclusionai

Ling-3.0-flash

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...

模型概览

模型规格

上下文
262,144 tokens
最大输出
32,768 tokens
架构
text->text
分词器
Other
知识截止时间
暂未提供
内容审核
能力与模态

输入输出

输入
text
输出
text
推理能力

API

支持的 API 参数

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
可用提供方

2 个提供方

Novita

unknown
上下文
262K
最大输出
33K
输入
$0.021
输出
$0.063
缓存读取
$0.0042
缓存写入
暂未提供

DeepInfra

bf16
上下文
131K
最大输出
33K
输入
$0.060
输出
$0.180
缓存读取
$0.012
缓存写入
暂未提供
inclusionai

个模型

全部模型