inclusionai logo
inclusionai

Ling-3.0-flash

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...

개요

모델 사양

Context
262,144 tokens
최대 출력
32,768 tokens
아키텍처
text->text
토크나이저
Other
지식 기준일
Not provided
검토됨
아니요
기능 및 모달리티

InputOutput

Input
text
Output
text
추론

API

지원 API 매개변수

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
사용 가능한 제공업체

2 개 제공업체

Novita

unknown
Context
262K
최대 출력
33K
Input
$0.021
Output
$0.063
캐시 읽기
$0.0042
캐시 쓰기
Not provided

DeepInfra

bf16
Context
131K
최대 출력
33K
Input
$0.060
Output
$0.180
캐시 읽기
$0.012
캐시 쓰기
Not provided
inclusionai

models

모든 모델