inclusionai logo
inclusionai

Ling-3.0-flash

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...

Überblick

Modellspezifikationen

Context
262.144 tokens
Maximale Ausgabe
32.768 tokens
Architektur
text->text
Tokenizer
Other
Wissensstand
Nicht angegeben
Moderiert
Nein
Fähigkeiten und Modalitäten

EingabeAusgabe

Eingabe
text
Ausgabe
text
Reasoning

Ja

API

Unterstützte API-Parameter

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
Verfügbare Anbieter

2 Anbieter

Novita

unknown
Context
262K
Maximale Ausgabe
33K
Eingabe
$0.021
Ausgabe
$0.063
Cache-Lesen
$0.0042
Cache-Schreiben
Nicht angegeben

DeepInfra

bf16
Context
131K
Maximale Ausgabe
33K
Eingabe
$0.060
Ausgabe
$0.180
Cache-Lesen
$0.012
Cache-Schreiben
Nicht angegeben
inclusionai

models

Alle Modelle