qwen logo
qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Présentation

Caractéristiques du modèle

Context
1 048 576 tokens
Sortie maximale
262 144 tokens
Architecture
text->text
Tokenizer
Qwen
Limite des connaissances
Not provided
Modéré
Non
Capacités et modalités

InputOutput

Input
text
Output
text
Raisonnement

Oui · xhigh · medium · low

API

Paramètres API pris en charge

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Fournisseurs disponibles

7 fournisseurs

DeepInfra

fp4
Context
262K
Sortie maximale
131K
Input
$2.00
Output
$6.00
Lecture du cache
$0.200
Écriture du cache
Not provided

Modal

unknown
Context
1M
Sortie maximale
262K
Input
$2.00
Output
$6.00
Lecture du cache
$0.250
Écriture du cache
Not provided

Novita

unknown
Context
1M
Sortie maximale
131K
Input
$2.00
Output
$6.00
Lecture du cache
$0.250
Écriture du cache
Not provided

SiliconFlow

fp8
Context
1.0M
Sortie maximale
131K
Input
$2.00
Output
$6.00
Lecture du cache
$0.250
Écriture du cache
Not provided

Alibaba

unknown
Context
1M
Sortie maximale
131K
Input
$2.00
Output
$6.00
Lecture du cache
$0.250
Écriture du cache
$2.50

Together

unknown
Context
1.0M
Sortie maximale
909K
Input
$2.50
Output
$6.25
Lecture du cache
$0.500
Écriture du cache
Not provided

Venice

unknown
Context
262K
Sortie maximale
66K
Input
$2.50
Output
$7.50
Lecture du cache
$0.313
Écriture du cache
Not provided