qwen logo
qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Panoramica

Specifiche del modello

Context
1.048.576 tokens
Output massimo
262.144 tokens
Architettura
text->text
Tokenizer
Qwen
Limite di conoscenza
Not provided
Moderato
No
Capacità e modalità

InputOutput

Input
text
Output
text
Ragionamento

· xhigh · medium · low

API

Parametri API supportati

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Provider disponibili

7 provider

DeepInfra

fp4
Context
262K
Output massimo
131K
Input
$2.00
Output
$6.00
Lettura cache
$0.200
Scrittura cache
Not provided

Modal

unknown
Context
1M
Output massimo
262K
Input
$2.00
Output
$6.00
Lettura cache
$0.250
Scrittura cache
Not provided

Novita

unknown
Context
1M
Output massimo
131K
Input
$2.00
Output
$6.00
Lettura cache
$0.250
Scrittura cache
Not provided

SiliconFlow

fp8
Context
1.0M
Output massimo
131K
Input
$2.00
Output
$6.00
Lettura cache
$0.250
Scrittura cache
Not provided

Alibaba

unknown
Context
1M
Output massimo
131K
Input
$2.00
Output
$6.00
Lettura cache
$0.250
Scrittura cache
$2.50

Together

unknown
Context
1.0M
Output massimo
909K
Input
$2.50
Output
$6.25
Lettura cache
$0.500
Scrittura cache
Not provided

Venice

unknown
Context
262K
Output massimo
66K
Input
$2.50
Output
$7.50
Lettura cache
$0.313
Scrittura cache
Not provided