qwen logo
qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Resumen

Especificaciones del modelo

Context
1.048.576 tokens
Salida máxima
262.144 tokens
Arquitectura
text->text
Tokenizador
Qwen
Corte de conocimiento
Not provided
Moderado
No
Capacidades y modalidades

InputOutput

Input
text
Output
text
Razonamiento

· xhigh · medium · low

API

Parámetros API compatibles

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Proveedores disponibles

7 proveedores

DeepInfra

fp4
Context
262K
Salida máxima
131K
Input
$2.00
Output
$6.00
Lectura de caché
$0.200
Escritura de caché
Not provided

Modal

unknown
Context
1M
Salida máxima
262K
Input
$2.00
Output
$6.00
Lectura de caché
$0.250
Escritura de caché
Not provided

Novita

unknown
Context
1M
Salida máxima
131K
Input
$2.00
Output
$6.00
Lectura de caché
$0.250
Escritura de caché
Not provided

SiliconFlow

fp8
Context
1.0M
Salida máxima
131K
Input
$2.00
Output
$6.00
Lectura de caché
$0.250
Escritura de caché
Not provided

Alibaba

unknown
Context
1M
Salida máxima
131K
Input
$2.00
Output
$6.00
Lectura de caché
$0.250
Escritura de caché
$2.50

Together

unknown
Context
1.0M
Salida máxima
909K
Input
$2.50
Output
$6.25
Lectura de caché
$0.500
Escritura de caché
Not provided

Venice

unknown
Context
262K
Salida máxima
66K
Input
$2.50
Output
$7.50
Lectura de caché
$0.313
Escritura de caché
Not provided