qwen logo
qwen

Qwen: Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Overview

Model specifications

Context
1,048,576 tokens
Maximum output
262,144 tokens
Architecture
text->text
Tokenizer
Qwen
Knowledge cutoff
Not provided
Moderated
No
Capabilities

InputOutput

Input
text
Output
text
Reasoning

Yes · xhigh · medium · low

API

Supported API parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Available providers

7 providers

DeepInfra

fp4
Context
262K
Maximum output
131K
Input
$2.00
Output
$6.00
Cache read
$0.200
Cache write
Not provided

Modal

unknown
Context
1M
Maximum output
262K
Input
$2.00
Output
$6.00
Cache read
$0.250
Cache write
Not provided

Novita

unknown
Context
1M
Maximum output
131K
Input
$2.00
Output
$6.00
Cache read
$0.250
Cache write
Not provided

SiliconFlow

fp8
Context
1.0M
Maximum output
131K
Input
$2.00
Output
$6.00
Cache read
$0.250
Cache write
Not provided

Alibaba

unknown
Context
1M
Maximum output
131K
Input
$2.00
Output
$6.00
Cache read
$0.250
Cache write
$2.50

Together

unknown
Context
1.0M
Maximum output
909K
Input
$2.50
Output
$6.25
Cache read
$0.500
Cache write
Not provided

Venice

unknown
Context
262K
Maximum output
66K
Input
$2.50
Output
$7.50
Cache read
$0.313
Cache write
Not provided
qwen

models

All models