Model specifications
- Context
- 1,000,000 tokens
- Maximum output
- 131,072 tokens
- Architecture
- text+image+video->text
- Tokenizer
- Qwen
- Knowledge cutoff
- Not provided
- Moderated
- No
Input → Output
Yes · xhigh · medium · low
Supported API parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p11 providers
Reka
fp8- Context
- 262K
- Maximum output
- 131K
- Input
- $0.350
- Output
- $2.55
- Cache read
- $0.050
- Cache write
- Not provided
AkashML
fp8- Context
- 262K
- Maximum output
- 131K
- Input
- $0.350
- Output
- $2.55
- Cache read
- $0.050
- Cache write
- Not provided
Chutes
fp8- Context
- 262K
- Maximum output
- 66K
- Input
- $0.350
- Output
- $2.75
- Cache read
- $0.035
- Cache write
- Not provided
Parasail
fp8- Context
- 262K
- Maximum output
- 236K
- Input
- $0.350
- Output
- $3.20
- Cache read
- $0.050
- Cache write
- Not provided
Phala
unknown- Context
- 262K
- Maximum output
- 236K
- Input
- $0.400
- Output
- $3.00
- Cache read
- $0.150
- Cache write
- Not provided
CoreWeave
fp8- Context
- 262K
- Maximum output
- 236K
- Input
- $0.400
- Output
- $3.00
- Cache read
- $0.150
- Cache write
- Not provided
Novita
unknown- Context
- 1M
- Maximum output
- 131K
- Input
- $0.420
- Output
- $3.00
- Cache read
- $0.085
- Cache write
- Not provided
Alibaba
unknown- Context
- 1M
- Maximum output
- 131K
- Input
- $0.425
- Output
- $2.55
- Cache read
- $0.085
- Cache write
- $0.531
Cloudflare
unknown- Context
- 262K
- Maximum output
- 236K
- Input
- $0.450
- Output
- $3.20
- Cache read
- $0.050
- Cache write
- Not provided
Venice
fp8- Context
- 262K
- Maximum output
- 66K
- Input
- $0.450
- Output
- $3.20
- Cache read
- Not provided
- Cache write
- Not provided
Io Net
fp8- Context
- 66K
- Maximum output
- 59K
- Input
- $0.480
- Output
- $3.40
- Cache read
- $0.250
- Cache write
- Not provided