Especificaciones del modelo
- Context
- 40.960 tokens
- Salida máxima
- 36.864 tokens
- Arquitectura
- text->rerank
- Tokenizador
- Qwen3
- Corte de conocimiento
- Not provided
- Moderado
- No
Input → Output
Parámetros API compatibles
frequency_penaltylogit_biaslogprobsmax_tokenspresence_penaltyrepetition_penaltyresponse_formatstopstructured_outputstemperaturetop_ktop_logprobstop_p1 Proveedor
Fireworks
unknown- Context
- 41K
- Salida máxima
- 37K
- Input
- Free
- Output
- Free
- Lectura de caché
- Not provided
- Escritura de caché
- Not provided