Specifiche del modello
- Context
- 262.144 tokens
- Output massimo
- 131.072 tokens
- Architettura
- text->text
- Tokenizer
- Other
- Limite di conoscenza
- Not provided
- Moderato
- No
Input → Output
No
Parametri API supportati
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p2 provider
DeepInfra
bf16- Context
- 262K
- Output massimo
- 131K
- Input
- $0.080
- Output
- $0.200
- Lettura cache
- $0.040
- Scrittura cache
- Not provided
CoreWeave
bf16- Context
- 262K
- Output massimo
- 236K
- Input
- $0.100
- Output
- $0.250
- Lettura cache
- $0.050
- Scrittura cache
- Not provided