Caractéristiques du modèle
- Context
- 1 000 000 tokens
- Sortie maximale
- 65 536 tokens
- Architecture
- text->text
- Tokenizer
- Other
- Limite des connaissances
- Not provided
- Modéré
- Non
Input → Output
Non
Paramètres API pris en charge
include_reasoningmax_tokensreasoningseedtemperaturetool_choicetoolstop_p2 fournisseurs
DeepInfra
bf16- Context
- 262K
- Sortie maximale
- 131K
- Input
- $0.080
- Output
- $0.200
- Lecture du cache
- $0.040
- Écriture du cache
- Not provided
CoreWeave
bf16- Context
- 262K
- Sortie maximale
- 236K
- Input
- $0.100
- Output
- $0.250
- Lecture du cache
- $0.050
- Écriture du cache
- Not provided