Model specifications
- Context
- 1,000,000 tokens
- Maximum output
- 65,536 tokens
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- Not provided
- Moderated
- No
Input → Output
No
Supported API parameters
include_reasoningmax_tokensreasoningseedtemperaturetool_choicetoolstop_p2 providers
DeepInfra
bf16- Context
- 262K
- Maximum output
- 131K
- Input
- $0.080
- Output
- $0.200
- Cache read
- $0.040
- Cache write
- Not provided
CoreWeave
bf16- Context
- 262K
- Maximum output
- 236K
- Input
- $0.100
- Output
- $0.250
- Cache read
- $0.050
- Cache write
- Not provided