Model specifications
- Context
- 1,048,576 tokens
- Maximum output
- 64,000 tokens
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- Not provided
- Moderated
- No
Input → Output
Yes · high · low · none
Supported API parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools1 Provider
Tencent
fp8- Context
- 1.0M
- Maximum output
- 64K
- Input
- $0.834
- Output
- $2.50
- Cache read
- $0.042
- Cache write
- Not provided