Model specifications
- Context
- 8,192 tokens
- Maximum output
- 4,096 tokens
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- Not provided
- Moderated
- No
Input → Output
Supported API parameters
max_completion_tokensmax_tokensresponse_formatstopstructured_outputstemperature1 Provider
Tencent
fp8- Context
- 8K
- Maximum output
- 4K
- Input
- $0.074
- Output
- $0.295
- Cache read
- Not provided
- Cache write
- Not provided