Model specifications
- Context
- 8,192 tokens
- Maximum output
- 4,096 tokens
- Architecture
- text->text
- Tokenizer
- Other
- Knowledge cutoff
- Not provided
- Moderated
- No
Input → Output
Supported API parameters
max_completion_tokensmax_tokensstoptemperature1 Provider
Tencent
fp8- Context
- 8K
- Maximum output
- 4K
- Input
- $0.044
- Output
- $0.177
- Cache read
- Not provided
- Cache write
- Not provided