Model specifications
- Context
- 1,050,000 tokens
- Maximum output
- 128,000 tokens
- Architecture
- text+image+file->text
- Tokenizer
- GPT
- Knowledge cutoff
- 2026-02-16
- Moderated
- Yes
Input → Output
Yes · max · xhigh · high · medium · low · none
Supported API parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstool_choicetools7 providers
OpenAI
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.100
- Output
- $0.600
- Cache read
- $0.010
- Cache write
- $0.125
Azure
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.200
- Output
- $1.20
- Cache read
- $0.020
- Cache write
- $0.250
OpenAI
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.200
- Output
- $1.20
- Cache read
- $0.020
- Cache write
- $0.250
Amazon Bedrock
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.220
- Output
- $1.32
- Cache read
- $0.022
- Cache write
- $0.275
Azure
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.220
- Output
- $1.32
- Cache read
- $0.022
- Cache write
- $0.275
Azure
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.220
- Output
- $1.32
- Cache read
- $0.022
- Cache write
- $0.275
OpenAI
unknown- Context
- 1.1M
- Maximum output
- 128K
- Input
- $0.400
- Output
- $2.40
- Cache read
- $0.040
- Cache write
- $0.500