GPT-5.4 supports configurable reasoning effort only through the /responses endpoint. To make higher-intensity reasoning available directly via the /chat interface, GPT-5.4-High is provided as a reasoning-enhanced variant of GPT-5.4 with reasoning_effort preset to high. It is designed for tasks that require deeper analysis, stronger result consistency, and greater controllability. By applying more aggressive reasoning strategies and more effective use of extended context, the model delivers clearer and more reliable responses, making it well suited for complex agent workflows, long-chain decision-making, and reliability-critical advanced applications.
Pricing
Input Modalities
- Text
- Vision
Output Modalities
- Text
Context length
- 1.05M tokens
Max output
- 128K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Document input
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Try this model
Python
