Are GPT-5.4 and GPT-5.4 Thinking two different models?
GPT-5.4 Thinking is the ChatGPT product form of GPT-5.4; API calls use gpt-5.4. Experiences such as process prompts and interactive adjustments in ChatGPT cannot be directly regarded as part of an API response; gpt-5.4-pro is a separate variant.
How should images be submitted to GPT-5.4?
When using Chat Completions, place text and image_url in the content array of the same message, and describe what you want recognized, compared, or explained. Images can be set to auto, low, or high detail; when analyzing small text and complex interfaces, clear screenshots are usually more important than vague questions.
Can GPT-5.4 automatically execute code and click web pages?
The model has relevant planning and instruction-generation capabilities, but executing actions requires tools provided by the application. After a function call is returned, the program should execute it and send back the results; browser or desktop operations also require the corresponding environment. Simply sending an operational request will not automatically establish a complete execution and verification workflow.
How can I control GPT-5.4's reasoning depth?
GPT-5.4's official native evaluations used reasoning settings from none to xhigh. On this platform, Responses provides the reasoning object, and Chat Completions provides the reasoning_effort field, but the same set of levels cannot be directly applied to both: the current Chat Completions parameter enumeration does not include none. When configuring, use levels valid for gpt-5.4 in the selected endpoint, and compare response quality and wait time using real tasks; values such as minimal and max in shared enumerations should not be directly treated as settings supported by this model.
How can GPT-5.4 maintain multi-turn conversations?
When using Chat Completions, put relevant history into messages; when using Responses, organize input and related conversation content according to the documentation. Provide the latest materials, revision goals, and key constraints in each turn; for longer tasks, retain interim summaries and a final version that can be independently reviewed.