OpenAI guidance favors clear requests with a purpose, subject, action, setting, and visual style. Add framing, lighting, materials, and constraints only when they affect the result. A straightforward request may need just one to three sentences.
This generator writes the model-specific brief from the visual choices in your reference. It identifies what the image depicts and how the composition works, then expresses the result as instructions a conversational image system can follow. The wording can say “preserve,” “replace,” “remove,” or “keep unchanged” because ChatGPT Images supports iterative creation and editing through natural language.
The generated brief is not the original source prompt. The image cannot reveal the conversation that preceded it, uploaded references, selection masks, system behavior, model version, seed, policy decisions, or later edits. The service creates a new brief from what is visible. You should review every inferred detail before treating it as intentional.
One request on this site produces six prompt versions. This version differs from the keyword-focused phrasing used for Midjourney, the relationship-heavy FLUX description, and the conditioning fields used in Stable Diffusion. It explains the task, the desired output, and the boundaries of the change in language that fits a conversational editing loop.