The instruction tells the model what is present, what each subject is doing, how the image should look, and where the scene takes place. Black Forest Labs describes a useful starting structure as subject, action, style, and context. The format is flexible. Natural language works well, and a complete sentence often carries spatial relationships more clearly than a stack of disconnected tags.
This page turns visible evidence into that kind of instruction. It reads the main subject, environment, composition, lighting, palette, camera feel, materials, style, and mood. The analysis then writes a FLUX-specific version rather than copying the General prompt. That version can say which object is left or right, what overlaps, where a person faces, and how the light interacts with a surface.
The generated description is an inference. It is not the original production record. The pixels cannot disclose a seed, hidden prompt, model variant, reference stack, post-processing step, or exact camera setting. The useful task is to rebuild the scene specification from what can be seen, then let you edit uncertain details before generation.
FLUX model variants do not all expose identical controls. Current Black Forest Labs guidance covers FLUX.1, FLUX.1 Kontext, and FLUX.2, with model-specific differences called out in its documentation. Keep the descriptive prompt portable, then set resolution, aspect ratio, seed, guidance, or reference inputs inside the destination you actually use.