
You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt. Think step by step about the request before writing the answer: - What is the subject and mood? - What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption. - What composition, framing, and grounded details will help the text-to-image model? Then output a single expanded prompt paragraph. Follow these rules strictly: 1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them. 2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout. 3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body. 4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes. 5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them. 6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown. 7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction. 8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy. 9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent. User's Input: Three women are seated side-by-side on a light-colored sofa with patterned cushions in an indoor setting featuring a green wall and a tall white decorative vase on the right. Each woman has a red ball gag in her mouth and is bound with beige rope. The woman on the left, with dark hair, wears a striped black and white long-sleeved top, a black skirt, black stockings, and black high-heeled sandals. Her torso and arms and ankles are tightly bound. The middle woman, also with dark hair, is dressed in a dark brown v-neck top and a black skirt, with light-colored stockings and black high-heeled sandals. She is bound around her torso and arms and ankles. The woman on the right, with blonde hair, wears a white partially unbuttoned shirt, a denim skirt, light-colored stockings, and black boots. Her torso and arms are boud, and her ankles secured . All three women maintain an upright posture, looking generally forward. The scene is in an elegant living room in a modern luxury house, featuring clean architectural lines, floor-to-ceiling windows, a refined neutral color palette, designer furniture, a marble coffee table, soft ambient lighting, Realistic textures, sophisticated atmosphere, cinematic interior photography, highly detailed, photorealistic.
Parameters used to generate this content
The node graph used to generate this content — pan and zoom to explore, or download it to run in ComfyUI.
AI models used to generate this content