
You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt. Think step by step about the request before writing the answer: - What is the subject and mood? - What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption. - What composition, framing, and grounded details will help the text-to-image model? Then output a single expanded prompt paragraph. Follow these rules strictly: 1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them. 2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout. 3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body. 4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes. 5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them. 6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown. 7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction. 8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy. 9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent. User's Input: A full-length, eye-level shot captures a woman and a man in an indoor setting. The woman, with long blonde wavy hair and a German face, is dressed in a traditional Bavarian-style costume consisting of a white, ruffled off-the-shoulder blouse, a brown laced bodice with pink ribbons, a green skirt with patterned trim, and a white apron. She wears black sheer stockings and black high-heeled pumps. Her wrists are tightly bound behind her back with thick beige rope, which wraps multiple times to secure her arms together and also is wrapped over her torso in a knotted shibari. A white cloth cleave gag is pulled taut through her open mouth and knotted behind her head. The woman is clearly annoyed and tries to turn her head to protest. Standing behind the woman to her right is a middle aged man dressed in casual clothes including a short-sleeved polo shirt, leather trousers with a black belt, and black boots. His face is rugged and he has a long beard, but he appears happy and enjoying. He has both hands positioned near the woman's cheeks one hand on each side, adjusting the cleave gag on her face. The background features a traditional bavarian tavern full of bavarans enjoying their beers. The lighting is bright and even, highlighting the textures of the clothing and materials. The overall mood is dramatic and staged, focusing on the captive subject.
Parameters used to generate this content
The node graph used to generate this content — pan and zoom to explore, or download it to run in ComfyUI.
AI models used to generate this content