🖼️ Loading...
    You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt.

Think step by step about the request before writing the answer:
- What is the subject and mood?
- What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption.
- What composition, framing, and grounded details will help the text-to-image model?

Then output a single expanded prompt paragraph.

Follow these rules strictly:
1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them.
2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout.
3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body.
4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes.
5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them.
6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown.
7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction.
8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy.
9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent.

User's Input:

This is a hyper-realistic photograph depicting a tense, dramatic scene inside a rustic, dimly lit room. The central focus is on two individuals: a muscular man with a rugged, weathered appearance, and a young woman. The man, who has dark skin, a well-defined physique, and short, dark hair, is wearing a white T-shirt and blue jeans with simple black mary jane shoes. He is standing behind the woman, his hands firmly pressing her torso. His facial expression is stern and intense.
The woman, with light skin and shoulder-length brown hair, is dressed in a torn white T-shirt with the words \"WHAT R U LOOKING?\" printed in bold black letters, and blue jeans. Her expression is one of distress and fear as she is bound with ropes around her wrists and ankles, and her legs are tied together.
The girl is bound with her arms behind her back with lots of light brown rope and her legs joined and held together by ropes, wrapped tighly around her ankles, feet, knees and thighs. She is cloth gagged with a white bandana, her lips pressed tightly onto it. Her crossed wrists are wrapped with a rope, behind her. Her arms are tied behind her back. Ropes are forming elegant patterns over her chest. 
The room has a vintage, Western theme, with wooden walls and a window that reveals an expansive, sunlit landscape outside. There are various objects scattered around: a large barrel, a rolled-up poster, and a desk cluttered with papers and a typewriter. A man in a cowboy hat and shirt is seated at the desk, seemingly oblivious to the scene. The overall color palette is earthy, with warm tones dominating the scene, adding to the gritty, realistic atmosphere.
    Prompt

    You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt. Think step by step about the request before writing the answer: - What is the subject and mood? - What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption. - What composition, framing, and grounded details will help the text-to-image model? Then output a single expanded prompt paragraph. Follow these rules strictly: 1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them. 2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout. 3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body. 4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes. 5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them. 6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown. 7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction. 8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy. 9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent. User's Input: This is a hyper-realistic photograph depicting a tense, dramatic scene inside a rustic, dimly lit room. The central focus is on two individuals: a muscular man with a rugged, weathered appearance, and a young woman. The man, who has dark skin, a well-defined physique, and short, dark hair, is wearing a white T-shirt and blue jeans with simple black mary jane shoes. He is standing behind the woman, his hands firmly pressing her torso. His facial expression is stern and intense. The woman, with light skin and shoulder-length brown hair, is dressed in a torn white T-shirt with the words \"WHAT R U LOOKING?\" printed in bold black letters, and blue jeans. Her expression is one of distress and fear as she is bound with ropes around her wrists and ankles, and her legs are tied together. The girl is bound with her arms behind her back with lots of light brown rope and her legs joined and held together by ropes, wrapped tighly around her ankles, feet, knees and thighs. She is cloth gagged with a white bandana, her lips pressed tightly onto it. Her crossed wrists are wrapped with a rope, behind her. Her arms are tied behind her back. Ropes are forming elegant patterns over her chest. The room has a vintage, Western theme, with wooden walls and a window that reveals an expansive, sunlit landscape outside. There are various objects scattered around: a large barrel, a rolled-up poster, and a desk cluttered with papers and a typewriter. A man in a cowboy hat and shirt is seated at the desk, seemingly oblivious to the scene. The overall color palette is earthy, with warm tones dominating the scene, adding to the gritty, realistic atmosphere.

    Generation Settings

    Parameters used to generate this content

    Created with
    ComfyUI
    CFG Scale1
    Sampler
    Euler
    Seed3063273866
    Steps10
    ComfyUI Workflow

    The node graph used to generate this content — pan and zoom to explore, or download it to run in ComfyUI.

    Info
    Image
    Likes
    0
    Created
    7/3/2026
    Source
    CivitAI
    Models Used

    AI models used to generate this content

    Actions