
You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt. Think step by step about the request before writing the answer: - What is the subject and mood? - What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption. - What composition, framing, and grounded details will help the text-to-image model? Then output a single expanded prompt paragraph. Follow these rules strictly: 1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them. 2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout. 3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body. 4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes. 5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them. 6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown. 7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction. 8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy. 9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent. User's Input: A black-and-white manga panel showing a person viewed from behind, standing waist-deep in a body of water. The person has long, straight black hair that falls down their back, and their bare shoulders and upper back are visible. The water ripples around them, with concentric circles emanating from their body. Steam or mist rises from the water surface, particularly around the person's shoulders and head. In the background, a range of snow-capped mountains rises under a gray sky. The mountains are detailed with shading to indicate texture and depth, and their slopes are covered with patches of snow and dark evergreen trees. On the right side of the frame, the dark silhouette of a coniferous tree is visible. A speech bubble is positioned to the upper right of the person's head, containing the text "I AM COLD..". The overall composition is centered on the person, with the mountains creating a strong vertical backdrop. The lighting is diffuse, with soft shadows on the mountains and subtle highlights on the snow and water. The color palette is monochromatic, using various shades of gray, black, and white to create contrast and depth. The texture of the water is rendered with fine lines to show ripples and reflections, while the rocks surrounding the water have rough, detailed surfaces. The hair is rendered with smooth, solid black areas, and the skin is depicted with minimal shading to emphasize its smoothness. The speech bubble is a simple white oval with a thin black outline, and the text inside is in a clean, sans-serif font, all uppercase, with the words "I AM COLD.." centered within the bubble.
Parameters used to generate this content
The node graph used to generate this content — pan and zoom to explore, or download it to run in ComfyUI.
AI models used to generate this content