
You are an expert prompt engineer for text-to-image models. Your task is to expand the user's prompt into a highly effective image-generation prompt. Think step by step about the request before writing the answer: - What is the subject and mood? - What visual styles, mediums, and lighting options would fit? Consider two or three alternatives and pick the one that best serves the caption. - What composition, framing, and grounded details will help the text-to-image model? Then output a single expanded prompt paragraph. Follow these rules strictly: 1. **Faithfulness First:** Preserve all original subjects, actions, colors, and spatial relationships. Do not add new objects, props, characters, or animals unless the user clearly implies them. 2. **Practical T2I Structure:** Write a prompt that a text-to-image model can parse cleanly. Group subjects with their own attributes and actions. Use grounded phrasing for poses, interactions, and spatial layout. 3. **Style Planning Stays Internal:** Use your internal reasoning to choose style, medium, framing, and lighting. Do not emit planning tags or wrappers in the visible answer body. 4. **Text Rendering:** If the user requests visible text, quotes, labels, or typography, specify the exact text clearly and wrap requested words in quotes. 5. **Avoid Over-Specification:** Do not invent highly specific clothing, colors, materials, or scene details unless the input supports them. 6. **Structure:** Write one cohesive paragraph after the thinking block. No bullets, JSON, or markdown. 7. **Respect Existing Detail:** If the user's prompt is already detailed, lightly polish and finalize rather than heavily expanding — preserve their phrasing and direction. 8. **Respect the Human Form:** Treat depictions of people with dignity. Assume clothing covers genitals and intimate anatomy. 9. **Preserve User Medium:** When the user explicitly requests a medium (e.g. "photo of", "photograph of", "illustration of", "painting of", "sketch of", "3D render of"), honor it. Do not pivot to a different medium to avoid difficulty — match the user's stated intent. User's Input: A masterfully rendered portrait captures Getou Suguru from Juvenile Kaisen, standing in a dynamic cowboy shot with wide sleeves and long black hair pulled back into a single hair bun adorned with subtle piercings. His upper body is illuminated by sunlight filtering through a blue sky above, casting soft shadows across his face as he gazes directly at the viewer, eyes intense and black framed within sharp focus. A hand rests gently on his chin, stroking thoughtfully while his other hand rises slightly in quiet contemplation, conveying emotional depth without words. He wears a traditional black kimono layered over Japanese-inspired clothing, its texture rich with fine silk weave and subtle folds catching the light. Surrounding him is an ethereal atmosphere: rainbows shimmer across clouds, tiny sparkles drift like dust motes in the sunbeams, and pastel hues blend seamlessly into warm amber tones that enhance his figure against a soft backdrop of misty sky. The composition employs a slight Dutch angle for visual tension, with foreshortening drawing attention to his expressive face and delicate features. Shot at ultra-high resolution with 32K UHD clarity, every detail-from the texture of fabric to the subtle glint in his eyes-is rendered with breathtaking precision, evoking an emotional impact through masterful craftsmanship. The scene balances fashion photography elegance with a sense of quiet intensity, creating a new high-quality visual masterpiece that feels both timeless and strikingly fresh, capturing the essence of youth, strength, and inner calm in a moment suspended between worlds.
Parameters used to generate this content
The node graph used to generate this content — pan and zoom to explore, or download it to run in ComfyUI.
AI models used to generate this content