What problem does it solve?
Raw user requests for image generation are often vague or unstructured, leading to unstable results from image models. This Skill rewrites those requests into complete, model-ready visual prompts that lock down what must stay unchanged while filling in composition, lighting, material, and style details.
Core Features & Use Cases
- Three Generation Modes: Handles text-to-image, reference-guided, and image-to-image workflows, with mode-specific rules for how much creative freedom the prompt allows.
- Multi-Reference Role Assignment: When multiple reference images exist, the final prompt explicitly assigns each one a role, such as locking subject identity, supplying environment, or defining materials.
- Text and Layout Safety: Prevents the model from rendering prompt text, layout labels, watermarks, or gibberish into the image unless the user explicitly asks for visible text.
- Use Case: A content creator asks for a Xiaohongshu-style cover image from a product photo. The Skill produces a final prompt that preserves the product's identity and outline, adds commercial lighting and a clean background, and reserves a clear title area.
Quick Start
Before generating an image, ask the assistant to optimize your request into a final image prompt, for example: turn my idea of a cozy coffee shop product shot into a ready-to-use generation prompt.