What problem does it solve?
Gemini image generation tasks require precise prompts, reference-based edits, and careful watermark handling to achieve consistent visuals across iterations.
This skill provides a structured approach to generating and editing images with Gemini models while preserving character identity, pose, and composition, and guiding watermark removal through prompt instructions.
Core Features & Use Cases
- Prompt-driven text-to-image and reference-based image-to-image generation using Gemini models.
- Maintain character and pose consistency across multiple outputs by reusing references from previous results.
- Instruct watermark removal via prompts rather than patching source images, ensuring cleaner finals.
- Flexible deployment workflow with automatic API key resolution (Vertex AI or Gemini) and safe prompting guidance.
Quick Start
Describe the visual you want, specify the target Gemini model, and generate images that preserve identity and remove watermarks.