What problem does it solve?
This Skill removes the friction of working with the Gemini API for image tasks, eliminating the need to write boilerplate code for API calls, handle response parsing, or troubleshoot common format issues like mismatched image file extensions.
Core Features & Use Cases
- Text-to-Image Generation: Create custom images from text prompts with control over resolution and aspect ratio, ideal for marketing assets, social media content, and concept art.
- Image Editing & Composition: Modify existing images or combine up to 14 reference images into new compositions, perfect for product mockups, photo retouching, and creative visual projects.
- Iterative Multi-Turn Refinement: Chat with the model to progressively tweak images through conversational feedback, great for logo design, visual prototyping, and fine-tuning creative outputs.
- Use Case: A small business owner can generate product mockups for their online store, edit product photos to match brand aesthetics, and refine designs through simple chat commands without hiring a professional designer.
Quick Start
Use the gemini-imagegen skill to generate a 16:9 widescreen product photo of a wireless coffee maker on a marble countertop with soft studio lighting and save it as product_shot.jpg.