gemini-imagegen

Generate and edit images via the Gemini API with custom resolutions.

Updated Feb 1, 2026
One-click install
npx skills add https://github.com/isndotbiz/website --skill gemini-imagegen-isndotbiz
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: gemini-imagegen
Source: https://github.com/isndotbiz/website/tree/main/.claude/skills/gemini-imagegen
Command: npx skills add https://github.com/isndotbiz/website --skill gemini-imagegen-isndotbiz

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, Pillow, and includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill streamlines the creation and modification of visual content, enabling users to generate unique images from text prompts and refine existing ones with ease.

Core Features & Use Cases

  • Text-to-Image Generation: Create diverse images from detailed text descriptions.
  • Image Editing & Refinement: Modify existing images by adding elements, changing styles, or making specific edits.
  • Multi-Turn Iteration: Engage in conversational refinement for iterative image development.
  • Use Case: Generate a photorealistic product mockup for a new gadget, then ask the Skill to add a specific background and adjust the lighting for a more professional look.

Quick Start

Use the gemini-imagegen skill to create a logo for 'Acme Corp' with a blue gradient.

Frequently Asked Questions about gemini-imagegen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using the Gemini API?

Image editing with the Gemini API allows you to modify existing images by adding elements, changing styles, or making specific adjustments. You can achieve iterative image development through conversational multi-turn refinement.

Can I create product mockups and logos using text-to-image generation?

Yes, you can specify custom resolutions and aspect ratios when generating images. This allows you to tailor the text-to-image output dimensions to fit specific design requirements for logos, mockups, or stickers.

Do I need a GEMINI_API_KEY to use this image generation Skill?

This Skill depends on the google-genai library and Pillow for image processing. These dependencies enable the application to handle image generation, apply edits, and process the resulting visual content efficiently.

What are the limitations of using Gemini API for AI art and image editing?

Limitations of using the Gemini API for AI art include the need for a valid GEMINI_API_KEY and reliance on external dependencies like Pillow. Complex multi-turn refinement may require iterative prompting to achieve the exact desired visual result.