gemini-imagegen

Generate and edit images via the Gemini API with customizable resolutions and aspect ratios.

1|Updated Feb 23, 2026
One-click install
npx skills add https://github.com/hackefeller/ghostwire --skill gemini-imagegen-hackefeller
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: gemini-imagegen
Source: https://github.com/hackefeller/ghostwire/tree/main/.github/skills/gemini-imagegen
Command: npx skills add https://github.com/hackefeller/ghostwire --skill gemini-imagegen-hackefeller

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill automates the creation and modification of images using the Gemini API, streamlining visual content generation for various applications.

Core Features & Use Cases

  • Text-to-Image Generation: Create images from detailed text prompts.
  • Image Editing & Manipulation: Modify existing images, apply style transfers, and compose elements from multiple sources.
  • Use Case: Generate a logo for a new startup, create product mockups for an e-commerce site, or design custom stickers for social media.

Quick Start

Use the gemini-imagegen skill to create an image of a futuristic city at sunset.

Frequently Asked Questions about gemini-imagegen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using the Gemini API?

Text-to-image generation with the Gemini API requires submitting detailed text prompts to the skill. It uses Nano Banana Pro to create custom visual content, supporting resolutions from 1K to 4K and various aspect ratios.

Can I edit existing images and compose elements from multiple reference images?

Image editing and composition with the Gemini API allow you to modify existing images and combine elements from multiple reference images. The skill supports multi-turn refinement to iteratively adjust visual content.

Do I need an API key to use Gemini for AI art generation?

You need to set the GEMINI_API_KEY environment variable to use Gemini for AI art generation. This key authenticates your requests to the Gemini API for text-to-image creation and image editing.

What image resolutions and aspect ratios does text-to-image generation support?

Text-to-image generation supports customizable resolutions of 1K, 2K, and 4K. You can generate visual content with aspect ratios including 1:1, 2:3, and 3:2 to fit various design layouts.

What's the best way to create product mockups and logos using AI image generation?

AI image generation for product mockups and logos is best handled by providing detailed text descriptions to the skill. It uses the Gemini API to generate tailored visual content for e-commerce sites and startups.

Are there limitations when doing multi-turn refinement for image editing?

Multi-turn refinement for image editing relies on the Gemini API's context handling. While it allows iterative adjustments to visual content, complex multi-reference compositions may require precise prompting to achieve desired results.