gemini-imagegen

Generate and edit images using Google Gemini multimodal models via the API.

1.1k|99|Updated Feb 12, 2026
One-click install
npx skills add https://github.com/MooseGoose0701/skill-compose --skill gemini-imagegen-moosegoose0701
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: gemini-imagegen
Source: https://github.com/MooseGoose0701/skill-compose/tree/main/skills/gemini-imagegen
Command: npx skills add https://github.com/MooseGoose0701/skill-compose --skill gemini-imagegen-moosegoose0701

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, and includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill enables users to generate and edit images using Google Gemini's advanced AI models, simplifying the creation of visual content.

Core Features & Use Cases

  • Text-to-Image Generation: Create images from textual descriptions.
  • Image Editing: Modify existing images based on prompts.
  • Multi-Image Composition: Combine multiple images or maintain character consistency.
  • Use Case: A user wants to create a unique illustration for a blog post. They can describe the desired image, and the Skill will generate it using Gemini's capabilities.

Quick Start

Use the gemini-imagegen skill to generate an image of a cat wearing a hat.

Frequently Asked Questions about gemini-imagegen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate AI illustrations from text descriptions?

AI illustration generation from text descriptions uses Google Gemini models to convert written prompts into high-quality images. This Skill handles text-to-image creation to produce visual content for blogs or creative projects.

Can I edit existing photos using a text prompt with Gemini?

Yes, editing existing photos using text prompts with Gemini is supported. This Skill processes your reference images and applies modifications based on the specific textual instructions you provide to achieve the desired edits.

What is the best way to maintain character consistency across multiple AI images?

Maintaining character consistency across multiple AI images is achieved through multi-image composition. This Skill uses Gemini to process reference images alongside text prompts, ensuring visual coherence across generated outputs.

Do I need a GEMINI_API_KEY to generate images?

Yes, a GEMINI_API_KEY environment variable is required to generate images. This Skill needs this key to authenticate requests sent to the Google Gemini multimodal models for executing text-to-image and image editing tasks.

Does image generation with Google Gemini support multi-image composition?

Yes, image generation with Google Gemini supports multi-image composition. This feature allows you to combine multiple reference images and maintain character consistency within the generated AI illustrations.