generate-image

Generate and edit images with text prompts and configurable AI models.

270|16|Updated Feb 3, 2026
One-click install
npx skills add https://github.com/gupsammy/Claudest --skill generate-image-gupsammy
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: generate-image
Source: https://github.com/gupsammy/Claudest/tree/main/plugins/claude-content/skills/generate-image
Command: npx skills add https://github.com/gupsammy/Claudest --skill generate-image-gupsammy

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, pillow, and includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill empowers users to generate novel images from text descriptions and edit existing images with natural language commands, streamlining visual content creation.

Core Features & Use Cases

  • Text-to-Image (t2i): Generate original images based on detailed prompts.
  • Image-to-Image (i2i): Edit existing images by describing desired changes.
  • Multi-Reference Composition: Combine elements from multiple images into a new composition.
  • Use Case: Create a unique logo for a new brand, generate concept art for a game, or modify a product photo to showcase different variations.

Quick Start

Use the generate-image skill to create a picture of a cat wearing a party hat.

Frequently Asked Questions about generate-image

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text descriptions using AI?

To generate images from text, you provide a detailed prompt describing your desired visual content. The Skill uses advanced AI models to perform text-to-image generation, creating photorealistic or stylized art based on your specific instructions.

Can I edit existing images by describing changes in natural language?

Yes, you can edit existing images by describing desired changes in natural language. The image-to-image feature processes your commands to modify visual content, allowing you to showcase different variations or update product photos without manual editing.

Do I need a specific API key and package manager to run text-to-image generation?

Yes, you need a GEMINI_API_KEY and the uv package manager to execute the text-to-image generation scripts. These dependencies are required to authenticate requests and run the underlying AI models for processing your visual content.

What is multi-reference composition for AI image generation?

Multi-reference composition combines elements from multiple existing images into a new composition. This technique allows you to merge distinct visual components, such as combining a specific product with a different background, into a cohesive generated image.

Does this Skill support configurable aspect ratios and resolutions for AI art?

Yes, the Skill supports configurable aspect ratios and resolutions for AI art generation. You can specify these parameters alongside choosing between different models, such as Nano Banana or Nano Banana Pro, to achieve precise text rendering and photorealism.

What are the limitations of using natural language commands for image editing?

Natural language image editing relies heavily on prompt clarity and model interpretation. Complex or highly abstract instructions might not translate perfectly into the desired visual changes, requiring iterative prompting to achieve precise text rendering or specific photorealistic details.