nano-banana-pro

Generate and edit images with Gemini 3 Pro Image API.

5|Updated Jan 31, 2026
One-click install
npx skills add https://github.com/kcns008/clusterclaw --skill nano-banana-pro-kcns008
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nano-banana-pro
Source: https://github.com/kcns008/clusterclaw/tree/main/skills/nano-banana-pro
Command: npx skills add https://github.com/kcns008/clusterclaw --skill nano-banana-pro-kcns008

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, pillow, and includes scripts (resource) components.

What problem does it solve?

It removes the friction of creating or revising images by turning a natural-language prompt and optional reference images into a finished visual output.

Core Features & Use Cases

  • Image Generation: Create a brand-new image from a text prompt using Gemini 3 Pro Image.
  • Image Editing: Modify one image or combine multiple images into a single composed result.
  • Practical Workflow: Use it for concept art, marketing visuals, social graphics, and rapid visual iteration with automatic output saving and resolution handling.

Quick Start

Use the nano-banana-pro skill to create a 2K image of a futuristic neon city at sunset and save it as output.png.

Frequently Asked Questions about nano-banana-pro

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate and edit images using the Gemini API?

You can generate and edit images using the Gemini API by providing a natural-language text prompt and optional reference images. The skill handles text-to-image creation, single-image editing, and multi-image composition, outputting a finished PNG file.

Can I combine multiple images into one with Gemini 3 Pro Image?

Yes, you can combine multiple images into one composed result with Gemini 3 Pro Image. The skill supports multi-image composition workflows, allowing you to merge several input images into a single visual output.

What do I need to set up before generating images with this API tool?

Before generating images, you need a valid Gemini API key and Python image-processing dependencies installed. The environment requires the google-genai and pillow libraries to manage resolution selection and PNG output.

What is the best way to create marketing visuals from text prompts?

The best way to create marketing visuals from text prompts is using an AI image generation tool that supports natural-language input. This skill turns your descriptive text directly into high-resolution PNG assets for rapid visual iteration.

Does this image generation method automatically save the output file?

Yes, this image generation method automatically saves the output file. It manages resolution selection and error-checked API execution to ensure your generated or edited image is securely saved as a PNG file.