nano-banana-pro

Generate and edit images using the Gemini 3 Pro Image model.

16|1|Updated Feb 23, 2026
One-click install
npx skills add https://github.com/Flexasaurusrex/OpenPaw --skill nano-banana-pro-flexasaurusrex
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nano-banana-pro
Source: https://github.com/Flexasaurusrex/OpenPaw/tree/main/skills/nano-banana-pro
Command: npx skills add https://github.com/Flexasaurusrex/OpenPaw --skill nano-banana-pro-flexasaurusrex

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, pillow, and includes scripts (resource) components.

What problem does it solve?

This Skill allows users to generate new images from text descriptions or edit existing images using AI, streamlining creative workflows and content creation.

Core Features & Use Cases

  • Image Generation: Create unique images based on detailed text prompts.
  • Image Editing: Modify existing images with specific instructions (e.g., change style, add elements).
  • Multi-Image Composition: Combine multiple images into a single scene.
  • Use Case: A marketer needs a banner image for a new product launch. They can use this Skill to generate several options based on a description and then select and refine the best one.

Quick Start

Generate an image of a futuristic cityscape at sunset with the filename 'cityscape.png'.

Frequently Asked Questions about nano-banana-pro

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using Gemini?

To generate images from text prompts, this Skill uses the Gemini 3 Pro Image model to create unique visuals based on your detailed descriptions. You simply provide the text prompt and specify an output filename like 'cityscape.png' to save the result.

Can I edit existing images with AI using a text instruction?

Yes, you can edit existing images with AI by providing specific text instructions. The Skill uses the Gemini 3 Pro Image model to modify elements, change styles, or add new components to your input image based on your prompt.

Do I need an API key to use the Gemini image generation model?

Yes, you need a GEMINI_API_KEY to authenticate and use the Gemini image generation model. You must also have the 'uv' binary installed on your system to execute the Python scripts that run the generation and editing processes.

What is the maximum number of images I can combine for multi-image composition?

For multi-image composition, you can combine up to 14 input images into a single scene. The Skill processes these inputs through the Gemini 3 Pro Image model to merge and blend multiple visual elements seamlessly.

How does multi-image composition work for content creation?

Multi-image composition works by taking up to 14 separate input images and merging them into a single cohesive scene. This streamlines content creation workflows by allowing you to generate complex visuals from multiple source images using AI.