nano-banana-pro

Generate and edit images via the Gemini 3 Pro Image API.

1|Updated Feb 1, 2026
One-click install
npx skills add https://github.com/zkcpku/verdentClaw --skill nano-banana-pro-zkcpku
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nano-banana-pro
Source: https://github.com/zkcpku/verdentClaw/tree/main/skills/nano-banana-pro
Command: npx skills add https://github.com/zkcpku/verdentClaw --skill nano-banana-pro-zkcpku

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, pillow, and includes scripts (resource) components.

What problem does it solve?

This Skill allows users to easily generate new images from text descriptions or edit existing images using advanced AI models, streamlining creative workflows.

Core Features & Use Cases

  • Image Generation: Create unique images based on detailed text prompts.
  • Image Editing: Modify existing images with specific instructions.
  • Multi-Image Composition: Combine multiple images into a single scene.
  • Use Case: Generate a photorealistic image of a "cyberpunk cat wearing sunglasses" or edit a landscape photo to "add a starry night sky."

Quick Start

Use the nano-banana-pro skill to generate an image of a futuristic city at sunset and save it as 'cityscape.png'.

Frequently Asked Questions about nano-banana-pro

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text descriptions using the Gemini API?

To generate images from text, you provide a detailed text prompt to the image generation API, which then creates a unique image based on your description. This streamlines creative workflows by automating visual content creation.

Can I edit existing images and combine multiple pictures into one scene?

Yes, you can edit existing images with specific instructions and perform multi-image composition to combine multiple pictures into a single scene. This allows for complex image manipulation directly through text commands.

Do I need a GEMINI_API_KEY to use this image generation and editing tool?

Yes, you must configure the GEMINI_API_KEY environment variable to authenticate requests. This key is required to access the underlying Gemini 3 Pro Image API for both generation and editing tasks.

What image resolutions are supported for AI generation and editing?

The AI image generation and editing process supports 1K, 2K, and 4K resolutions. You can specify the desired output resolution when submitting your text-to-image prompts or image editing instructions.

How does multi-image composition work with AI image editing?

Multi-image composition works by taking multiple input images and merging them into a single cohesive scene based on your instructions. This feature leverages the Gemini API to seamlessly blend visual elements from different sources.

What is the best way to add specific elements like a starry sky to an existing photo?

The best way to add elements to an existing photo is through AI image editing, where you provide the original image and a text instruction like 'add a starry night sky'. The API processes this to modify the image accordingly.