nano-banana-pro

Generate and edit images using the Gemini 3 Pro Image API.

5.1k|1.2k|Updated Feb 23, 2026
One-click install
npx skills add https://github.com/linuxhsj/openclaw-zero-token --skill nano-banana-pro-linuxhsj
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nano-banana-pro
Source: https://github.com/linuxhsj/openclaw-zero-token/tree/main/skills/nano-banana-pro
Command: npx skills add https://github.com/linuxhsj/openclaw-zero-token --skill nano-banana-pro-linuxhsj

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, pillow, and includes scripts (resource) components.

What problem does it solve?

This Skill allows users to generate new images from text prompts or edit existing images using natural language instructions, eliminating the need for complex graphic design software.

Core Features & Use Cases

  • Image Generation: Create unique images based on detailed text descriptions.
  • Image Editing: Modify existing images by providing instructions like "add a hat" or "change the background".
  • Multi-Image Composition: Combine multiple input images into a single scene.
  • Use Case: Generate a realistic image of a "cyberpunk cat wearing sunglasses in a neon-lit alley" or edit a photo to "make the sky look like a sunset."

Quick Start

Use the nano-banana-pro skill to generate an image of a 'fluffy white cat sitting on a windowsill' and save it as 'cat.png'.

Frequently Asked Questions about nano-banana-pro

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I edit existing images using natural language instructions?

To edit existing images using natural language, you provide descriptive instructions like "add a hat" or "change the background" to the Gemini 3 Pro Image API, which modifies the visuals without requiring complex graphic design software.

What is multi-image composition in AI image generation?

Multi-image composition in AI image generation is the process of combining multiple input images into a single unified scene using descriptive instructions, enabled by the Gemini 3 Pro Image API.

Does the Gemini 3 Pro Image API support different aspect ratios and resolutions?

Yes, image generation with the Gemini 3 Pro Image API supports various resolutions and aspect ratios, allowing you to create new visuals from text prompts tailored to specific dimensional requirements.

Do I need the Google GenAI library and Pillow to generate images from text?

Yes, you need the Google GenAI library and Pillow to generate images from text, as these dependencies are required for interacting with the API and handling the subsequent image manipulation.

What is the best way to generate AI art without complex graphic design software?

The best way to generate AI art without complex graphic design software is using text prompts with the Gemini 3 Pro Image API, which creates unique visuals based on detailed text descriptions.