nano-banana-2

Generate and edit images from text prompts using gemini-3.1-flash-image-preview.

Updated Aug 27, 2026
One-click install
npx skills add https://github.com/pr3t3l/openclaw-config --skill nano-banana-2-pr3t3l
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: nano-banana-2
Source: https://github.com/pr3t3l/openclaw-config/tree/main/workspace/skills/nano-banana-2-gemini
Command: npx skills add https://github.com/pr3t3l/openclaw-config --skill nano-banana-2-pr3t3l

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Gemini image generation, editing, and search-grounded image creation via gemini-3.1-flash-image-preview (Nano Banana 2). Automates creating visuals from text prompts, transforming existing images with natural language instructions, and grounding results with live references. All outputs are saved under .nano-banana/ in the project directory.

Core Features & Use Cases

  • Create from prompts: Generate new visuals from descriptive prompts using Gemini.
  • Edit existing images: Apply transformations to local images with text instructions.
  • Grounded results: Produce images informed by current web and image references.

Quick Start

Provide a text prompt to generate an image.

Frequently Asked Questions about nano-banana-2

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using Gemini?

Gemini image generation creates visuals from descriptive text prompts by submitting your input to the gemini-3.1-flash-image-preview model and saving the resulting output files directly in your local .nano-banana/ directory.

Can I edit existing local images with natural language instructions?

Editing existing images applies natural language transformations to local image files, processing your requested modifications through the Gemini model and saving the updated visual assets into the project's local .nano-banana/ directory.

How does grounded image generation work with live search data?

Grounded image generation produces visuals informed by current web and image references, using live search data to ensure the generated output reflects up-to-date styling and factual context for your creative workflows.

Do I need a Gemini API key to use this image generation model?

A GEMINI_API_KEY is required to authenticate requests and access the gemini-3.1-flash-image-preview model for text-to-image generation, image editing, and search-grounded visual creation tasks within your project directory.

What is the best way to organize generated Gemini image assets locally?

Organizing generated image assets is handled automatically by storing all visual outputs under the .nano-banana/ folder in your project directory, using local directory organization for straightforward asset management and iterative refinement.

Why does grounded image generation require live search references?

Grounded image generation requires live search references to produce images informed by current web data, ensuring that creative workflows requiring rapid ideation and up-to-date reference styling maintain factual accuracy across diverse subjects.