gpt-image-2

Generate and edit images with embedded text using OpenAI GPT Image 2 via RunComfy CLI.

12|2|Updated May 18, 2026
One-click install
npx skills add https://github.com/runcomfy-com/skills --skill gpt-image-2-runcomfy-com
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: gpt-image-2
Source: https://github.com/runcomfy-com/skills/tree/main/gpt-image-2
Command: npx skills add https://github.com/runcomfy-com/skills --skill gpt-image-2-runcomfy-com

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

It helps users generate and edit high-quality images with reliable prompt-following, especially when the output must include accurate embedded text, logos, and brand-safe layout details.

Core Features & Use Cases

  • Text- and logo-accurate image generation: Produces signage, packaging mockups, UI concepts, and multilingual typography with strong instruction precision.
  • Edit with reference images: Uses up to 10 provided image URLs to preserve the intended subject characteristics while changing backgrounds, layout, and text placement.
  • Model routing guidance for sibling models: Recommends switching to Flux 2 for heavy stylization, Nano Banana Pro for photorealistic portraits, and Seedream for cinematic/aesthetic-first shots.

Quick Start

Use the gpt-image-2 skill to generate an image by asking for your exact subject, setting, and any embedded text you want included.

Frequently Asked Questions about gpt-image-2

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images with accurate embedded text and logos?

To generate images with accurate embedded text and logos, you use text-to-image generation by providing a precise prompt that includes your exact subject, setting, and the specific text or typography you require. The system processes these instructions to produce brand-safe layouts with reliable prompt-following for signage and packaging mockups.

Can I edit existing images while preserving the original subject?

Yes, you can edit existing images while preserving the original subject by providing up to 10 public HTTPS reference image URLs. The system uses these references to maintain intended subject characteristics while allowing you to change backgrounds, layouts, and text placement.

What is the best way to create multilingual typography for UI concepts?

The best way to create multilingual typography for UI concepts is through text-accurate image generation. You provide exact text strings and layout instructions in your prompt, and the model outputs high-quality visual mockups with strong instruction precision for the required typography.

Does this image generation approach work for heavy stylization or photorealistic portraits?

This approach is optimized for text-accurate visual requirements and brand-safe layouts rather than heavy stylization or photorealistic portraits. For heavy stylization, you should route to Flux 2; for photorealistic portraits, use Nano Banana Pro; and for cinematic, aesthetic-first shots, use Seedream.

How do I use reference images for ad creatives and UI mockups?

To use reference images for ad creatives and UI mockups, you invoke the CLI with the correct model endpoint and provide your prompt alongside up to 10 public HTTPS reference image URLs. You can also specify optional size constraints to meet your creative formatting requirements.