What problem does it solve?
This Skill helps AI agents reliably generate, edit, composite, and review still images with OpenAI GPT Image models while avoiding deprecated integrations, invalid parameters, unsafe likeness workflows, and incomplete deliverables.
Core Features & Use Cases
- Model and API Selection: Choose GPT Image 2, distinguish the direct Image API from conversational Responses workflows, and migrate legacy GPT Image and DALL-E integrations.
- Production Image Workflows: Create detailed prompts, generate variants, edit and composite reference images, use masks and multiple inputs, preserve visual invariants, and handle exact in-image text.
- Output and Reliability Controls: Decode and save base64 image output, process partial-image streams, handle moderation and transient failures, estimate cost and rate limits, and record request metadata.
- Safety and Release Review: Apply consent, rights, privacy, provenance, authenticity, brand, and final-asset QA checks for customer-facing media.
- Use Case: Create a branded product poster, replace an object in a photographed scene with a masked reference edit, or run a conversational image workflow with a controlled follow-up change.
Quick Start
Use the openai-gpt-image skill to plan, generate, or edit a still image with GPT Image 2 and return the saved asset, metadata, safety checks, and final QA findings.