What problem does it solve?
This Skill centralizes and automates AI image generation across multiple providers so users can produce high-quality images, edits, and reference-based modifications without manually handling different APIs and formats.
Core Features & Use Cases
- Multi-provider support: Works with Google Gemini, Google Imagen, OpenAI GPT Image, DashScope (阿里通义万象), and Replicate.
- Reference image edits & prompts: Accepts reference images for multimodal edits, supports aspect ratios, explicit sizes, and quality presets.
- Configurable preferences and model resolution: Loads EXTEND.md or environment variables, allows CLI overrides, and displays chosen provider/model before generation.
- Robust CLI workflow: Validates inputs, auto-detects providers, retries transient failures, supports sequential and optional parallel batch generation, and writes output PNG files.
Quick Start
Generate a 2k landscape image of a futuristic city using the Google provider and save it as out.png.