What problem does it solve?
This Skill removes the friction of switching between disparate image generation tools and handling inconsistent parameter requirements across different AI models, enabling seamless creation of custom visual assets from natural language prompts.
Core Features & Use Cases
- Multi-model support: Unified interface for 4 leading image generation backends (OpenAI gpt-image-2, Google Gemini 3.1 flash, TensorsLab Seedream v5, Alibaba Wan2.7-image-pro) with aligned core parameters.
- Flexible customization: Control aspect ratio, output quality, resolution, file format, and number of generated images to match specific use case needs.
- Robust error handling: Automatic retry for transient network or API errors, clear error messages for content moderation or invalid API keys, and context-aware output directory selection to keep generated assets organized.
- Use case: When preparing a research report, generate a custom cover image directly saved to the report's asset folder without manual file management or switching between different image generation platforms.
Quick Start
Use the create_image skill to generate a 16:9 high-quality cyberpunk Tokyo street scene image and save it to the workspace/reports/cyberpunk-tokyo/ directory.