baoyu-image-gen

Generate images from text prompts via OpenAI, Gemini, DashScope, and Replicate.

10|2|Updated Mar 3, 2026
One-click install
npx skills add https://github.com/wx-chevalier/Awesome-Agent-Skills --skill baoyu-image-gen-wx-chevalier
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: baoyu-image-gen
Source: https://github.com/wx-chevalier/Awesome-Agent-Skills/tree/main/baoyu/baoyu-image-gen
Command: npx skills add https://github.com/wx-chevalier/Awesome-Agent-Skills --skill baoyu-image-gen-wx-chevalier

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill automates the creation of images from textual descriptions, enabling users to visualize concepts, generate artwork, and create visual assets without needing advanced design skills.

Core Features & Use Cases

  • Multi-Provider Support: Integrates with OpenAI, Google (Gemini), DashScope, and Replicate APIs for flexible image generation.
  • Customization Options: Supports various aspect ratios, sizes, and quality settings.
  • Reference Image Input: Allows using existing images as a basis for generation (e.g., for style transfer or image editing).
  • Use Case: A marketer needs a unique banner image for a social media campaign. They provide a prompt like "A futuristic cityscape at sunset, vibrant colors, digital art style" and the Skill generates several options.

Quick Start

Use the baoyu-image-gen skill to generate an image of a cat wearing a hat.

Frequently Asked Questions about baoyu-image-gen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts across different AI providers?

Text-to-image generation is enabled across OpenAI, Google Gemini, DashScope, and Replicate APIs. You provide a text prompt describing the desired visual, and the skill outputs AI-generated images configurable by aspect ratio, size, and quality.

Can I use a reference image for AI art generation and editing?

Reference image input is supported for AI art generation. You can supply an existing image as a basis for the process, enabling use cases like style transfer and image editing without needing advanced manual design skills.

What text-to-image parameters can I configure for visual content creation?

Configurable parameters for visual content creation include aspect ratio, size, and quality settings. These options allow you to tailor the generated output for specific marketing, design, or artistic endeavor requirements.

Does this AI image generation skill work with OpenAI and Google Gemini?

The skill works with both OpenAI and Google Gemini, along with DashScope and Replicate. This multi-provider support integrates various APIs to facilitate flexible text-to-image generation directly from textual descriptions.

What is the best way to create marketing banner images using AI?

The best way to create marketing banners is using text-to-image generation. You provide a descriptive prompt like a futuristic cityscape at sunset, and the skill generates multiple visual asset options matching the requested digital art style.