baoyu-danger-gemini-web

Generate text and images via the Gemini Web API with vision input.

13|1|Updated Mar 4, 2026
One-click install
npx skills add https://github.com/ideacco/baoyu-skills-openclaw --skill baoyu-danger-gemini-web-ideacco
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: baoyu-danger-gemini-web
Source: https://github.com/ideacco/baoyu-skills-openclaw/tree/main/skills/baoyu-danger-gemini-web
Command: npx skills add https://github.com/ideacco/baoyu-skills-openclaw --skill baoyu-danger-gemini-web-ideacco

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill provides a powerful AI backend for generating text and images, acting as a versatile tool for creative and informational content creation.

Core Features & Use Cases

  • Text Generation: Create written content from prompts, suitable for articles, summaries, or creative writing.
  • Image Generation: Produce images based on textual descriptions, ideal for illustrations, social media posts, or concept art.
  • Vision Input: Analyze and understand content from reference images to inform text or image generation.
  • Multi-turn Conversation: Engage in back-and-forth dialogue for iterative content refinement.
  • Use Case: Generate a blog post about sustainable living and then create accompanying featured images for the post.

Quick Start

Use the baoyu-danger-gemini-web skill to generate an image of a futuristic city at sunset.

Frequently Asked Questions about baoyu-danger-gemini-web

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using Gemini Web API?

Multi-turn conversation allows iterative content refinement by engaging in back-and-forth dialogue. You can generate an initial text or image response and then provide follow-up prompts to update and improve the output.

Can I use reference images for vision input with AI generation?

Yes, you can use reference images for vision input. The Skill analyzes and understands content from provided images to inform subsequent text or image generation based on that visual context.

Do I need Node.js and Bun to run the Gemini Web Skill?

Yes, you need Node.js, Bun, and Chrome installed. These dependencies are strictly required for execution and authentication to access the reverse-engineered Gemini Web API functionalities.

What are the limitations of using a reverse-engineered Gemini Web API for image generation?

The main limitation of using a reverse-engineered Gemini Web API is potential instability, as it lacks official support and relies on Chrome for authentication, which may break with interface updates.

How does multi-turn conversation work for iterative content refinement?

Multi-turn conversation allows iterative content refinement by engaging in back-and-forth dialogue. You can generate an initial text or image response and then provide follow-up prompts to update and improve the output.