What problem does it solve?
This skill provides seamless access to Google's Gemini API for rapid text generation, image generation, multimodal analysis, function calling, and search grounding, enabling you to implement advanced AI capabilities without building from scratch.
Core Features & Use Cases
- Unified access to Gemini models: text generation, image creation, multimodal analysis, and function calling through REST endpoints.
- Rapid prototyping and automation: generate marketing copy, summarize content, create visuals, and ground results with live search in a single workflow.
- Use Case: Example: you need a product description and a hero image for a new feature; Gemini can produce both text and visuals in a coordinated response.
Quick Start
Use the Gemini API with your GOOGLE_API_KEY to generate content and visuals by selecting appropriate models (gemini-2.5-flash for text, gemini-2.5-flash-image for images) and, if needed, enable image generation and grounding.