gemini-omni-flash-api

Generate and edit videos with the Gemini Omni Flash model via the google-genai SDK.

3.9k|396|Updated Feb 6, 2026
One-click install
npx skills add https://github.com/google-gemini/gemini-skills --skill gemini-omni-flash-api
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: gemini-omni-flash-api
Source: https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-omni-flash-api
Command: npx skills add https://github.com/google-gemini/gemini-skills --skill gemini-omni-flash-api

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires google-genai, ffmpeg, and includes scripts (resource) and assets (resource) components.

What problem does it solve?

This Skill addresses the challenge of manually creating video content by offering advanced text-to-video generation and editing, allowing users to create professional-quality videos with ease.

Core Features & Use Cases

  • Text to Video: Generate videos from text prompts, perfect for storytelling, presentations, and explainer videos.
  • Image to Video: Create videos from single images or image sets, ideal for animations or quick content generation.
  • Video Editing: Edit existing videos with style changes, inpainting/outpainting, and audio regeneration.
  • Use Case: Need a promotional video for a product launch? Use this Skill to generate a video from a single image, with text descriptions, custom styles, and audio generated from scratch.

Quick Start

Generate a promotional video for a new product using the gemini-omni-flash-api skill. Provide a text prompt, an image of the product, and set the desired aspect ratio.

Frequently Asked Questions about gemini-omni-flash-api

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate video from text prompts for professional video production?

Generating video from text prompts uses the Gemini Omni Flash model and google-genai SDK to automate video production. This text-to-video approach creates professional content for storytelling, presentations, and explainer videos without manual editing.

Can I create animations from a single image using an AI video generator?

Creating animations from a single image uses the image-to-video feature via the google-genai SDK. This approach generates dynamic video content from static images, ideal for quick content generation and product promotional videos.

Does video editing with the Gemini API support style changes and audio regeneration?

Video editing with the Gemini API supports style changes, inpainting, outpainting, and audio regeneration. These features allow you to modify existing videos and generate custom audio directly through the google-genai SDK.

Do I need Python and ffmpeg to run text-to-video generation?

Running text-to-video generation requires Python, ffmpeg, and the google-genai SDK for operation. You must configure these dependencies locally to successfully execute the Gemini Omni Flash model and generate video content.

What is the best way to generate a promotional video from a product image?

Generating a promotional video is best achieved using the image-to-video feature with the Gemini Omni Flash model. Provide a product image, text descriptions, and set the desired aspect ratio to automatically generate video with custom styles and audio.