What problem does it solve? Creating visual and audio assets requires switching between multiple AI generation platforms with different APIs and parameters. This Skill unifies image, video, and audio generation through a single fal.ai MCP interface with ready-to-use model configurations. ## Core Features & Use Cases - Image Generation: Text-to-image and image editing with Nano Banana 2 for fast drafts and Nano Banana Pro for high-fidelity production output. - Video Generation: Text-to-video and image-to-video using Seedance 1.0 Pro, Kling Video v3 Pro, and Veo 3, with duration and aspect ratio controls. - Audio Generation: Text-to-speech with CSM-1B, video-to-audio with ThinkSound, plus ElevenLabs and VideoDB integration options. - Use Case: A content creator needs a thumbnail, a short promo clip, and a voiceover for a video. They generate the thumbnail with Nano Banana Pro, animate it into a 5-second clip with Seedance, and produce narration with CSM-1B, all through one MCP server. ## Quick Start Configure the fal.ai MCP server with your FAL_KEY, then ask the assistant to generate an image of your chosen subject using the fal-ai-media skill.