What problem does it solve? Creating AI-generated media requires navigating dozens of models with different parameters, pricing, and capabilities. This Skill consolidates image, video, and audio generation into a single workflow via the fal.ai MCP server, so you can produce media without learning each model's API. ## Core Features & Use Cases - Image Generation: Text-to-image and image editing with Nano Banana 2 for fast drafts and Nano Banana Pro for high-fidelity production output. - Video Generation: Text-to-video and image-to-video using Seedance 1.0 Pro, Kling Video v3 Pro, and Veo 3, with duration and aspect ratio controls. - Audio Generation: Conversational text-to-speech with CSM-1B, video-to-audio with ThinkSound, plus ElevenLabs and VideoDB integration options. - Cost Control: Estimate generation costs before running expensive jobs and discover models via search. - Use Case: A content creator needs a thumbnail, a 5-second intro clip, and voiceover narration for a video. They generate the thumbnail with Nano Banana 2, animate it with Seedance, and produce the narration with CSM-1B, all from one conversation. ## Quick Start Configure the fal.ai MCP server with your FAL_KEY, then ask the AI to generate a landscape image of a sunset cityscape using Nano Banana 2.