What problem does it solve? Creating AI-generated media requires navigating dozens of models with different parameters, pricing, and capabilities. This Skill unifies image, video, and audio generation through the fal.ai MCP server so you can produce media assets from a single interface without learning each model's API. ## Core Features & Use Cases - Image Generation: Create images with Nano Banana 2 for fast drafts or Nano Banana Pro for high-fidelity production output, including image editing and style transfer from uploaded sources. - Video Generation: Produce text-to-video and image-to-video clips with Seedance 1.0 Pro, Kling Video v3 Pro, or Veo 3, with control over duration, aspect ratio, and seed. - Audio Generation: Synthesize speech with CSM-1B, generate video-matched audio with ThinkSound, or use ElevenLabs and VideoDB for voice, music, and sound effects. - Use Case: A content creator needs a thumbnail, a 5-second intro clip, and a voiceover for a video. They generate the thumbnail with Nano Banana Pro, animate it with Seedance, and narrate it with CSM-1B, checking cost estimates before each run. ## Quick Start Ask the assistant to generate an image of a specific scene using fal.ai, for example a landscape product photo, and it will select the right model and parameters.