video-gen

Generate videos from images and text prompts using Higgsfield DOP and fal.ai Kling 3.0.

267|63|Updated Feb 15, 2026
One-click install
npx skills add https://github.com/modu-ai/cowork-plugins --skill video-gen-modu-ai
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: video-gen
Source: https://github.com/modu-ai/cowork-plugins/tree/main/moai-media/skills/video-gen
Command: npx skills add https://github.com/modu-ai/cowork-plugins --skill video-gen-modu-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill simplifies the production of high-quality videos by combining image and text inputs with advanced AI models, reducing the need for complex video editing skills.

Core Features & Use Cases

  • Image-to-Video Conversion: Transform static images into cinematic motion videos with customizable motion presets.
  • Text-to-Video Generation: Generate dynamic videos from textual prompts, suitable for marketing or storytelling.
  • Lip-Sync and Character Scenes: Create videos with animated characters speaking or lip-syncing based on images and prompts.
  • Use Case: For marketing, a user can turn a product image into an engaging promotional video using preset camera movements like orbit or dolly.

Quick Start

Provide an image URL and prompt describing the desired video motion to generate a cinematic video from your image.

Frequently Asked Questions about video-gen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
Can I create text-to-video animations for marketing content?▼

You can generate a dynamic video from a text prompt using fal.ai Kling 3.0. This text-to-video generation feature creates dynamic sequences suitable for marketing campaigns or visual storytelling.

Does lip-sync video generation work with fal.ai Kling 3.0?▼

Creating dynamic motion videos from images requires an image URL and a text prompt. The prompt should describe the desired video motion, which the AI then uses to automate the cinematic transformation.

Does lip-sync video generation work with fal.ai Kling 3.0?▼

Yes, lip-sync and character scene video generation is handled by fal.ai Kling 3.0. This model supports text-to-video applications and animates characters speaking based on your provided images and prompts.

What inputs are required to create dynamic motion videos from images?▼

Creating dynamic motion videos requires an image URL and a text prompt describing the desired motion. The AI uses these inputs to automate the cinematic video transformation without manual editing.