lipsync

Generates lip-synced videos from audio tracks and source media via the RunComfy CLI.

12|2|Updated Aug 12, 2026
One-click install
npx skills add https://github.com/genmedia-labs/skills --skill lipsync-genmedia-labs
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: lipsync
Source: https://github.com/genmedia-labs/skills/tree/main/lipsync
Command: npx skills add https://github.com/genmedia-labs/skills --skill lipsync-genmedia-labs

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires @runcomfy/cli.

What problem does it solve? Matching a face's mouth movements to a separate audio track requires choosing between many AI lip-sync models with different input requirements. This Skill routes your request to the right RunComfy endpoint — Sync Labs, OmniHuman, Kling, or Creatify — based on whether you have a source video, a portrait still, or only a script. ## Core Features & Use Cases - Mouth-swap on existing video: Sync Labs sync v2 / Pro applies state-of-the-art lip-sync onto existing footage while preserving everything else in the frame. - Avatar from a portrait: ByteDance OmniHuman turns a single portrait photo plus an audio file into a talking-head video. - Script-to-synced-video: Kling lipsync text-to-video and HappyHorse generate speech audio in-pass when no audio file exists. - Use Case: Dub a brand video into five languages by running Sync Labs sync v2 Pro with the original video and each translated voiceover MP3. ## Quick Start Ask the agent to lip-sync your source video to a voiceover file, for example: sync the lips in product-demo.mp4 to the audio in voiceover-es.mp3 using RunComfy.

Frequently Asked Questions about lipsync

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I lip-sync a video to an audio file with AI?▼

Provide a source video URL and an audio URL, then run Sync Labs sync v2 via the runcomfy CLI with both URLs in the JSON input. The model replaces only the mouth region while preserving the original camera, lighting, and body motion.

What is the difference between Sync Labs and OmniHuman for lip sync?▼

Sync Labs sync v2 applies mouth-sync onto an existing video, preserving the original footage. OmniHuman generates a new talking-head video from a single portrait photo plus audio, making it the default for avatar-style content rather than editing existing footage.

Can I lip-sync a video without a pre-recorded audio file?▼

Yes. Kling lipsync text-to-video and HappyHorse 1.0 generate speech audio in-pass from a written script and sync it to the resulting video. The tradeoff is that the audio is regenerated each call, so you cannot lock the mouth to a specific voiceover.

Why does my lip-sync output drift or look misaligned?▼

Drift usually comes from a significant duration mismatch between the audio and video. Trim the audio or extend the video so lengths match, and use clean voiceover audio without a music bed, since audio quality directly drives mouth quality.

Is it allowed to lip-sync a real person's face to someone else's voice?▼

Only with consent from both parties. The skill instructs agents to refuse requests targeting real public figures without consent or aiming at defamatory or sexually explicit synthetic media, and operators must hold rights to both the face and the voice.