long-video-agent

Automate multi-scene planning, parallel segment generation, and timeline assembly for narrative videos.

12|3|Updated Jun 17, 2026
One-click install
npx skills add https://github.com/phuhao00/bony-agent --skill long-video-agent
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: long-video-agent
Source: https://github.com/phuhao00/bony-agent/tree/main/.agent/skills/long-video-agent
Command: npx skills add https://github.com/phuhao00/bony-agent --skill long-video-agent

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill solves the complexity of producing long-form, multi-scene narrative videos by automating the orchestration of script-to-video workflows, ensuring narrative consistency and rhythmic pacing.

Core Features & Use Cases

  • Multi-Scene Planning: Automatically decomposes scripts into structured shot lists with specific visual descriptions and timing.
  • Parallel Generation: Utilizes advanced models like Tongyi Wan to generate video segments concurrently, significantly reducing production time.
  • Seamless Assembly: Handles the technical stitching of video segments, transitions, and timing to create a cohesive final product.
  • Use Case: A content creator can provide a 2-minute short-story script, and the agent will generate the storyboard, produce individual scenes, and assemble the final video with appropriate pacing.

Quick Start

Use the long-video-agent to generate a 60-second cinematic video based on the provided script about a futuristic city.

Frequently Asked Questions about long-video-agent

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate long-form narrative video production from a script?▼

To automate long-form narrative video production, this Skill decomposes scripts into structured shot lists, generates segments in parallel, and stitches scenes into a cohesive final video with consistent pacing.

What is multi-scene planning for video generation and how does it work?▼

Multi-scene planning for video generation is the process of automatically decomposing a script into a structured storyboard with specific visual descriptions and timing, ensuring narrative consistency before parallel segment synthesis begins.

Can I use Tongyi Wan models for parallel video segment generation?▼

Yes, you can use Tongyi Wan models for parallel video segment generation. The Skill orchestrates concurrent rendering of individual scenes to significantly reduce overall media production time.

How do I seamlessly stitch multiple video segments into a single timeline?▼

To seamlessly stitch multiple video segments into a single timeline, the Skill handles technical assembly by aligning transitions, timing, and scene pacing to produce a cohesive long-form video.

Does automated scene stitching maintain consistent visual style across shots?▼

Automated scene stitching maintains consistent visual style across shots by orchestrating multi-scene planning and synthesis parameters, ensuring the final assembled timeline reflects a unified narrative aesthetic.

What are the limitations of AI-driven storyboard generation for complex storytelling?▼

A limitation of AI-driven storyboard generation for complex storytelling is its reliance on script structure clarity; highly abstract narratives may require manual adjustments to achieve precise rhythmic control and desired visual pacing.