video-prompt-scaffold

Generates structured seven-part video prompts for Veo, Sora, and Kling generation models.

5|2|Updated Jul 12, 2026
One-click install
npx skills add https://github.com/LongLeo287/seosona-flow --skill video-prompt-scaffold-longleo287
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: video-prompt-scaffold
Source: https://github.com/LongLeo287/seosona-flow/tree/main/.claude/skills/video-prompt-scaffold
Command: npx skills add https://github.com/LongLeo287/seosona-flow --skill video-prompt-scaffold-longleo287

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve? Writing video generation prompts that models like Veo, Sora, and Kling actually follow is hard: vague prompts produce distorted motion, inconsistent characters, and generic AI slop. This Skill enforces a disciplined seven-part prompt structure so every scene has one clear subject, one action, and one camera move. ## Core Features & Use Cases - Seven-Part Prompt Scaffold: Fills Subject, Action, Shot & framing, Camera move, Setting & light, Style & lens, and Duration & audio in a fixed order, with a one-line template for fast drafting. - Model-Specific Guidance: Adapts prompt style per model — cinematic natural language with silent b-roll for Veo, multi-beat story sequences for Sora, and motion descriptions from a start frame for Kling image-to-video. - Anti-Slop Discipline: Bans empty filler words like cinematic or epic and forces concrete lens, lighting, and material specifications; integrates with the SEOSONA Flow gallery of ~24 pre-filled video prompt scaffolds. - Use Case: You need a 6-second product b-roll clip for an ad. The Skill produces a prompt like a specific subject performing one action, one dolly-in move, defined lighting, 85mm shallow depth of field, and no dialogue for later voiceover. ## Quick Start Ask the AI to write a Veo video prompt for your scene using the seven-part structure with one camera move and one action.

Frequently Asked Questions about video-prompt-scaffold

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I write a good prompt for Veo video generation?▼

Structure the prompt in seven parts: subject, action, shot and framing, one camera move, setting and light, style and lens, then duration and audio. Keep one action and one camera move per scene, and mark b-roll as no dialogue, ambient only for later voiceover.

What is the difference between prompting Veo, Sora, and Kling?▼

Veo responds best to natural cinematic descriptions with explicit audio directives around 8 seconds. Sora handles longer multi-beat story sequences in one prompt. Kling works image-to-video, so you provide a start frame and describe only the motion from it.

Why does my AI video come out distorted or inconsistent?▼

Distortion usually comes from stacking multiple camera moves or actions in one scene, which confuses the model. Split complex ideas into separate scenes, keep character, wardrobe, and palette consistent across scenes, and state duration explicitly.

Should I include words like cinematic or 8k in video prompts?▼

No. Empty modifiers like cinematic, epic, or 8k are treated as slop and removed. Replace them with concrete specifications such as lens focal length, depth of field, light direction, and film stock style.

Can I add dialogue or text directly in a generated video?▼

Dialogue can be included in parentheses if you intentionally want it baked in, but b-roll should stay silent with ambient audio only so voiceover can be added later. On-screen text should be added with a text overlay node rather than baked into the video.