Agency OS Skill Library

Video Prompt Builder

Turns a creative brief into a shot-by-shot video prompt with effects timeline, density map, energy arc and paste-ready blocks.

Video First result: 30 minutes video-prompt-builder
Back to all skills

What it does

Takes a creative brief, anything from one line to a full storyboard, and returns a complete five-section video prompt. You get a shot-by-shot effects timeline, a master inventory of every effect used, a density map showing where the video is busy and where it breathes, a three-act energy arc, and a paste-ready prompt block with one self-contained prompt per clip plus a STYLE LOCK footer that keeps every clip visually consistent. The prompts are model-agnostic architecture: shot structure, camera language, lighting and constraints transfer to whichever video tool your workspace has on record.

Say this to start

This skill has no button. You start it by saying what you want. Any of these will do it:

/build-video-prompt
> build me a video prompt
> shot list for this video
> b-roll planning
> ad concept shot by shot
> brand film shot list

When to reach for it

When NOT to use it

If you actually wantUse this instead
a still image prompt or the overall visual strategycreative-director
the spoken script of a sales videovsl-architect
editing, cutting or grading footage you already haveauto-editor
slicing an existing long video into short clipsvideo-repurposer

Before you start

What you needWhy
A creative brief, even a rough oneif the brief is too vague to build from, it will ask one focused clarifying question before proceeding, which costs you a round tripRequired
The workspace tool-stack file at protocol/current-tool-stack.mdit names the video tool the prompts are targeted at, and the paste-ready section is formatted for itRequired
voice-profile.md in the active workspacesupplies brand personality, colour preferences, energy level and aesthetic direction; without it the visual tone is genericOptional
positioning.md in the active workspacemarket tier, luxury through to accessible, shapes visual treatment, pacing and effects densityOptional
offer-output.md in the active workspacegives product context and audience, which shape the shot list's narrative and CTA framingOptional
An account, credits and API access for the video tool on recordonly needed if you want it to chain straight into generation. The default path just hands you the promptsOptional
A duration targetwithout one it defaults to 10 to 15 seconds, which may not be what your placement needsOptional

How it runs

  1. You give it the briefAnything from 'a runner in a stadium, 10 seconds' to a detailed storyboard. Useful additions are subject, setting, mood, brand context, specific camera moves, duration target, colour palette and any reference films or ads.
  2. It loads the rules and the tool of recordIt reads the prompting reference that holds the model constraints, checks the workspace tool-stack file for which video tool to target, and pulls the breakdown reference closest to your format: athletic brand film, product demo, short social ad or longer brand story.
  3. It picks up workspace contextBrand personality and colour cues from your voice profile, market tier from positioning, product context from your offer, and any recent feedback logged against this skill or creative-director. It mentions the adaptations it made.
  4. It calibrates shot count to durationDuration drives everything: 4 to 6 seconds gets 3 to 5 shots and one density peak, 20 to 30 seconds gets 14 to 20 shots and a full three-act arc. If you did not name a duration it uses 10 to 15 seconds.
  5. It writes the four analysis sectionsThe shot-by-shot timeline with effect, visual description, camera, timing, lighting, constraints and transition per shot; the master effects inventory with usage counts; the density map rating each 3 to 6 second segment high, medium or low; and the energy arc across acts. The most impactful shot is explicitly flagged as the signature visual effect.
  6. It writes the paste-ready blockEach shot collapses into a standalone 50 to 100 word prompt using the six-element formula: subject, action, environment, camera, style, constraints. A STYLE LOCK footer holds the style, lighting and constraint keywords you paste into every single generation so the clips match.
  7. You pick a pathThe default is prompt only: take the paste-ready block to whichever tool you use, no API keys needed. If you want end-to-end generation and have the platform's tools and key configured, it hands the prompts to the execution skill for the tool on record.
  8. Feedback loopIt asks whether the prompt is ready as-is, needs minor tweaks, or needs rework, and logs the answer with the video type and duration so the next prompt adapts.

What you get

Honest limits

Read this before you rely on it

Where people go wrong

The mistakeDo this instead
Writing one sentence that moves both the camera and the subjectSplit them. 'The dancer spins slowly in the centre of the stage. Camera holds fixed framing.' This is the most common failure mode across every current video model.
Stacking simultaneous camera movements to get a richer shotOne primary camera instruction per shot. Sequential compounds are fine, 'push-in then subtle rise'; simultaneous ones cause jitter.
Writing longer prompts to get more controlStay at 50 to 100 words per shot. Longer prompts accumulate conflicting instructions and degrade quality. Be specific, not verbose.
Asking for fast motion throughoutUse fast sparingly, one fast element per shot at most. Unqualified 'fast' degrades output. Slow and medium have their own vocabulary that models read reliably.
Generating each clip from its own prompt and wondering why they do not matchPaste the STYLE LOCK block into every single generation. It exists to hold style, lighting and constraints identical across clips.
Worth knowing

If you can only add one detail to a shot, add lighting. It is the single highest-impact element in a video prompt, and every shot in the timeline gets a lighting line for that reason.