groups-compose
Designs and maintains semantic groupings and readable layouts on the filmmaking canvas — scenes, character-reference sets, act beats, and other titled visual…
Orchestrates story, script, screenplay, concept, product promo, and multi-shot idea work into finished video. Use first when the user asks to make a video from a story or script; asks what next in a story video project; or needs a decision spanning script splitting, image refs,
$ npx -y skills add Utopai-Research/pai-pro --skill story-to-video-workflow --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/story-to-video-workflowContext preview
The summary Claude sees to decide when to auto-load this skill.
Orchestrates story, script, screenplay, concept, product promo, and multi-shot idea work into finished video. Use first when the user asks to make a video from a story or script; asks what next in a story video project; or needs a decision spanning script splitting, image refs,
name: story-to-video-workflow description: >- Orchestrates story, script, screenplay, concept, product promo, and multi-shot idea work into finished video. Use first when the user asks to make a video from a story or script; asks what next in a story video project; or needs a decision spanning script splitting, image refs, voices or VO, video clips, render strategy, Timeline ordering, or final Timeline handoff. Routes execution to script-compose, image-compose, voice-compose, and video-compose before those skills' CLIs are used.
Use this ladder unless the user skips, reorders, supplies refs, or asks for a rough direct render:
1. Clarify only blockers. 2. Raw idea/story -> `script-compose` production script; existing screenplay -> capture/adapt. 3. `script-compose` splits <=15s dialogue-aware shots and extracts characters, material variants, detailed locations/location variants, and speaker/VO needs. 4. `image-compose` creates useful visual anchors: base/variant character sheets and detailed location/detail anchors. 5. `voice-compose` creates reusable anchors for every speaker and VO/narrator. 6. Confirm shot count, durations, continuity needs, and first blocker. 7. Default render path: straight-to-video from refs. Storyboard only if requested, hard to control, or needed for diagnosis. 8. Default dispatch: hybrid. Chain continuous dependent shots; render independent scenes/shots in parallel. 9. Render clips, assign Timeline `shot_id` when sequence order is unambiguous, then hand off to Timeline.
Plan ahead internally, but only ask the next meaningful user-facing choice; the Consent and gates ladder fixes when render path and dispatch become askable.
| Need | Load next | |---|---| | Script capture, rewrite, split, or analysis | `script-compose` | | Character, location, storyboard, starting frame, or visual anchor | `image-compose` | | Narration, dialogue read, character voice, or audio node | `voice-compose` | | Clip render, continuation, audio refs, storyboard animation, or video prompt | `video-compose` | | Scene/ref grouping or canvas layout frames | `groups-compose` |
Capability skills own CLI flags, node grammar, refs, and recovery hints. `PROJECT_AGENT.md` owns shared failure handling.
Follow the project `PROJECT_AGENT.md` § "Recommendation and choice shape". Recommend one concrete next step. Add a second option only when there is a real tradeoff.
Before recommending refs/video, inspect `workflow.json` when needed and summarize only:
If the story implies more than roughly 3 minutes, recommend narrowing scope before clip planning.
After shot notes, missing video-bound character/location/voice anchors are the default next step; include a rough-direct skip when speed matters. Once anchors/user refs/rough-direct are settled, offer only a short ref review or clip-plan confirmation if ambiguity remains.
Ask only after the script/shot plan is settled and anchors, usable refs, rough-direct, or a simple single-clip case make rendering real. If anchors are still missing, return to Planning checkpoint.
Use project choice shape:
description: `Fastest path to motion.`
description: `Generate storyboard images first for composition control.`
For storyboard-first, load `image-compose` Pattern 6: one composite mosaic per clip/<=15s shot note, subtype `storyboard`.
Ask only after render path is picked and a multi-clip plan exists. Skip for one clip. Use project choice shape:
description: `Chain within continuous scenes; render separate scenes independently.`
description: `Render all clips independently.`
description: `Each clip continues from the previous one; boundaries default to a hard cut to a new angle (avoids the same-shot seam) — keep a boundary same-shot only for an unbroken oner.`
Signals: continuous scene/state -> sequential (hard-cut handoffs between clips); a single unbroken action the viewer must read as ONE motion -> one ≤15s clip, else sequential with a same-shot handoff; separate scenes/time jumps/wardrobe changes/montage -> parallel; continuous clusters separated by hard cuts -> hybrid. Do not chain video refs acr
The local AI filmmaking studio, driven from your coding agent. ![Discord][discord-url] [][claude-code-url] [][codex-url]
Repo: Utopai-Research/pai-pro
Designs and maintains semantic groupings and readable layouts on the filmmaking canvas — scenes, character-reference sets, act beats, and other titled visual…
Generates/edits filmmaking canvas images via generate_image.js and generate_image_pro.js. Use before image CLIs for character/location design, refs, starting…
Handles explicit screenplay/story work on the filmmaking canvas. Triages screenplay (use verbatim), story/concept (iterate then rewrite), or neither (defer).…
Generates and prompts video clips on the filmmaking canvas. Use when the user asks to generate, render, animate, continue, restyle, edit, shoot, or compose a…
Designs and attaches voice samples or final narration/line audio on the filmmaking canvas via the local generate_voice.js CLI. Use before calling…