--- name: short-form-video description: Build and iterate short-form vertical (9:16) videos in Hyperframes - TikTok/Reels/Shorts style. Use when the user says "short-form video", "vertical video", "TikTok/Reels/Shorts", "make a short", "talking-head + motion graphics", or when the target is a 1080x1920 composition with face video + synced scene overlays + karaoke captions. Encodes the full May Shorts 19 playbook: face-mode choreography, audio-synced scene timing, karaoke captions, and the 10-rule quality checklist. --- # Short-Form Vertical Video (Hyperframes) ## ⚖️ STYLE DESIGN - HIGHEST PRIORITY RULE (read this first) If the edit prompt contains a **"STYLE DESIGN (BẮT BUỘC TUÂN THỦ 100%)"** section: the palette, fonts and tone in that section **COMPLETELY REPLACE** every color/font/branding rule in this skill (including "dark fintech blue", the GĐT hex codes, the default gradients...). Keep only: animation technique, layout, cut rhythm, render workflow, and the bug fixes. Illustrations generated via /api/illustrations MUST pass the correct styleId in the prompt. NO Style Design section -> use this skill's default branding. If the edit prompt ALSO contains a **"PHONG CÁCH DỰNG"** (video style) section: that section **WINS over this skill** for ALL visual and motion language - materials, transitions, camera moves, effects. This skill then keeps only PROCESS: step order, cutting, captions/keys, draft->final, QC. Do not mix the skill's default visuals into the video style - half-and-half is exactly the failure this rule exists to prevent. Short-form = 1080x1920 vertical, 10–30s, talking-head face + motion-graphic scene overlays + karaoke captions. Everything in this skill is distilled from the May Shorts 19 iteration autopsy (v1 → v4) and should be applied on every new short. **Always invoke `/hyperframes` first.** This skill sits on top of it - it does not replace the framework rules (`data-*` attributes, `window.__timelines`, composition structure). Those are non-negotiable regardless of the format. ## When this skill fires - "Make a short-form video", "TikTok post", "Reels", "Shorts", "vertical video" - Any build starting from a talking-head recording + script/transcript intended for social - Retiming, recutting, or re-syncing an existing short - Adding karaoke captions synced to a voiceover ## The playbook (high-level) 1. **Audio is source of truth.** Edit audio FIRST (cut retakes, pauses). Save as `-edit.mp4`. Measure exact duration with `ffprobe` - this is the composition's `data-duration`. 2. **Transcribe the edited audio** with `npx hyperframes transcribe .mp4 --model small.en --json` (English only - for Vietnamese use faster-whisper per the noti skills), or if retiming an existing build with a `shift()` function in captions, keep the existing captions and just shift scene starts. 3. **Author scene boundaries in edited-time** - NEVER mix original-time and edited-time anchors in the same file. See "Audio-sync protocol" below. 4. **Build the composition scaffold** (4 layers: ambient-bg, seam-treatment, captions, face) - see "Composition scaffold" below. 5. **Author scenes with LOCAL offsets** relative to each scene's `data-start`. Each scene is its own sub-composition under `compositions/scene-