In the Skill's Own Terms
Turns arbitrary text — an article, notes, a topic, a brief — into a faceless explainer video with invented scene visuals (typography, abstract graphics, diagrams, data-viz). Works through seven orchestrated steps: setup with brief confirmation, synthetic capture package (no website), design system selection from shipped frame presets with optional brand token remixing, storyboard and script using narrative design to reshape input, audio generation with TTS and BGM, visual design adding time-coded shot sequences with invented elements, parallel frame building via sub-agents, and final render.
User-gated checkpoints at Steps 0, 3, and 6 in collaborative mode. Faceless means no capture step and no real asset inventory — every visual is invented downstream by workers.
What it produces
- STORYBOARD.md with frame-by-frame teaching plan, time-coded shot sequences, and invented focal/roles
- compositions/frames/NN-*.html — one HTML composition per frame built by sub-agents
- renders/video.mp4 — the final rendered video in landscape 1920x1080, portrait 1080x1920, or square 1080x1080
How It Works
- 01Confirm Brief and Initialize Project
BRIEF.md exists and determines the mode — read it and ask nothing. If missing and project is fresh, run the intent layer to lock the brief.
- 02Create Synthetic Capture Package
Save the user's full input verbatim as capture/extracted/visible-text.txt (the source of information).
- 03Choose Frame Preset and Build frame.md
Pick one shipped preset from hyperframes-creative/frame-presets/ that fits the topic and tone.
- 04Write Storyboard with Narrative Design
Turn text into frame-by-frame teaching plan using story-design.md structure.
- 05Generate Audio and Add Visual Design
Run audio.mjs in background for TTS, word timings, BGM from HeyGen library.
- 06Build Frames and Assemble Index
Sync durations, build per-frame packets, dispatch one sub-agent per frame in parallel. Each worker writes compositions/frames/NN-*.html. Build captions in background.