text-to-video-direct
Installation
SKILL.md
Direct Text-to-Video Pipeline
Single model call from a text prompt to a clip. No intermediate still.
When to use
- Exploratory or B-roll shots where you don't need to lock the opening frame.
- Motion-heavy shots where the model's interpretation of motion matters more than composition.
- Faster iteration when image-then-animate is overkill.
When NOT to use
- Character-specific shots needing visual consistency — direct text-to-video drifts heavily across takes. Use
text-to-image-to-videoinstead. - Talking heads — use
voice-to-lip-syncover a generated character clip.
Inputs
- Shot brief from
scripts/storyboards/NN-*.md. - Text-to-video model from
brief/tools-and-models.md.