skills/skills.volces.com/ppt-audio-to-video

ppt-audio-to-video

Installation
SKILL.md

PPT Audio To Video

Use this skill when the source video has narration audio but no usable slide visuals, and the final deliverable should be a slide-based lecture video.

Resolve bundled scripts relative to this skill directory. If the runtime has already opened this SKILL.md, prefer paths like scripts/extract_slide_outline.py and scripts/render_from_timing_csv.py instead of machine-specific absolute paths.

Core workflow

  1. Inventory inputs.

    • Confirm which of these exist: audio-only mp4/m4a/mp3/wav, ppt/pptx, pdf, and any pre-rendered slide images.
    • Prefer an existing pdf or image directory for rendering. Treat pptx as the source of slide text and as a fallback for export.
  2. Prepare tools.

    • Required for deterministic steps: ffmpeg, ffprobe, pdftoppm.
    • Required for transcription: whisper-cli from whisper-cpp plus a multilingual model such as ggml-small.bin.
    • If only pptx exists and no pdf/images exist, prefer Keynote or PowerPoint export on macOS. Use soffice only as fallback because profile or rendering issues are common.
Installs
2
First Seen
Apr 23, 2026
ppt-audio-to-video from skills.volces.com