fun-asr
Installation
SKILL.md
Fun-ASR: Audio Transcription
Transcribe audio files using Alibaba Cloud DashScope's Fun-ASR non-real-time speech recognition model. Supports speaker diarization, multi-language recognition, and multiple output formats (plain text, JSON, SRT subtitles).
Workflow
Audio → S3 → Fun-ASR async → Poll → Save file → Agent reads
Requirements
Environment Variables
Set these in ~/.wmyskills/.env (shared across skills), or scripts/.env (takes priority). Use scripts/.env.example as the template — add the variables to ~/.wmyskills/.env, never overwriting an existing file: