hume-evi
Installation
SKILL.md
Hume EVI production guide
Use this skill when the task is to design, implement, troubleshoot, or review a Hume Empathic Voice Interface (EVI) realtime voice agent. Do not use it for offline-only Hume Text-to-Speech/Octave work unless the user is explicitly configuring EVI voices or comparing EVI to TTS.
Hume's EVI is a realtime speech-to-speech agent interface. It streams user audio, measures expressive vocal modulation, generates a language response, and produces expressive assistant speech. Treat it as a live conversation system, not as a batch transcription or batch TTS API. Documented facts below were verified on 2026-07-10 unless a source date is explicitly noted.
Primary official sources:
- Hume EVI overview: https://dev.hume.ai/docs/speech-to-speech-evi/overview
- EVI version guide: https://dev.hume.ai/docs/speech-to-speech-evi/configuration/evi-version
- Configuration guide: https://dev.hume.ai/docs/speech-to-speech-evi/configuration/build-a-configuration
- Chat WebSocket API reference: https://dev.hume.ai/reference/speech-to-speech-evi/chat
- Audio guide: https://dev.hume.ai/docs/speech-to-speech-evi/guides/audio
- Tool use guide: https://dev.hume.ai/docs/speech-to-speech-evi/features/tool-use
- Prompting guide: https://dev.hume.ai/docs/speech-to-speech-evi/guides/prompting
- Privacy controls: https://dev.hume.ai/docs/resources/privacy
- Pricing page: https://www.hume.ai/pricing
- Acceptable Use Policy: https://www.hume.ai/acceptable-use-policy