faion-multimodal-ai
Installation
SKILL.md
Entry point:
/faion-net— invoke this skill for automatic routing to the appropriate domain.
Multimodal AI Skill
Communication: User's language. Code: English.
Purpose
Handles multimodal AI applications. Covers vision, image generation, video generation, speech, and voice synthesis.
Scope
| Area | Coverage |
|---|---|
| Vision | GPT-4o Vision, Gemini Vision, image understanding |
| Image Generation | DALL-E 3, Midjourney, Stable Diffusion |
| Video Generation | Sora, Runway, Pika |
| Speech-to-Text | Whisper, Deepgram, AssemblyAI |
| Text-to-Speech | OpenAI TTS, ElevenLabs, Google TTS |
| Voice | Real-time voice, voice cloning |