synthesia
synthesia
The avatar-video tool skill — have a reason for an avatar, use consented likeness only, make the script
spoken-word, assemble + localize, note the disclosure. The agent scripts and plans; the human approves every
video; WoopSocial publishes. (Ships with tools/integrations/synthesia.md.)
The POV: the presenter is synthetic — the standards stay human
Synthesia is the enterprise avatar category leader: one locked script becomes a consistent presenter in 140+
languages with no re-shoots, which makes it unbeatable for training, onboarding, explainers, and localization
at scale. The top-1% operator holds four lines. (1) The fit test comes first: avatars read
polished-but-clinical — they lose to a real face for trust-led founder content and testimonials (route
those to talking-head-and-piece-to-camera); the pro move is the hybrid — the founder films the trust layer,
the avatar scales the informational layer. (2) Consent is the architecture, not friction: stock avatars are
paid consenting actors; a personal avatar requires your live consent recording on an unspliced single-take
source — and nobody gets an avatar of a competitor, celebrity, or anyone who hasn't verifiably consented.
(3) The script is most of avatar quality — and it locks before render: spoken-word writing (short
sentences, SSML, read aloud), because a comma-level edit forces a full ~8–12-minute re-generation off the
minute cap. (4) Disclosure, always: a synthetic presenter is labeled — platform AI tags and the EU AI Act's
synthetic-media obligations make undisclosed avatars a channel-level risk.