Synthesia

  • Synthesia (Synthesia Ltd., London 2017, Active)

    Status: Active (Victor Riparbelli CEO; 4B post-money Jan 2026; 90% Fortune 100 / 95% DAX 40 客户; SOC 2 Type II + GDPR + ISO 42001 + SAML/SSO). Category ai-productivity-data; pipeline_stage=none. Not text-to-video like Runway/Kling (those are scene-gen), not a clip repurposer (→ Opus Clip/Submagic)​ — it is the enterprise talking-head avatar platform: script → 240+ stock or Personal digital-twin avatar → Express-2 micro-expression engine (lip-sync/gesture/eye-contact) → 160+ lang voice + AI Dubbing → SCORM/LMS-ready MP4. Avatar realism + L&D governance (SCORM/SSO/translation) is the moat; per-minute video caps + Enterprise-only SCORM are the gripes.

    1. English Introduction

    Synthesia is “the AI presenter studio for business — no camera, no actor, no studio.” Paste a script, pick a 240+ stock avatar (or train a Personal Avatar from a 2-min phone video, or pay $1,000/yr for Studio Express-1), choose from 160+ languages and natural-sounding voices, and the Express-2 engine renders facial micro-movements, hand gestures, and lip-sync matched to speech. Edit in a slide-like canvas (templates, screen recorder, PPT import, brand kit), enable 1-click translation into 80+ languages with re-dubbed lip-sync, export MP4 or SCORM for LMS, or call the REST API (Creator+ 360 min/yr included) to generate programmatically. Used by Merck, SAP, Mondelez, Bosch for training/onboarding/compliance/localized explainers. Pitch: HeyGen is marketer-friendly; Synthesia is the audited, SCORM-and-SSO version for L&D and compliance teams.

    2. Core Features

    Layer Capability Notes
    Avatars 240+ stock (Express-2 expressive), Personal Avatar (phone video, 3/5/unlimited by tier), Studio Express-1 ($1,000/yr add-on) Closest to real presenter in lane
    Voice/locale 160+ languages, AI voices, voice clone, AI Dubbing w/ lip-sync re-sync, 1-click translate 80+ langs (Ent) Translation = 40% of generated vols
    Canvas Slide-like editor, 250+ templates, screen recorder, PPT import, multi-avatar scenes, interactive CTAs/quiz, auto-captions Not timeline NLE
    Governance SCORM export (Ent), SAML/SSO (Ent), Brand Kit, LMS embed, priority content moderation, audit trails L&D wedge
    API REST video gen, Creator+ 360 min/yr, Enterprise custom Was Ent-only pre-2026, now on Creator
    Analytics Views, completion, drop-off, per-share-page
    Access Web studio only No desktop, no mobile editor, no MCP
    No No text-to-scene (Runway/Kling), no real-time conversational (Tavus), no free-tier commercial use, credits no rollover Minute-metered

    3. Pricing (2026-08, USD, credit-pool = video min + dub min shared, no rollover)

    Tier Monthly Annual eff/mo Video min/yr Avatars Gates
    Basic/Free $0 $0 ~3 min/mo (10/yr) 9 stock Watermark, no download, no commercial
    Starter $29 **216/yr) 120 min/yr (10/mo) 125+ stock, 3 Personal Logo off, download, AI Dubbing 120 min/yr, 1 editor+3 guests
    Creator $89 **768/yr) 360 min/yr (30/mo) 180+ stock, 5 Personal API 360 min/yr, multi-avatar, interactive, brand kit, priority support
    Enterprise Custom Custom Unlimited 240+ stock, unlimited Personal SCORM, SSO, 1-click 80+ lang, CSM, DPA, custom credits

    pricing_type=freemium-per-seat-video-minutes-credit-pool; starting_price=$18/mo (first paid); annual ≈ 38% off; Studio Express-1 avatar $1,000/yr add-on on annual plans; unused minutes/credits no roll; overage = upgrade or wait cycle.

    4. Access Type

    🌐 studio.synthesia.io web · REST API (Creator+ / Enterprise) · no desktop app, no mobile editing, no DAW/plugin, no MCP.

    5. Reviews (2026)

    • G2 4.7/5 (2,542), Capterra 4.6/5 (313), aggregate ~4.7/5 (2,855)​ — avatar realism/Express-2, lip-sync, 160-lang dubbing, L&D SCORM fit most praised.
    • Complaints (Trustpilot mixed): auto content-moderation flags compliant scripts, slow human review, auto-renew non-refundable (Capterra), 10 min/mo on Starter too tight, uncanny in close-up, Enterprise opaque quoting, credit pool eats fast if using AI Dubbing + Playground clips.
    • Verdict: best enterprise/L&D avatar video with compliance trail; wrong for real-time conversational agents (Tavus), cinematic text-to-scene (Runway/Kling/Sora), or budget solo creators (HeyGen/Colossyan cheaper).

    6. Best For / Not For

    Best For Not For
    L&D / HR / compliance teams shipping multilingual training to LMS (SCORM+SSO) Real-time chat avatars / voice agents (→ Tavus)
    Enterprises localizing explainers into 80+ langs with audit trail Cinematic scene generation from prompts (→ Runway/Kling/Sora)
    Companies wanting digital-twin of execs for repeat internal comms Solo creators needing >30 min/mo under $30 (→ HeyGen/Colossyan)
    Regulated industries (SOC2/GDPR/ISO42001, DPA) Premium ad campaigns needing photoreal talent (uncanny risk)
    Teams already in PPT/Slides workflow Mobile-first / no-web-access buyers

    7. Competitors (Lane Map)

    Lane Tool Synthesia edge / gap
    Avatar video HeyGen​ (in your AI-Avatar lane) HeyGen marketer-friendly, cheaper entry, digital-twin focus; Synthesia SCORM/SSO/translate depth
    Avatar video Colossyan Colossyan e-learning scenarios stronger; Synthesia avatar realism + Fortune 100 trust
    Photo-talker D-ID D-ID cheap API photo-talker; Synthesia full-body expressive + governance
    Real-time conv Tavus Tavus conversational/realtime; Synthesia async presenter only
    Text-to-scene Runway, Kling, Sora Scene gen, not talking-head; different lane
    Edit+transcript Descript Descript edits real footage+voice; Synthesia generates avatar from text
    In-list overlap Async(former Podcastle), Submagic, Opus Clip Those record/repurpose real media; Synthesia synthesizes presenter
    Voice lane ElevenLabs/Murf/Auphonic Voice-only/master; Synthesia binds voice to avatar

     

Do Not Sell or Share My Personal Information Cookie Settings