{"slug": "synthesia-interactive-avatar-api", "title": "Synthesia Interactive Avatar API", "summary": "Synthesia launched its Interactive Avatar API today for Enterprise customers, letting developers embed real-time, photorealistic, lip-synced avatars into web products using the company's new Express-3 model. The API takes a Bring-Your-Own-Stack approach, allowing customers to connect their own LLM from OpenAI, Anthropic, Google or other providers, plus their own speech-to-text and optional text-to-speech, while Synthesia handles avatar rendering via LiveKit. Synthesia cited Gartner's forecast that by 2028 seven in 10 customer journeys will begin with a conversational third-party interface, and said Enterprise customers are already shipping Interactive Avatars in production for HR onboarding, customer service and in-person kiosks.", "body_md": "[Blog](https://www.synthesia.io/blog)\n\n# Interactive Avatar API: A New Way to Embed Real-Time Avatars Anywhere\n\nCreate AI videos with 240+ avatars in 160+ languages\n\n### Summary\n\n- Developers can now embed a real-time, photorealistic, lip-synced interactive avatar or our brand new Style Avatars powered by our latest [Express-3 model](https://www.synthesia.io/post/synthesia-launches-new-ai-assistant-for-faster-video-creation-and-express-3-avatar-model) into their web product via our API.\n- We’re taking an open approach: you can quickly combine your own stack, including your own LLM, speech-to-text (STT) and text-to-speech (TTS) systems, with our API to create an integrated system. Synthesia handles the avatar rendering; you have the freedom to choose the rest.\n- Interactive Avatar API launches today for Enterprise customers: you can use it as a support concierge, a live FAQ agent, or anywhere you need a conversational interface that actually listens and responds.\n\n## Interactive Avatars API: a new way to interact with users\n\nGartner [forecasts](https://www.gartner.com/en/newsroom/press-releases/2025-02-10-traditional-customer-service-channels-are-losing-ground-to-mobile-and-ai-innovations) that by 2028, seven in 10 customer journeys will begin with a conversational third-party interface. Until now, embedding a real-time, photorealistic interactive avatar into a website or app meant choosing between closed-stack vendors which offered limited customization and vendor lock-in or hand-stitched integrations with high latency and poor reliability.\n\nToday, we’re launching the Synthesia Interactive Avatar API to shift from broadcast video to actual conversation: real-time, interruptible, personalized, and agentic. Our Avatars can respond in any language and in real-time, and will soon be able to take actions such as scheduling follow-up meetings.\n\nEnterprise customers are already shipping Interactive Avatars in production, from HR onboarding and customer service to in-person kiosk experiences. You can experience it yourself by talking to [Synthesia’s own press officer here](https://www.synthesia.io/post/synthesia-ai-press-officer). \n\n## Three key capabilities\n\nSynthesia is taking a Bring-Your-Own-Stack (BYO) approach to Interactive Avatars, enabling customers to supply their own LLM, speech-to-text and text-to-speech providers while Synthesia handles the real-time, lip-synced avatar rendering.\n\n**1. Bring your own AI stack, by default.** Connect your own LLM from OpenAI, Anthropic, Google or any other provider, STT (any provider), and optionally TTS. Synthesia layers the avatar on top via LiveKit, and we also offer a Synthesia-hosted TTS inside the plugin as an add-on.\n\n**2. Real-time, streaming lip sync that feels natural.** The avatar responds with idle and listening behaviour, so conversations flow smoothly even during pauses. Interruptions and turn-taking work out of the box. Customers can build completely customizable avatars from photorealistic to branded characters like an animated mascot, with custom backgrounds.\n\n**3. Language-agnostic and customizable.** Lip sync works with a number of languages including French, German or Spanish, as well as accents. Use stock synthetic avatars or create custom ones, including branded Style Avatars from Avatar Builder, and make them interactive via API.\n\n## The pipeline, input to interaction\n\nHere is technology stack needed to build an interactive video experience:\n\n- **Input (TTS):** When the user speaks, the voice captured in real-time is converted to text and streamed in real-time.\n- **LLM layer:** It is the brain of your agent. It has all the context, guardrails and system prompt to respond to the user.\n- **Voice synthesis (STT):** Responses delivered in the avatar’s cloned voice - consistent, on-brand, and instant.\n- **Real-time rendering:** Photorealistic or cartoon based avatars with lip-syncs, expressions and gestures that drive the engagement with the user.\n- **Agentic calls:** Help you extend beyond conversation! Let your agent take follow-up actions – like schedule a meeting or fill a form.\n\nOwning the conversation layer gives customers control over data, guardrails and system prompts. Today’s release also unlocks agentic functions, so an avatar can book meetings, complete a form or trigger a workflow mid-conversation rather than just respond to it.\n\n## Developer experience\n\nWe’ve built this for teams that ship production-ready AI experiences, not demos:\n\n- Official LiveKit Agents plugin (livekit-plugins-synthesia, Python): one import, minimal scaffolding\n- Public quickstart repo with minimal bootstrap, RAG examples, and tool-use patterns: [github.com/synthesia-ai/interactive-avatar-quickstarts](http://github.com/synthesia-ai/interactive-avatar-quickstarts)\n- Cursor/Claude skill for agent-led integration: [github.com/synthesia-ai/skills](http://github.com/synthesia-ai/skills)\n- Full API docs with reference and guides: [docs.synthesia.io/reference/interactive-avatars](http://docs.synthesia.io/reference/interactive-avatars)\n\n## Enterprise-grade from day one\n\nInteractive Avatar runs on the same platform businesses already trust for their video, offering enterprise-grade security (SOC 2 Type II, ISO 27001), systems management (ISO 42001), data privacy (GDPR and CCPA compliance), and workspace controls (SSO and SCIM). Because you bring your own AI, your data governance model doesn't change: the Avatar never sees your knowledge base, and your model never leaves your infrastructure.\n\n## Available now\n\nInteractive Avatars are available now for Enterprise customers. The full-stack pipeline (Synthesia-hosted LLM + TTS + avatar, no code needed) will launch on October 1, 2026, for teams who prefer a managed approach.\n\nTry it out here: [https://www.synthesia.io/features/avatars/interactive-avatars](https://www.synthesia.io/features/avatars/interactive-avatars)\n\n## Alessandra Venier\n\nAlessandra Venier is a Corporate Affairs Manager at Synthesia\n\n[Go to author's profile](https://www.synthesia.io/blog/authors/alessandra-venier)", "url": "https://wpnews.pro/news/synthesia-interactive-avatar-api", "canonical_source": "https://www.synthesia.io/post/interactive-avatar-api-real-time-avatars", "published_at": "2026-09-21 21:11:27+00:00", "updated_at": "2026-09-21 21:23:29.465188+00:00", "lang": "en", "topics": ["ai-products", "ai-agents", "generative-ai", "large-language-models", "ai-tools"], "entities": ["Synthesia", "Interactive Avatar API", "Express-3", "LiveKit", "OpenAI", "Anthropic", "Google", "Gartner"], "alternates": {"html": "https://wpnews.pro/news/synthesia-interactive-avatar-api", "markdown": "https://wpnews.pro/news/synthesia-interactive-avatar-api.md", "text": "https://wpnews.pro/news/synthesia-interactive-avatar-api.txt", "jsonld": "https://wpnews.pro/news/synthesia-interactive-avatar-api.jsonld"}}