Interactive Avatars

Build real-time, conversational avatar experiences on your own stack. You bring the LLM and speech-to-text; Synthesia renders the avatar, with optional built-in text-to-speech if you don't want to bring your own.

You'll need Python 3.10+ and a Synthesia API key with Interactive Avatar access to get started.

Get oriented

Start with the Minimal quickstart.

Building something more specific?

Or integrate directly

Not using the Python plugin? You can call the API yourself. You still connect to a LiveKit room, but you mint the token and manage the session instead of the plugin doing it for you.

Troubleshooting


API updates · Status