NEWS · COMPANIES · #1074
Synthesia creates interactive digital avatar of a journalist using its video and voice models
Synthesia built an interactive digital avatar (a “digital twin”) of a journalist at its new New York office, training it on a single article and making it deterministic so it only answers questions related to that story. The avatar was produced from photos and a two‑minute voice sample and uses Synthesia’s own video and voice models (customers can opt to use alternatives like Cartesia, ElevenLabs, Google or OpenAI and choose hosting).
KEY POINTS
- Synthesia built an interactive digital avatar (a “digital twin”) of a journalist at its new New York office, training it on a single article and making it deterministic so it only answers questions related to that story.
- The avatar was produced from photos and a two‑minute voice sample and uses Synthesia’s own video and voice models (customers can opt to use alternatives like Cartesia, ElevenLabs, Google or OpenAI and choose hosting).
- This demonstrates a practical enterprise use of agentic avatars and the tech stack (voice‑to‑text, language model, text‑to‑voice, video animation), highlighting both product capability and consent/liability tradeoffs for likenessed AI.
WHY IT MATTERS
This demonstrates a practical enterprise use of agentic avatars and the tech stack (voice‑to‑text, language model, text‑to‑voice, video animation), highlighting both product capability and consent/liability tradeoffs for likenessed AI.