Home/industry/Synthesia rolls out a conversational digital twin for a journalist
Create an original premium technology-news editorial illustration featuring a realistic digital avatar displayed on a high‑resolution monitor as the dominant element, positioned opposite a human journalist seated at a sleek studio desk with a microphone; the scene is set inside Synthesia’s modern New York office studio, with subtle branding on a wall sign that reads “Synthesia” and visible lighting rigs and cameras in the background; the avatar is animated, mouth moving, conveying conversation, while the journalist looks
IndustryPublished 27 September 2026 · 18:523 min read

Synthesia rolls out a conversational digital twin for a journalist

Synthesia’s Interactive Avatar Launch

Synthesia, a video‑generation startup, unveiled an interactive digital avatar of a journalist, marking its first public avatar for anyone outside the company.

The avatar is programmed to answer press‑related questions about Synthesia and is limited to a single story supplied by the journalist.

The demonstration occurred during a visit to Synthesia’s new New York office, an expansion from its United Kingdom headquarters.

Earlier this year the company announced a $4 billion valuation and reported more than $100 million in annual recurring revenue.

Synthesia competes with other avatar providers such as D‑ID, HeyGen and Colossyan.

Technical Build of the Conversational Twin

To create the avatar, the team set up a mini film studio inside the office and captured a series of photographs of the journalist.

They also recorded a two‑minute audio sample of his voice, which serves as the source for speech synthesis.

The journalist gave explicit consent for the avatar’s creation and use.

Synthesia produced four versions: a personal avatar that reads any script, with and without glasses, and an interactive avatar that can listen and respond, also with and without glasses.

The interactive version combines voice‑to‑text, a language model, text‑to‑voice, and a video animation model.

The voice‑to‑text component transcribes spoken questions into text for processing.

The language model interprets the text, decides on an action, and generates a textual reply.

The text‑to‑voice engine converts that reply into spoken audio.

Finally, Synthesia’s proprietary video model animates the avatar’s facial movements to match the audio.

While Synthesia supplies its own video and voice models, customers may opt for alternatives such as Cartesia, ElevenLabs, Google or OpenAI.

Enterprises also have the choice to host avatars on their preferred cloud provider or pay Synthesia for hosted services.

The entire pipeline was assembled in a matter of days, according to the development team.

This rapid turnaround demonstrates the modularity of Synthesia’s stack.

By exposing an API, Synthesia enables developers to integrate its video and voice capabilities into custom applications.

The API can be combined with third‑party services to build bespoke interactive avatars or other multimedia products.

Business and Ethical Implications

The journalist’s avatar is deliberately constrained to answer only about his story on venture‑backed startups and fraud.

This narrow focus illustrates how companies can limit an avatar’s knowledge base for compliance or branding reasons.

The interactive avatar represents a shift from static, pre‑recorded videos toward conversational experiences.

Earlier avatar offerings required users to write scripts that the avatar would recite verbatim.

The new capability aligns with growing interest in AI‑driven training and customer‑facing interactions.

Organizations can deploy similar avatars for internal knowledge bases, product demos, or support desks.

However, the technology also raises questions about consent, data privacy, and the authenticity of AI‑generated speech.

Synthesia’s model permits clients to select cloud hosting, which may affect data residency and regulatory compliance.

The company’s valuation and ARR suggest strong market demand for such AI video solutions.

As more firms adopt interactive avatars, the competitive landscape among D‑ID, HeyGen and Colossyan is likely to intensify.

The journalist’s experience provides a concrete example of how a personalized digital twin can be created quickly and used for targeted communication.

Why This Matters: The avatar shows that enterprises can now deploy conversational digital twins at scale, offering efficient, controllable outreach while foregrounding consent and data‑governance challenges.

#industry#ai#digest#auto

This digest was compiled from:

Share this digest

Share on XWhatsAppLinkedInTelegram

People Also Read

Share your thoughts

Reactions, corrections, or insights — all welcome.

0/2000