May 24, 2021 · 26m · mad

Fireside Chat: Victor Riparbelli (Founder & CEO, Synthesia) with Matt Turck (Partner, FirstMark)

Victor Riparbelli · 12m spoken Matthias Niessner · 5m spoken Matt Turck · 4m spoken Synthesia AI Avatar · 1m spoken
0:00 / 0:00
▶ Watch on YouTube →

gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions

In this Data Driven NYC fireside chat, Matt Turck interviews Synthesia co-founders Victor Riparbelli and Matthias Niessner about the technological foundations, enterprise use cases, safety ethics, and future roadmap of AI-generated video avatars.

How this conversation actually went

Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. Matt holds 19.1% of the talking time here. How this is scored →

Matt as informed peer 3.3 Guest teaching 4.3 Guest disagreement 0.3 Matt pushing back 0.4
05100:0010:0020:000:22–4:21 · Matt as informed peer 3/10 Synthesia AI Avatar Product Presentation Video Matt provides a clear introductory overview combining computer vision and voice AI before asking the guest to explain underlying generative networks. Matthias details how computer graphics methods evolved from film industry stunt editing into deep learning and GANs.4:21–8:04 · Matt as informed peer 2/10 Technical Roots and Sci-Fi Vision of Holograms Matt prompts the founders to explain how academic computer graphics research translated into Synthesia and how the founding team met. Matthias and Victor share sci-fi inspirations like Star Trek holodecks and the academic papers that sparked the company.8:04–10:27 · Matt as informed peer 2/10 Synthesia Product Walkthrough and Corporate Learning Use Cases Matt asks for a walkthrough of the browser-based software product and core enterprise applications. Victor clarifies a common misconception, pointing out that Synthesia primarily replaces static text and training PDFs rather than replacing high-end camera productions.10:27–13:56 · Matt as informed peer 4/10 Scaling Enterprise Video, Customer Experience, and API Vision Matt highlights the strategic importance of providing the technology as an API. Victor elaborates on enterprise use cases, explaining how automated video allows banking data to be converted directly into personalized visual content.13:56–19:24 · Matt as informed peer 5/10 Digital Twin Avatars vs. Synthetic Humans Matt directly addresses the ethics of deepfakes, asking for clear definitions and safeguards against non-consensual visual cloning. Victor presents a structured response covering identity verification, public awareness campaigns, and media provenance architectures.19:24–24:16 · Matt as informed peer 3/10 Multilingual Video Generation and AI Model Synergies Matt relays an audience question about multilingual features and asks about the future roadmap of AI capabilities over the next 3 to 4 years. Matthias explains the synergy of converging AI models and foresees generating entire movies directly from text prompts.24:16–25:26 · Matt as informed peer 4/10 Q&A: Bespoke Avatars and Personalized Video Generation Matt brings in a audience question about personalized avatar creation and wraps up the interview by citing William Gibson regarding unevenly distributed technology. Victor explains the current process for custom avatars and the technical path toward single-image generation.0:22–4:21 · Guest teaching 4/10 Synthesia AI Avatar Product Presentation Video Matt provides a clear introductory overview combining computer vision and voice AI before asking the guest to explain underlying generative networks. Matthias details how computer graphics methods evolved from film industry stunt editing into deep learning and GANs.4:21–8:04 · Guest teaching 4/10 Technical Roots and Sci-Fi Vision of Holograms Matt prompts the founders to explain how academic computer graphics research translated into Synthesia and how the founding team met. Matthias and Victor share sci-fi inspirations like Star Trek holodecks and the academic papers that sparked the company.8:04–10:27 · Guest teaching 5/10 Synthesia Product Walkthrough and Corporate Learning Use Cases Matt asks for a walkthrough of the browser-based software product and core enterprise applications. Victor clarifies a common misconception, pointing out that Synthesia primarily replaces static text and training PDFs rather than replacing high-end camera productions.10:27–13:56 · Guest teaching 4/10 Scaling Enterprise Video, Customer Experience, and API Vision Matt highlights the strategic importance of providing the technology as an API. Victor elaborates on enterprise use cases, explaining how automated video allows banking data to be converted directly into personalized visual content.13:56–19:24 · Guest teaching 5/10 Digital Twin Avatars vs. Synthetic Humans Matt directly addresses the ethics of deepfakes, asking for clear definitions and safeguards against non-consensual visual cloning. Victor presents a structured response covering identity verification, public awareness campaigns, and media provenance architectures.19:24–24:16 · Guest teaching 5/10 Multilingual Video Generation and AI Model Synergies Matt relays an audience question about multilingual features and asks about the future roadmap of AI capabilities over the next 3 to 4 years. Matthias explains the synergy of converging AI models and foresees generating entire movies directly from text prompts.24:16–25:26 · Guest teaching 3/10 Q&A: Bespoke Avatars and Personalized Video Generation Matt brings in a audience question about personalized avatar creation and wraps up the interview by citing William Gibson regarding unevenly distributed technology. Victor explains the current process for custom avatars and the technical path toward single-image generation.0:22–4:21 · Guest disagreement 0/10 Synthesia AI Avatar Product Presentation Video Matt provides a clear introductory overview combining computer vision and voice AI before asking the guest to explain underlying generative networks. Matthias details how computer graphics methods evolved from film industry stunt editing into deep learning and GANs.4:21–8:04 · Guest disagreement 0/10 Technical Roots and Sci-Fi Vision of Holograms Matt prompts the founders to explain how academic computer graphics research translated into Synthesia and how the founding team met. Matthias and Victor share sci-fi inspirations like Star Trek holodecks and the academic papers that sparked the company.8:04–10:27 · Guest disagreement 1/10 Synthesia Product Walkthrough and Corporate Learning Use Cases Matt asks for a walkthrough of the browser-based software product and core enterprise applications. Victor clarifies a common misconception, pointing out that Synthesia primarily replaces static text and training PDFs rather than replacing high-end camera productions.10:27–13:56 · Guest disagreement 0/10 Scaling Enterprise Video, Customer Experience, and API Vision Matt highlights the strategic importance of providing the technology as an API. Victor elaborates on enterprise use cases, explaining how automated video allows banking data to be converted directly into personalized visual content.13:56–19:24 · Guest disagreement 1/10 Digital Twin Avatars vs. Synthetic Humans Matt directly addresses the ethics of deepfakes, asking for clear definitions and safeguards against non-consensual visual cloning. Victor presents a structured response covering identity verification, public awareness campaigns, and media provenance architectures.19:24–24:16 · Guest disagreement 0/10 Multilingual Video Generation and AI Model Synergies Matt relays an audience question about multilingual features and asks about the future roadmap of AI capabilities over the next 3 to 4 years. Matthias explains the synergy of converging AI models and foresees generating entire movies directly from text prompts.24:16–25:26 · Guest disagreement 0/10 Q&A: Bespoke Avatars and Personalized Video Generation Matt brings in a audience question about personalized avatar creation and wraps up the interview by citing William Gibson regarding unevenly distributed technology. Victor explains the current process for custom avatars and the technical path toward single-image generation.0:22–4:21 · Matt pushing back 0/10 Synthesia AI Avatar Product Presentation Video Matt provides a clear introductory overview combining computer vision and voice AI before asking the guest to explain underlying generative networks. Matthias details how computer graphics methods evolved from film industry stunt editing into deep learning and GANs.4:21–8:04 · Matt pushing back 0/10 Technical Roots and Sci-Fi Vision of Holograms Matt prompts the founders to explain how academic computer graphics research translated into Synthesia and how the founding team met. Matthias and Victor share sci-fi inspirations like Star Trek holodecks and the academic papers that sparked the company.8:04–10:27 · Matt pushing back 0/10 Synthesia Product Walkthrough and Corporate Learning Use Cases Matt asks for a walkthrough of the browser-based software product and core enterprise applications. Victor clarifies a common misconception, pointing out that Synthesia primarily replaces static text and training PDFs rather than replacing high-end camera productions.10:27–13:56 · Matt pushing back 0/10 Scaling Enterprise Video, Customer Experience, and API Vision Matt highlights the strategic importance of providing the technology as an API. Victor elaborates on enterprise use cases, explaining how automated video allows banking data to be converted directly into personalized visual content.13:56–19:24 · Matt pushing back 3/10 Digital Twin Avatars vs. Synthetic Humans Matt directly addresses the ethics of deepfakes, asking for clear definitions and safeguards against non-consensual visual cloning. Victor presents a structured response covering identity verification, public awareness campaigns, and media provenance architectures.19:24–24:16 · Matt pushing back 0/10 Multilingual Video Generation and AI Model Synergies Matt relays an audience question about multilingual features and asks about the future roadmap of AI capabilities over the next 3 to 4 years. Matthias explains the synergy of converging AI models and foresees generating entire movies directly from text prompts.24:16–25:26 · Matt pushing back 0/10 Q&A: Bespoke Avatars and Personalized Video Generation Matt brings in a audience question about personalized avatar creation and wraps up the interview by citing William Gibson regarding unevenly distributed technology. Victor explains the current process for custom avatars and the technical path toward single-image generation.

speaking balance: gold is Matt, purple is the guest (3 minute bins)

0:00 · Matt 29.2% · guest 70.8%0:00 · Matt 29.2% · guest 70.8%3:00 · Matt 23.4% · guest 76.6%3:00 · Matt 23.4% · guest 76.6%6:00 · Matt 10.3% · guest 89.7%6:00 · Matt 10.3% · guest 89.7%9:00 · Matt 14.4% · guest 85.6%9:00 · Matt 14.4% · guest 85.6%12:00 · Matt 14.2% · guest 85.8%12:00 · Matt 14.2% · guest 85.8%15:00 · Matt 20.8% · guest 79.2%15:00 · Matt 20.8% · guest 79.2%18:00 · Matt 13.2% · guest 86.8%18:00 · Matt 13.2% · guest 86.8%21:00 · Matt 1.5% · guest 98.5%21:00 · Matt 1.5% · guest 98.5%24:00 · Matt 57.6% · guest 42.4%24:00 · Matt 57.6% · guest 42.4%
Sharpest disagreement ▶ 9:15 Reframing Product Purpose

Victor gently counters a widespread misconception, pointing out that Synthesia is not intended as a camera replacement for film shoots, but as a replacement for long-form text and PDFs.

Hardest push from Matt ▶ 15:40 Addressing Deepfake Risks

Matt presses the guest on the controversial topic of deepfakes, defining consent violations clearly and demanding a breakdown of technical and ethical safeguards.

Biggest teaching moment ▶ 18:20 Media Provenance Framework

Victor educates the host and audience on structural media verification, detailing how a Shazam-like central database or blockchain provenance system can verify video authenticity better than simple detection tools.

Matt holds his own ▶ 12:43 Highlighting API Distribution Model

Matt demonstrates commercial sharpness by identifying the API release as the key architectural shift that transforms video from recorded content into scalable software.

the scores for every segment, with the reasoning behind each
ChapterTopicMatt as informed peerGuest teachingGuest disagreementMatt pushing backWhy
Synthesia AI Avatar Product Presentation Video 3400 Matt provides a clear introductory overview combining computer vision and voice AI before asking the guest to explain underlying generative networks. Matthias details how computer graphics methods evolved from film industry stunt editing into deep learning and GANs.
Technical Roots and Sci-Fi Vision of Holograms 2400 Matt prompts the founders to explain how academic computer graphics research translated into Synthesia and how the founding team met. Matthias and Victor share sci-fi inspirations like Star Trek holodecks and the academic papers that sparked the company.
Synthesia Product Walkthrough and Corporate Learning Use Cases 2510 Matt asks for a walkthrough of the browser-based software product and core enterprise applications. Victor clarifies a common misconception, pointing out that Synthesia primarily replaces static text and training PDFs rather than replacing high-end camera productions.
Scaling Enterprise Video, Customer Experience, and API Vision 4400 Matt highlights the strategic importance of providing the technology as an API. Victor elaborates on enterprise use cases, explaining how automated video allows banking data to be converted directly into personalized visual content.
Digital Twin Avatars vs. Synthetic Humans 5513 Matt directly addresses the ethics of deepfakes, asking for clear definitions and safeguards against non-consensual visual cloning. Victor presents a structured response covering identity verification, public awareness campaigns, and media provenance architectures.
Multilingual Video Generation and AI Model Synergies 3500 Matt relays an audience question about multilingual features and asks about the future roadmap of AI capabilities over the next 3 to 4 years. Matthias explains the synergy of converging AI models and foresees generating entire movies directly from text prompts.
Q&A: Bespoke Avatars and Personalized Video Generation 4300 Matt brings in a audience question about personalized avatar creation and wraps up the interview by citing William Gibson regarding unevenly distributed technology. Victor explains the current process for custom avatars and the technical path toward single-image generation.

Statements from this episode (18)

Assertion Not checkable as stated
Synthesia has generated over one million videos across 40 countries
“Since our launch, we have generated more than one million videos and served thousands of customers in more than 40 countries.”
Synthesia AI Avatar May 24, 2021 ▶ 1:15
Assertion Not checkable as stated
Niessner: Generative AI can create video virtually indistinguishable from reality
“Now there's a lot of new stuff coming that you can actually make not just images out of it, but you can actually create full videos out of it and can make these things look very, very realistic. You can create very high resolution videos and make them pretty m…”
Matthias Niessner May 24, 2021 ▶ 4:05
Disclosure
Synthesia aims to enable video creation without cameras, actors, or studios
“Creating technology that would make it easy to create video for everyone and without having to deal with cameras and actors and studio equipment every time you want to create a video.”
Victor Riparbelli May 24, 2021 ▶ 7:53
Insight
Riparbelli: Synthetic video replaces enterprise text, not traditional camera production
“Our platform is not really a replacement for traditional video production as you know it with like a camera, right? It's actually more replacement for text.”
Victor Riparbelli May 24, 2021 ▶ 9:40
Prediction Not checkable as stated
Riparbelli: Video and media production will shift from cameras to code
“So the big idea that Synthesia is built around, right, is that video production and media production in general is going to go from something that we record with cameras and microphone to something we code with computers.”
Victor Riparbelli May 24, 2021 ▶ 11:59
Disclosure
Synthesia to launch V1 of its video generation API
“The API part of it, of which we're launching our V-One of the platform in a couple of weeks, will enable a whole new space for opportunities, most of which we probably haven't really thought of yet, right?”
Victor Riparbelli May 24, 2021 ▶ 13:26
Prediction Not checkable as stated
Riparbelli: AI will make video websites as easy to build as text websites
“Moving forward, this will make it as easy to create a video-driven website as a text-driven website, and I think that's going to have major implications for how we kind of go about the user experience online.”
Victor Riparbelli May 24, 2021 ▶ 13:46
Assertion Not checkable as stated
Riparbelli: 80% of Synthesia's enterprise clients use custom avatars
“One which is already used by roughly 80% of our enterprise clients today, which is that you can create a real avatar of yourself.”
Victor Riparbelli May 24, 2021 ▶ 14:20
Prediction Not checkable as stated
Riparbelli: Everyone will eventually use a personal digital avatar for video calls
“We definitely believe that we're all going to have a digital representation of ourselves kind of an avatar of ourselves that we can use for creating video, maybe even Zoom calls, stuff in the future.”
Victor Riparbelli May 24, 2021 ▶ 14:39
Prediction Not checkable as stated
Riparbelli: Synthetic video technologies will definitely be used for bad
“And these technologies will be used for bad for sure.”
Victor Riparbelli May 24, 2021 ▶ 16:41
Insight
Riparbelli: Public exposure to synthetic media builds essential deepfake literacy
“Exposure of this type of media is, is the most important part of it, right? Once you start getting personalized birthday messages from David Beckham and Lionel Messi, for example, you know that's not real. That, that builds that sort of embedded sense of this …”
Victor Riparbelli May 24, 2021 ▶ 17:42
Disclosure
Synthesia shares data with major tech firms to build deepfake detectors
“Both Synthesia as a company and Matthias is working with a lot of the larger companies by sharing data and helping them with effort to build these sort of deep fake detection tools.”
Victor Riparbelli May 24, 2021 ▶ 18:05
Assertion Supported
Riparbelli: Synthesia supports 55 languages for synthetic video generation
“We support 55 languages right now”
Victor Riparbelli May 24, 2021 ▶ 19:39
Insight
Niessner: Using AI in a company is no longer a differentiator
“Like, you're not unique anymore when you're using AI in a company at this point, you actually Basically you have to use it just to compete with the massive scale of data and these kinds of things.”
Matthias Niessner May 24, 2021 ▶ 21:40
Prediction Not checkable as stated
Niessner: AI will eventually generate full blockbuster movies directly from books
“In the long run, what I see is you basically can generate a whole movie or something like this from just a book, right? So you kind of, you're helping Hollywood in a sense creating fully featured blockbuster films just by looking at some texts.”
Matthias Niessner May 24, 2021 ▶ 23:26
Assertion Not checkable as stated
Riparbelli: Synthesia has close to 200 custom avatars on its platform
“We already have, I think, close to 200 custom avatars on the platform so far”
Victor Riparbelli May 24, 2021 ▶ 24:49
Assertion Partly supported
Riparbelli: Creating a custom Synthesia avatar requires three to four minutes of footage
“And the onboarding process is roughly three to four minutes of footage.”
Victor Riparbelli May 24, 2021 ▶ 24:58
Disclosure
Riparbelli: Synthesia is working to generate custom video avatars from a single image
“We're working on getting that down to just a single image.”
Victor Riparbelli May 24, 2021 ▶ 25:01
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.