Video Is About to Stop Being One-Way (and That Changes Everything) | Victor Riparbelli, Synthesia
July 28, 202639:29

Video Is About to Stop Being One-Way (and That Changes Everything) | Victor Riparbelli, Synthesia

Every company is creating content that nobody reads, nobody watches, and nobody remembers, and the CEO of the AI platform that 90% of Fortune 100 companies use to fix that just explained what comes next.

In this episode, Craig Smith sits down with Victor Riparbelli, co-founder and CEO of Synthesia, to discuss the $4 billion company that is now redefining what video communication means for the enterprise.

The conversation opens with the founding insight that still drives the company: AI is going to drive the marginal cost of creating video to zero, which changes not just how content is produced but who can produce it and for whom. Victor describes how Synthesia found its first real market not in Hollywood - which rejected the technology as too low quality - but in corporate trainers and educators who were comparing it not to a film but to a 10-page PDF no one was reading. The most forward-looking section of the conversation covers Synthesia's next product: moving video from a one-way broadcast into a two-way interactive conversation, where an AI avatar can conduct a real-time sales demo, simulate a customer for sales training, draw graphs on screen to explain pricing, and score whether the person on the other side actually understood the content. Victor also makes a sharp prediction about where AI entertainment will actually emerge, not in cinemas or on Netflix, but from film students with laptops posting 17-minute short films on Instagram, the same way synthesizers didn't replace pianos but created entirely new genres of music.

Key Topics Covered:

● How Synthesia found its first real market: corporate trainers creating content nobody was reading, who compared AI video not to Hollywood but to a PDF, and found it vastly superior

● The transition from one-way video broadcast to two-way interactive avatar conversations, and what that means for sales demos, corporate training, and education

● Why Hollywood will be the last industry to adopt AI video, and why the first AI-generated entertainment will come from broke film students on Instagram, not studios

● Why AI content won't replace real video, it will become its own genre, the same way synthesizers didn't replace guitars but created electronic music

● How the CEO uses Claude daily for strategic thinking, playing devil's advocate, and replacing the long memo with a voice note

As AI video tools proliferate, this conversation offers one of the clearest frameworks for understanding where the technology is actually headed, not toward Hollywood, but toward transforming the way every company communicates internally and externally, with interactive AI avatars replacing the static website as the primary interface between a business and its customers.

Subscribe to Eye on A.I. for weekly conversations with the people building and deploying the future of AI.

Craig Smith on X: https://x.com/craigss

EYE On A.I. on X: https://x.com/EyeOn_AI

Connect with Victor Riparbelli

LinkedIn: https://uk.linkedin.com/in/victorriparbelli

[00:00:00] [SPEAKER_01] I think Hollywood is going to be the last industry to adopt any of these things.

[00:00:03] [SPEAKER_00] The idea of doing full-length films, has that gone by the wayside as you found your market niche? Or do you still imagine going that? I think what you will see is AI-generated entertainment. It's not in the cinemas, it's not on Netflix. And I kept asking why don't you put an avatar on this that would engage students more? And this is for K-12 because students are bored by text on a screen.

[00:00:31] [SPEAKER_01] They kept telling us the same thing. I know I'm creating all this content and no one is reading it. And even if they do read it, they don't remember anything else. I'm wasting my time. I need to create video and audio content. I think for kids, if you could go, you could personalize their education in two ways. Interest-based and with a celebrity or a person that they know. It would be amazing, right?

[00:00:49] [SPEAKER_00] Hey, Victor. It's great to meet you. As I mentioned a moment ago, I was an early Synthesia. I have trouble pronouncing the company name and have used it for years. But I'd like you to start by introducing yourself to listeners, give your background so far as it's relevant, and the story of how you came to start the company.

[00:01:18] [SPEAKER_01] Yeah, sure. So, I'm Victor and I'm one of the co-founders and the CEO of Syntheja today. And I grew up in Copenhagen in Denmark. And I think my path to what founded the company started in my early childhood years. I loved reading. I loved computers. In my late teenage years, I figured out that I could make some money off my interest in computers. So, I began making websites for local businesses, e-commerce stores, kind of small projects like that, where I could make pretty good money for like a 17-year-old.

[00:01:48] [SPEAKER_01] I was pretty good at it. Started running the marketing and kind of did it for some years, but eventually kind of capped out in how intellectually interesting it could be to help my local tennis store sell more tennis rackets. So, I went to the Danish startup ecosystem. I was there for three years. Found myself in like product and growth roles. I have a bachelor in computer science and business, but honestly, I was never really there. I was more interested in building companies.

[00:02:16] [SPEAKER_01] After having worked in the Danish ecosystem for some years, I knew I wanted to build my own company. I also knew that I wasn't super interested in building accounting tools and bookkeeping tools and like normal business process stuff. I've always loved science fiction. I love fringe ideas. I love the weird, wonderful corners of the internet and what's happening in the world. And I wanted to combine my interest in building products with that. So, I moved to London.

[00:02:41] [SPEAKER_01] I spent a year figuring out what company I wanted to build, met a guy called Matthias Liester, who was an associate professor at Stainford at the time. And he'd done a research paper called Face to Face, which is the first research paper demonstrating a neural network producing, you know, very photo realistic video. When we look back at it today, it's less impressive. The world has moved on quite a lot, but I saw this and I felt like there's something very magical about it.

[00:03:04] [SPEAKER_01] And as someone who loves creative things in my spare time, you know, I make music, dabbled in like Photoshop and a bit of video editing, 3D rendering. I saw the technology. I was like, this is going to change everything we know about how we create video and media in general, and also how we consume it eventually. So, I decided to start a company around it, which took a while to convince Matthias. I was 25 years old at the time. And I think he thought I was, you know, he wasn't very interested in the beginning.

[00:03:33] [SPEAKER_01] But with enough persistence, we kind of got there in the end. And we ended up founding a company based on basically two big ideas. The first one being that AI is increasingly going to be able to generate content, not just analyze patterns and data. And that's going to change the equation around how we create content, right? It's going to drive the marginal cost of creating content to zero. So, the cost in dollars of creating content is going to be very low. The cost in time is going to be very low. And the time and skill required to create video is going to be very low.

[00:04:02] [SPEAKER_01] And I think that's the bits up until now. I think that's the market win right now where anyone can go on the internet. They can prompt something in a few minutes or they can produce something that's very, very high quality. And that changes, like, the nature of video production, right?

[00:04:16] [SPEAKER_00] When you reached out to him, first of all, how did you happen to read the paper? Because this is an experience I think anyone who follows Archive, you know, the repository of preprint papers has experienced. You see these great ideas. You think, wow, you could build a company around that. A lot of times that research just fades into the woodwork. How did you happen to read the paper?

[00:04:44] [SPEAKER_00] And how did you reach out to him?

[00:04:47] [SPEAKER_01] So, when I moved to London in 2016, VR and AR was like the hype cycle at the time. The Oculus headset that just came out. And I was very frailed by all this tech. And I actually thought I was going to start a VR company. So, I spent a lot of time just, like, meeting researchers, meeting companies. I worked as kind of like a freelance consultant to some big companies and what they should do with VR and AR. And I kind of pretty quickly came to the conclusion that outside of the headset being very clunky as a barrier for VR adoption.

[00:05:17] [SPEAKER_01] The other thing that was a big problem was that creating content for VR is very expensive, you know. Because basically what you're doing is you're building a computer game. And that comes at a very high cost even compared to video production. So, I kind of started to have this interest in how can we make content creation easier. That took me down a path of things like volumetric capture, which is you have, like, a lot of cameras and an array around you. You go in and you do a performance and you get a 3D asset out of that. You could port into VR. That was interesting. I was part of setting up a studio for this in London.

[00:05:48] [SPEAKER_01] And a whole bunch of other interesting ideas of, like, how content creation could be made easier. So, I was following that space. And Matthias Leesner had a lot of papers in this space around VR. But he also had this paper coming out called Face to Face. And when I saw that, I think I had this eureka moment where, you know, instead of trying to do this with VR, where we're both trying to bring a new type of media to market, which is hard and a clunky headset. And we're trying to make it easier to create content. That's kind of a double whammy of very hard problems.

[00:06:17] [SPEAKER_01] If this works with video, that's interesting, right? Because video, the infrastructure of video is already in place. We're all spending time on YouTube. Like, video is, you know, you could just take an AI video and replace it with a normal video. And so, the distribution problem wasn't really there. And I think that's where I kind of started to feel like this is actually, like, a very interesting business idea. Because if we can solve this for video, we, you know, we can distribute it to the world pretty quickly.

[00:06:42] [SPEAKER_00] So, you started the company then. And that's very funny. I was about the VR stuff. I was running the New York Times operations in China at the time. And when VR first came out, and there was a company, I'm sure they've long disappeared, that came through town. And they were showing us their cameras. And I got very excited.

[00:07:08] [SPEAKER_00] And I ended up partnering with Chinese media to do a big conference on VR. I thought, boy, that's the future. But as it turns out, it wasn't. And so, when did you start Synthesia? And tell me about the name. I always want to say Anesthesia.

[00:07:33] [SPEAKER_01] So, we started the company in 2017. But there was a lead up to that of, you know, trying to build a prototype and, you know, get the thesis down. And all that, trying to raise money, which was very, very, very difficult. Because this was a very non-consensus bet at the time. We were basically telling people, in 10 years, you can make a Hollywood film with nothing else but your imagination. And at that point in time, like, that was certainly like a sci-fi dream more than it was something that was, like, tangible to investors.

[00:08:02] [SPEAKER_01] And every investor just wanted to invest in fintech at the time. But in 2017, end of October 2017, raised the first funding round, got a million dollars from Mark Cuban. Who was a big believer and has been since day one.

[00:08:16] [SPEAKER_00] That process. So, how did you get to Mark Cuban? Just a cold outreach?

[00:08:21] [SPEAKER_01] Yeah, so we had a long list of investors that we tried to speak to. And we weren't, you know, very lucky. And we had a list over, like, business angels, you know, wealthy people that fit well in their kind of media and technology, Venn diagram. And Mark was one of them. So, we found his email online, actually, by downloading a hack. So, only got hacked a few years prior. And there was a lot of emails in that hack. We found his email. We sent him a cold email. And he responded back within four minutes saying, I think this is really interesting. Came back with some more questions.

[00:08:50] [SPEAKER_01] And then we had a 14-hour email conversation with him. Mark doesn't do any video calls or normal calls. He only does email. And so, 14 hours later, he'd asked a lot of great questions. And he said, I'll do a million dollars if the diligence checks out. And I think one thing that was very interesting about this process was Mark was aware of the research paper. He had implemented it himself at home. And he 100% believed in the vision around you're going to be able to make a Hollywood film from your laptop in 10.

[00:09:19] [SPEAKER_01] We didn't have to discuss that. He was just interested in how you're going to build a company around it. And what are you going to do until the technology actually works? Which is a very different conversation than every other investor we spoke to where we had to start at like, hey, this thing is going to be possible in 10 years. And they just, if they didn't, you know, Mark shared the same belief around us. So, he was evaluating us as a team. Whereas everyone else, we had to try and convince them about this like worldview. And that's very hard.

[00:09:44] [SPEAKER_01] So, we got turned down by I think almost 100 investors before Mark actually invested. And that became the, yeah, the start of the company.

[00:09:52] [SPEAKER_00] Yeah. And so, you started with a million, I mean, after your angels. That was a seed investment. You just completed Series E, I think 200 million. And you're now, your valuation's like 4 billion. Yeah, that's incredible. I mean, that's, I mean, all of the companies are growing extremely fast.

[00:10:19] [SPEAKER_00] So, the initial product was around avatars for advertising or different use cases. I'm particularly interested. You know, I was involved early on sort of as a media advisor with a company, AI education company.

[00:10:45] [SPEAKER_00] And I kept asking, why don't you put an avatar on this that would engage students more? Because, and this is for K through 12. Because students are bored by text on screen. And you've moved into education or you've been in education, corporate training primarily. Is that right?

[00:11:14] [SPEAKER_01] Yeah, that's right. So, what we discovered was that the first use case where this kind of technology was actually valuable and not just a cool demo or cool PR stunt was in explanatory or educational video. Especially the enterprise, which is what we've been focused mostly on up until now, right? Because we first started to try and create, you know, technology for Hollywood studios, video production agencies, the most obvious people to use this kind of technology.

[00:11:43] [SPEAKER_01] But what we bumped into was that the quality requirements was extremely high for these people. And the willingness to experiment was pretty low because they're already creating lots of great content, you know? They don't really need an AI solution. They're already making fantastic videos. But then we discovered there was this other group of people and this group of people were people who were desperate to make video content. Because they're educating someone on something. And the only way they could do it was to send them long pages of text or maybe a slide deck. And they kept telling us the same thing.

[00:12:12] [SPEAKER_01] I know I'm creating all this content and no one is reading it. And even if they do read it, they don't remember anything else. I'm wasting my time. I need to create video and audio content. That's how everybody, especially the younger generation, wants to consume their content, right? But I don't have a way of doing it. And when we then came up with the first iteration of our technology and our product, what we discovered was that for these people, when they evaluated the quality of the AI video, they would hold it up against a 10-page PDF document. Whereas in Hollywood, right, they would hold it up against a real Hollywood film that say the quality is terrible.

[00:12:42] [SPEAKER_01] So that was when we figured out that here's a really interesting first use case. And then we've evolved the company since then to still focus on educational, explanatory video and enterprise. We haven't gone... We work with a bunch of universities and schools and so on who also use the product for educational purposes. And I think we're going to evolve into doing more and more of that over the years. We're just now launching our second big product, which is evolving video from being a one-way broadcast, which is what a video is today, right?

[00:13:12] [SPEAKER_01] Like you make one video and then everybody watches the same thing. With this, it's a two-way conversation. So you can talk to the video. You can talk to the avatar. In a training or educational context, you know, if you're training your sales team, for example, then today what people do is that they create a video. The sales team sits down and watches the video, and that's better than text. What you can do now is you can insert agents into the video because that could, for example, be to watch the video, and then the agent pretends to be a customer for 15 minutes.

[00:13:40] [SPEAKER_01] And you have to answer questions, overcome objections, build empathy, set up next steps, and basically practice like you would with a real customer. But you're doing it in a setting with an AI agent, right? And the reason this is very powerful is because we learn so much better by practicing and saying things out loud. And we can also, through technology, we can actually score your eyes. We can actually give whoever created the video a signal that you've actually understood the content. You haven't just like passively sat through it.

[00:14:09] [SPEAKER_00] Has anyone taken that further and instead of just training, used it as a salesperson on a Zoom call?

[00:14:19] [SPEAKER_01] I think all that stuff is going to happen in the next 12 months where people are going to be using these things, like do more things with them. I think the consideration that's important in the different scenarios that you use it is that the quality really matters a lot, right? So if you want to put it in front of a customer, it has to be very, very high quality with the latency. It has to feel like talking to a real person.

[00:14:44] [SPEAKER_01] Whereas you're going to get away with a bit more for internal use cases where it's like the stakes are a little bit different. But I absolutely think that in the future, the way you're going to interact with a company is going to be probably less like a website, which is basically a brochure. And more like you go in and you open up a kind of a Zoom call with an avatar. The avatar can, you know, pull up content, share their screen, and you basically go through it more like a demo with a real salesperson than, you know, go to a website and like click your own way around it.

[00:15:13] [SPEAKER_00] Training that you're talking about. The avatar is the text or the content that they're speaking is all written. It's not live, right? Do you have, have you looked at doing, I mean, you were talking about having a salesperson trained. Can you do live interaction?

[00:15:36] [SPEAKER_01] So this is live. Yeah. So this is a live avatar you're talking to.

[00:15:38] [SPEAKER_00] Right. So how do you, what's the, is it just a large language model that you plug in on the back that's producing the text that the avatar is speaking?

[00:15:50] [SPEAKER_01] Yeah. So, I mean, for most use cases, it's going to be an LLM that's like driving the avatar. It can be whatever you want it to be. If you want it to be an API that feeds it information without an LLM, you can also do that. Basically what the system does, right? It takes a string of text and then it like generates in real time the avatar kind of saying whatever you want it to say.

[00:16:09] [SPEAKER_00] Right. And is that built into the product or does someone have to do some engineering on the back end?

[00:16:17] [SPEAKER_01] So we're offering it for the first use cases, what we call skills, which is training. It's an intern product that you could go in and we can create, you know, you can set up your own scoring system for how you evaluate the people who talk to it. You can change the prompt to be whatever it is you want it to be. We're also going to enable you to use a low level API. If you just want the avatar in real time and you don't want the rest of like our product around it. So both things are going to be possible. And in the future, we're going to launch more than just the skills, like the training use case.

[00:16:47] [SPEAKER_01] Maybe we'll launch the sales use case at one point. Recruiting is another one that's obvious. And we'll kind of go solution by solution.

[00:16:53] [SPEAKER_00] Yeah. But the live interaction, that's already available in the product?

[00:16:59] [SPEAKER_01] Not yet.

[00:17:00] [SPEAKER_00] It's in beta right now with a bunch of customers, but it's launched within the next couple of months. Okay. The other thing that I'm curious about, I have a son who's an actor who's very much against all of this. And I've been telling him, hey, license your likeness. I mean, you guys do license actors' likenesses for avatars. I don't think you license celebrity likenesses. Yeah.

[00:17:31] [SPEAKER_00] Have you thought about doing that?

[00:17:34] [SPEAKER_01] Yeah. It's an idea that comes up like pretty often. And I think there's a bunch of companies that have tried to do it. I think the best proxy is something like Cameo. If you remember that app. I think the company doesn't really exist anymore. But basically, you could go in, right? And you could buy these like clips and a real celebrity would take up their phone and they would do some of that. I think there is probably a business to be built around that. But it's not really our focus because we're much more like a corporate B2B platform. But I think something like that probably can work.

[00:18:03] [SPEAKER_01] And I think where it's interesting is probably the scale you can get. So if you have an e-commerce store and you want every customer to receive a personalized video from a celebrity, for example, you could do that. And maybe you pay the celebrity $1 for every video, right? So I think it's going to happen. But my sense with this stuff is always that until the technology is truly so good that you cannot tell the difference, it's not good enough.

[00:18:28] [SPEAKER_01] Because it's going to be – it's not going to feel the way that you want it to feel. Right, right. Is that right?

[00:18:35] [SPEAKER_00] But I think we're approaching that now. Because, again, on the education, the idea of having an avatar is the interface between the system and the student. If you had – you know, I'm old, so I think Tom Cruise is – I don't even know who a young sixth grader would be excited by.

[00:18:59] [SPEAKER_00] But if you had the avatar of a celebrity that they were attracted to, that would be a wonderful way to deliver content. So you think that's coming, that's possible. Does – is the uncanny valley, you know, between what you can produce now and what you think would be necessary, how wide is that or how deep is that currently?

[00:19:28] [SPEAKER_00] Because when I look at your avatars, they seem very realistic to me.

[00:19:37] [SPEAKER_01] Yeah. No, I think it's almost there now. And I think some of this stuff probably will happen in the next six to 12 months. Because the technologies by this point is like so mature that they really are of very high quality, right? And I totally agree, I think. For kids, if you could go – if you could personalize their education in two ways, interest-based and with a celebrity or a person that they know, it would be amazing, right?

[00:20:00] [SPEAKER_01] If you just really love soccer or music or like whatever, Ninja Turtles, and you could have all your math problems be framed as Ninja Turtle problems and delivered by your favorite celebrity, that's amazing, right? And I think we are getting pretty close to that being a reality where if you want David Beckham to teach your kids about mathematics, you probably can do that very soon.

[00:20:23] [SPEAKER_00] Yeah. And how does Synthesia – am I pronouncing it right? Well, first of all, where did you get the name? I know Synthesis and Asia. I don't know where they –

[00:20:36] [SPEAKER_01] Yeah. So it's – actually, it's a codename for like that started off and we just never changed it. But it comes from – I love music. And when I'm in my free time, I make music. Especially electronic music is my big passion. And electronic music is kind of interesting because it's a product of, you know, way back in the 70s and 80s, people were trying to create digital instruments, like drum machines and synthesizers. And a synthesizer's job is basically to imitate a piano or guitar or something.

[00:21:06] [SPEAKER_01] That's why they were created in the first place. Now, what happened back then was that the quality of the drum machines and the synthesizers was not very good. It didn't at all sound like a piano or guitar. So it wasn't used for what people thought it was going to be used for. But then there was a lot of people who took it and created new genres of music with it where you wanted to sound otherworldly and digital, right? And so if you look at music with sampling and synthesizers, you can basically make anything you want. And so that's kind of like where the name comes from because basically that's what we're doing, right?

[00:21:34] [SPEAKER_01] We're building synthesizers for video.

[00:21:37] [SPEAKER_00] Yeah. So how do you work with companies, say a corporate training company? Do you have an API that they plug into their system or do they come in and collaborate in a studio? How does that work?

[00:21:55] [SPEAKER_01] Yeah. So it's a web app. And I think the best way of thinking about the product is PowerPoint 2.0. It's used generally by people who use PowerPoint at Enterprise. So it's not video experts that use it. It's like your average office worker. And they log into the platform and making a video is kind of like making a slide deck. Go in, you create an avatar of yourself. You choose from the library. You type the scripts that you want. You put the visuals together. We have agents that can help you with all this stuff as well. And then you start to make videos.

[00:22:22] [SPEAKER_01] So it is really like a content creation tool, like making slide decks using Canva, Photoshop, those kind of products.

[00:22:33] [SPEAKER_00] On the interactivity, do you have any way of reading the attention of the person on the other side?

[00:22:49] [SPEAKER_01] So not yet, but we are building that because what you want is that if you do one of these like simulation calls, for example, you want the avatar to be able to show that it's bored or that it's unengaged or that it's very excited or something like that. Right. As a signal to like how you're doing yourself. It's a harder research problem, of course. But I think it's something we should, you know, we will be able to do within the like foreseeable future.

[00:23:16] [SPEAKER_00] There's competition out there, obviously. Where do you see this industry growing and where do you put yourselves in the competitive landscape?

[00:23:58] [SPEAKER_01] I think a lot of the very consumer level creation, I think it's going to end up happening inside of like Meta's apps, for example, inside of TikTok. I think they're going to launch all these AI creator tools inside the apps themselves. And then I think the big other push, right, is this like moving away from just video and to video that's interactive, two-way dynamic. As soon as you move into that space, the universe becomes a lot bigger, right?

[00:24:23] [SPEAKER_01] Like we'll probably focus more again on like business use cases, training use cases, HR use cases, which is more like kind of the world that Synthesia generally sits in. And there will probably be other companies who go and make amazing consumer level experiences of like talking to celebrities or, you know, almost like game show kind of that. I mean, there'll be a lot of like new experiences that's brought to life.

[00:24:45] [SPEAKER_01] So I think we're about to see the industry up until now it's been everyone operating pretty close to what each one does. But I just think the next couple of years we'll see a much bigger diversity in what the companies are doing and what they choose to focus on.

[00:24:59] [SPEAKER_00] Yeah, I mean, there's also a kind of a convergence between voice cloning and image cloning. Yeah, for sure. Do you do the voice cloning yourselves or do you have a third party?

[00:25:16] [SPEAKER_01] Yeah, so we train voice models ourselves. We train video models ourselves and real-time video models ourselves. We also use third party providers for language coverage. And I'm not like, you know, extremely beholden to the fact that we need to do everything in-house. But what we've seen is that across some of these modalities like voice, for example, our customers have some needs which are not available in the market. They really want to preserve their accent, for example. A lot of the existing solutions are not very good at that. So we built those ourselves. Yeah.

[00:25:44] [SPEAKER_00] This idea, you started out with this vision of producing a Hollywood movie and then, I mean, Sora has been retired, but there have been all these video generation platforms that have come up. How different do you consider yourself from the video generation platforms?

[00:26:08] [SPEAKER_01] I think there is going to be a foundation model layer and then there's going to be an application layer and then it's going to be some mix in between, right? And I think what we're very clearly seeing now is that the players that have the most compute and the most talent, like Google but VO3, can produce amazing models. Frankly, a world model like that is very hard for me to compete on, even though I've raised a lot of money and we're a pretty big company. I don't want to compete on like very general models that can do anything you want, which is basically what these companies do, right?

[00:26:38] [SPEAKER_01] The models we train is very specifically for avatars, people talking to the camera. They need to be able to produce clips of arbitrary length, not just eight minutes. We need to be able to have consistent identity. The avatars can't like change looks every time you make a new one with it and so on and so forth. So we partner with these big companies. They're a big part of our product. We offer many models inside their platform, both just as a standalone, but also if you want to have your avatar driving a car, for example, we can tap into these models and build a workflow around it.

[00:27:08] [SPEAKER_01] So I don't see them as competition. I see them more as partners. And my guess is that that's going to continue and they're going to be essentially like one of the primes that we use to build our product.

[00:27:18] [SPEAKER_00] Yeah. The idea of doing full length films, has that gone by the wayside as you've found your market niche or do you still imagine?

[00:27:31] [SPEAKER_01] No, I think someone else is going to do that better than we are. It's not the company we are. It's not the DNA. On a personal level, I love this stuff, but I don't think we're the right ones to do that. I also don't think it's going to be 90 minute feature films that's going to be made with AI. I think Hollywood is going to be the last industry to adopt any of these things. And I think what you will see AI generated entertainment is not in the cinemas. It's not on Netflix. It's going to be on TikTok and Instagram.

[00:27:58] [SPEAKER_01] It's going to be a film student with no money, but a laptop who's going to make something amazing, a short film that's 17 minutes long and posted on Instagram. It's going to be very different. And so in general, my confidence on someone saying I'm going to go and make technology for Hollywood to make Hollywood films. I'm very skeptical that that's going to happen anytime soon.

[00:28:17] [SPEAKER_00] Yeah. I mean, that's interesting. I spent a lot of my life in China and in China, it's starting here. They have these mini series of like a minute per episode and they become very popular. People consume them on the subway. Yeah.

[00:28:36] [SPEAKER_01] I actually watched a documentary about this phenomena like literally two days ago. Yeah. And I think exactly that's the kind of stuff that's going to happen, right? Because these are basically like factories where they create these like short clips that are 20 hour sets. And then you have like one full fledged content piece. And I think we're going to see the same thing with AI. Maybe it's this kind of content they'll create. People will build, especially like in that part of the world, things like live streaming is incredibly popular. You'll probably see the news there.

[00:29:06] [SPEAKER_01] I think going back to the idea of synthesizing in music, right? Like that created a new genre of music more than anything. And we still today use pianos and guitars, right? And I also think we're still going to be recording things with cameras. AI content is just going to be its own genre that's going to look fundamentally different.

[00:29:23] [SPEAKER_00] Yeah. Yeah. And you see AI content being mixed with live content or I don't know what you call non-synthetic content these days. For sure.

[00:29:36] [SPEAKER_01] I think, I mean, a lot of this stuff ultimately is visual effects, right? And visual effects have been around for a very long time. The change is just that now it's, you know, it's very, very realistic. And it's very, very easy to use. Whereas, you know, even just five years ago, if you wanted to create a scene with visual effects of like you floating in space, that would have been very expensive. It would take a lot of people. You had to be in a green screen somewhere and it'd be like a very involved project.

[00:30:02] [SPEAKER_01] And now you can go in and you can, in literally five minutes, we can create a clip of you floating in space. So I think we will see a lot of this stuff just being picked up as just a faster, easier way to produce visual effects. And I also think when I say Hollywood isn't going to be the last ones to adopt this, I mean in like fully adopting it. I'm sure that the way Hollywood is going to be doing next couple of years is instead of using old school green screen technology, they'll begin to dabble in these things.

[00:30:28] [SPEAKER_01] Instead of a lot of these like traditional workflows will probably slowly be replaced, but it will still predominantly be like real video with added visual effects. As opposed to the 100% AI generated content, which we see a lot of on socials right now.

[00:30:43] [SPEAKER_00] Yeah. Are you still following research? Is that feeding into the product at all? Because there is a lot of research going on with agentic AI and things like that.

[00:30:59] [SPEAKER_01] I mean, for sure. Like, I mean, we spend a lot of money on training models. We have a big R&D team that does the video of voice and audio models. And I mean, I wish I had more time to like really closely following everything that's going on. But I think the explosion of things you have to keep up with at 2026 around AI means that I have a great team that, you know, that is mostly concerned with like the specifics of the models. Yeah.

[00:31:26] [SPEAKER_00] I want to ask where you guys are going next. But before I do that, I've been asking people with the agentic ecosystem developing.

[00:31:38] [SPEAKER_00] There's been a lot of talk of having more precisely trained CEO co-pilots so that the founder or the leader of a company can talk to an AI chatbot or LLM or something.

[00:32:01] [SPEAKER_00] To think through how the company should develop, where the opportunities are, what the economics of various opportunities are. Do you do any of that?

[00:32:13] [SPEAKER_01] Yeah. I mean, I spent a lot of time on like feeding the models that I use. I just switched to Claude to understand like how I think. And I think I'm trying to figure out like what the best way of implementing that is. I think having something that people can ask for like my opinion is probably interesting. But I'm also want to make sure that it responds in the right way, right?

[00:32:39] [SPEAKER_01] So I think everyone, it feels like everyone's like trying to figure out like what do you do with all these models that are very smart? They can understand how you think, how do you operationalize it in the best way? I don't think I have to answer yet, but I definitely like every strategic question I try and answer. I use LLMs, right? They're really good for like, you know, playing devil's advocate to your thinking. They're really good for instead of me having to sit down for half an hour and write a long memo,

[00:33:09] [SPEAKER_01] I can just like voice note it and it'll understand, have a lot of context around how I think generally. So, yeah, I think it's going to be, it's going to play a big role in how we operate companies in the future.

[00:33:19] [SPEAKER_00] Yeah. And you're using, you switched to Claude recently. Is that Claude co-work or?

[00:33:26] [SPEAKER_01] Right now it's just a vanilla Claude. Yeah. The co-work. But, yeah.

[00:33:32] [SPEAKER_00] Yeah. I mean, it's a fascinating idea that as a founder, you know, you generally, hopefully you surround yourself with some very smart, like-minded people. But as a corporation gets larger, there's a problem of yes-men sort of reinforcing your own biases.

[00:33:53] [SPEAKER_00] And to have an extremely intelligent AI that you can bounce ideas off or brainstorm with. For example, are you guys looking into acquisitions? You have all this money now. How are you expecting to grow from here? Is that going into research, product development, acquisitions or what?

[00:34:20] [SPEAKER_01] I mean, I think it's all of them, right? Like we evaluate all those things on a consistent basis. Like I think this year we're going to grow headcount by 70%. We're opening new offices where, yeah. I think there's a part of this which is just like scaling our core business. And then there's like bringing new products to market. Those are two very different disciplines, right? And they require like different things. We haven't made any acquisitions yet. But, I mean, who knows what the future will bring.

[00:34:49] [SPEAKER_00] Yeah. And then within the workforce, are you implementing agentic workflows or agentic systems to take up a lot of the work?

[00:35:03] [SPEAKER_01] Yeah, everywhere. Yeah, everywhere. Like, I mean, you're very gifted with a very AI-native workforce. And people are finding all sorts of amazing ways to use these tools. From the smallest like minuscule problems all the way to dealing with things like content moderation, for example, for us, right? Which is an important piece of the product. So I love seeing how people are just picking up these tools and automating, you know, business processes that just save them time so they can be more productive and other things.

[00:35:31] [SPEAKER_00] Yeah. So where is Synthesia going? I mean, what's next on the roadmap?

[00:35:42] [SPEAKER_01] What we're trying right now is invent a new media format that, you know, is proprietary to us. And it's going to be all about interacting with agents in a visual manner, right? So it's going to be an avatar that you talk to. It's going to be graphics that's drawn on the screen in real time. It's going to be agents that can take action for you, that can help you do work better. And that's an audacious task. There's, you know, lots of things to do. We're doing it from the position of having a lot of amazing customers. We work with more than 90% of the Fortune 100.

[00:36:13] [SPEAKER_01] And we're working a bunch of deployments for them. And I'd love to see, like, you know, in two or three years, we've really managed to transform how companies communicate. And that the percentage of, like, one-way broadcast communication with two-way interactive communication will have taken a big swing towards the two-way interactive communication, which I think for, like, a majority of use cases is a better way of communicating with people. Yeah.

[00:36:37] [SPEAKER_00] Describe that in a little more detail. I mean, I was asking about, you know, having live interactions with an LLM, with an avatar as the interface. What's the architecture that you're looking at? Is it as simple as that?

[00:36:56] [SPEAKER_00] Or is there, we were talking about the avatar being able to read the emotions on the person it's interacting with. Is, you know, is that one of the directions? What are the things you're working on?

[00:37:12] [SPEAKER_01] I mean, that certainly is an important part. There's a lot of things that goes into this, right? And I think it's also why, for me, building a great company, in our case, is about combining a lot of different technology and a lot of different models into great workflows, ultimately. Right? More than it's about the individual technology. That is a very important part of it. Real-time avatars is important. Being able to draw motion graphics on the screen in real-time is very important.

[00:37:36] [SPEAKER_01] And we have an editor today where you put together your video or your experience, and that's evolving into becoming, instead of you just creating a scripted video, it's evolving into how do you want to structure this conversation with someone, right? So if it's the sales example you used yourself earlier, if I go on a website and I want a demo of the product, how do you want that demo to be conducted? You know, instead of just showing people a video of what the product looks like, it can be like, hey, what would you like to see? What are you interested in? Let's have that conversation.

[00:38:04] [SPEAKER_01] And then maybe sometimes the agent will fetch a video from over here that shows this in detail, a scripted video. Other times it may just, like, draw a graph on the screen to explain how the pricing works. All this is around, like, visual communication, right? And that's both the avatar and it's, like, all the stuff that happens on the canvas.

[00:38:23] [SPEAKER_00] And how far are you from having that general release for that? So general availability?

[00:38:31] [SPEAKER_01] So it's live with a bunch of, like, big customers today who are using it, and it's going to hit GA probably in the next, like, couple of months. You'll be able to have, like, the real-time avatar specifically for the training product to begin with. Yeah.

[00:38:44] [SPEAKER_00] Okay. Well, I'm running out of questions. Is there something that I haven't talked about that's top of mind? I think we got pretty fine-wired. Yeah. Are you talking to any education companies?

[00:39:00] [SPEAKER_01] We have. We do work with, like, universities and primary schools and that kind of thing.

[00:39:05] [SPEAKER_00] But no one building what I envision an AI? No, not yet. Yeah.

[00:39:11] [SPEAKER_01] I think it's going to happen, too. I'm actually going to ASU GSV, which is a conference in San Diego. Yeah, yeah. So I'm going there next week and hopefully meet a bunch of interesting customers there.

[00:39:20] [SPEAKER_00] Okay.

[00:39:23] Bye.