Skip to content
The Twenty Minute VCThe Twenty Minute VC

How to Build Your Own Data Center & Why Every Startup Should Do It

Cliff Weitzman is the co-founder and CEO of Speechify, the world's leading AI voice and text-to-speech platform, used by more than 60 million people globally. Diagnosed with dyslexia as a child, Cliff first built Speechify at Brown University to help him consume written material through audio. ----------------------------------------------- Timestamps: 00:00 - Intro 00:55 - Why Speechify Is Spending Tens of Millions on Its Own GPUs 07:12 - Why Buying GPUs Can Be Better Than Renting Them 12:28 - Are NVIDIA’s Circular Economy Fears Overblown? 13:49 - What No One Understands About Buying AI Chips 18:26 - Why Owning Compute Can Become a Competitive Advantage 20:55 - How Big Can the AI Data Market Become? 22:59 - How ElevenLabs Leapfrogged Speechify 25:23 - Is Speechify Making a Mistake Going Into B2B? 30:17 - Why Every Winning AI Company Becomes a Compound Startup 31:10 - Is This the Hardest Time Ever for Startups to Hire Great Talent? 37:26 - How AI Is Completely Changing Software Engineering 39:18 - How Should Companies Think About AI Token Spend? 45:29 - How Your Hiring Process Needs to Change in the AI Era 47:03 - Is Voice AI Becoming Completely Commoditized? 50:18 - Is Customer Support the Wrong AI Market to Bet On? 53:34 - Sierra vs ElevenLabs: Who Becomes Bigger? 55:18 - Quick-Fire Round ----------------------------------------------- Subscribe on Spotify: https://open.spotify.com/show/3j2KMcZTtgTNBKwtZBMHvl?si=85bc9196860e4466 Subscribe on Apple Podcasts: https://podcasts.apple.com/us/podcast/the-twenty-minute-vc-20vc-venture-capital-startup/id958230465 Follow Harry Stebbings on X: https://twitter.com/HarryStebbings Follow Cliff Weitzman on X: https://twitter.com/cliffweitzman Follow 20VC on Instagram: https://www.instagram.com/20vchq Follow 20VC on TikTok: https://www.tiktok.com/@20vc_tok Visit our Website: https://www.20vc.com Subscribe to our Newsletter: https://www.thetwentyminutevc.com/contact ----------------------------------------------- Legal Disclaimer: The content of this podcast is for informational and entertainment purposes only and does not constitute financial or investment advice. Any discussion of stocks, public markets, or investment strategies reflects the personal opinions of the speakers and should not be relied upon when making investment decisions. Figures, valuations, and financial data referenced may be estimates or subject to error. Always consult a qualified financial adviser before making any investment decision. The views expressed are those of the individual speakers and do not represent the views of 20VC or its affiliates. ----------------------------------------------- #20vc #harrystebbings #cliffweitzman #ceo #speechify #ai #voiceai #startup #elevenlabs #sierra

Cliff WeitzmanguestHarry Stebbingshost
Sep 5, 20261h 4mWatch on YouTube ↗

EVERY SPOKEN WORD

  1. 0:000:55

    Intro

    1. CW

      It was the biggest strategic mistake I made in the history of Speechify.

    2. HS

      How do you reflect on that?

    3. CW

      So-

    4. HS

      Today is a real freaking discussion. Cliff Weitzman, founder and CEO of Speechify, one of the fastest-growing text-to-speech startups in the world, on the show.

    5. CW

      The best way to lose is not to be in the race. Be in the race. You don't want to be a fat manager who's like a general sitting in the back saying, "Take that hill." You want to be the warrior who runs up with their sword and engages the enemy first.

    6. HS

      [claps hands] Ready to go? [upbeat music] Cliff, it is so good to have you back in the studio, dude. I, I was looking forward to this one, 'cause when I was writing it up, it's a very different thread of conversation to how I'd normally go. And so thank you so much for joining me again today, dude.

    7. CW

      My pleasure. Glad to be here as always.

    8. HS

      Now, I wanted to start with you're spending tens of millions of dollars on NVIDIA GPUs,

  2. 0:557:12

    Why Speechify Is Spending Tens of Millions on Its Own GPUs

    1. HS

      and you're paying an additional $100,000 per GPU to receive them four months early. Why? Like, what do you know that the market doesn't know?

    2. CW

      So in 2022, we bought a huge rack of GPUs from NVIDIA, and the reason we bought them is for training, right? We have a bunch of models. The newest Speechify Simba 3.2 model is ranked number one in the world for quality, um, above all the frontier labs, 10X more affordable than stuff like ElevenLabs. And we used to rent GPUs, and we found that engineers at Speechify would be parsimonious with how they use the GPUs 'cause they were like, "Oh my God. I'm costing the company tens of thousands of dollars. Like, I don't wanna do that." And the analogy my brother and I came up with is imagine you're Michael Jordan, and you wanna be in the NBA. It's the only thing you care about, and you need to pay $20 an hour just to train in a basketball center. Well, that sucks. You want one that you can go to whenever you want to. In fact, you want a hoop in your house. And so our initial idea was we want a hoop in our house, and so we bought a bunch of our own GPUs, and that deal ended up being really good for us, and we ended up training really good models. So with time, we invested more and more and more and more. So that's the first part. The second part is actually how the economics work out. So if you look at it, the Transformer was invented inside of Google in 2017. NVIDIA came out with A100 GPUs in 2019. Shortly after, they came out with H100 GPUs, right? The original ChatGPT was trained on A100s. And then they came out with Blackwells, so then B200s, B300s. And now they came out with Rubens, which is the GPUs that Elon is sending to space, and they're like liquid cool, cool. They're very, very cool. And we're like, "Okay. Huh." One, every class of GPU is more affordable per one trillion FLOPs, right? So a FLOP is addition, subtraction, multiplication, any mathematical operation, and you measure them in how many trillion of operations happen per second in a GPU, and so they're more affordable as it relates to this. Um, if I was to buy an H100 for, let's say, $30,000, that's how much the kinda a single card would cost. If I wanted to rent an H100 for one hour spot instance from GCP, it could cost me $5. If I rented it from like, you know, Azure or AWS, maybe it'll cost me $3.50 per hour. So if I multiply that times 24 hours and then times 365 days in a year, I'm actually gonna end up paying 35,000 to $50,000 to rent that GPU for one year, but I could buy it for $30,000. So it's 1.5x the cost of owning the hardware to rent the hardware for a year. Now, the hardware is typically, um, warrantied for three years to work properly, but it'll keep working after the warranty for, I imagine, I don't know, 10 years. So the math just maths where it makes way more sense to buy them. The other big part is if you want to do large-scale training like we do, you need the memory to be co-located with a large cluster of GPUs. I can't just rent from Google or Microsoft or even Base10 and run the size of training that I want because I need a gigantic memory card next to it with all of my data that all the GPUs are accessing. So that's why we first started buying them. The next thing that we found is actually if you run open source models for coding, you could pay Anthropic, and then, you know, you're paying for all the tokens and the fact that you're doing the, the, the branded, right, Fable one. Or you can run an open source model, and instead of running it on a spot instance from Azure or anyone else, you run it on your own hardware, and then you're paying a fraction of a fraction of a cent per token. And so for all those reasons, it made a ton of sense, but we can go into all the depth that you want.

    3. HS

      I just wanna dig in. The first thought that I have is I completely understand the rationale there, but chips depreciate. You have chip cycles, and they are accelerating. We are seeing newer and newer chips being created. We're seeing specialization within chips. By buying, you're locking yourself in, so to speak, to one chip architecture. How do you think about that?

    4. CW

      At Speechify, we still use K80s for a lot of specific operations for inference, and we use older models of GPUs constantly. Um, and there's essentially a difference between when you do inference and when you do training. For training, I'm like, "Okay, I have this hypothesis. I wanna know the answer to this hypothesis as soon as possible." Like, every minute that it doesn't come out, I'm in competition with everybody else. And so having a GPU architecture that is much faster by orders of magnitude is a huge advantage. But if you go speech-to-text or text-to-speech with Speechify, I can afford to give you a lower quality GPU, and it'll give you what you need still in, you know, 100 milliseconds, so it's like totally good. And so I can always use these older GPU models for inference. That's number one. Number two, we have so many experiments that we're running at every single point in time. Not all of them need to run on like the newest hardware. So the analogy I always give, let's say you bought an iPhone back in 2011, and it's an iPhone 3G, and then you bought another iPhone and another iPhone and another iPhone. You could have a drawer in your house with like five iPhones that are collecting dust 'cause you can only use one iPhone at a time. But if I own 100,000 GPUs, I'm still gonna use all of them at the same time. And so I'm not losing anything by having more GPUs because not only do I own a bunch, I still rent from the hy- hyperscalers all the time, and I rent both dedicated instances that I prepaid for and I rent spot instances. For example, more people use Speechify in September 'cause everybody goes back to school, so I need to, like, level out the load. And so the parts of that load that I know for sure I'm always going to use, whether it be training or it be inference, I might as well just own it. And then on top of that is also the case that I have so many other friends who are running training and running inference, I can always rent it out to other people if I have excess capacity, which I don't expect to have, but, like, every once in a while you have an interesting situation. So for all those reasons, it just makes mathematical financial sense. Lastly, if you have excess capital, really you either stick it in the bank or you buy a bond, right? Like, the best long year ... L- the best bond you can buy long tail, I don't know, will yield you, like, 5%, or you can buy a GPU. And because renting it would cost me 1.5X buying it for the year, the return is, like, way higher.

    5. HS

      So how many GPUs do you buy then?

    6. CW

      So let's talk about Rubens for example. So Rubens

  3. 7:1212:28

    Why Buying GPUs Can Be Better Than Renting Them

    1. CW

      come in the form of 72 cards in one rack. So we'll buy multiple racks of Rubens, and then on top of that we'll buy B300s, which are, like, the newest form of Blackwells, uh, because we can get them earlier. And then the same thing, like, you know, when we bought our first instances of, uh, DGX H100 GPUs, we bought just, like, a bunch of racks of those, and then those get delivered in a truck to the data center. Uh, we rent the data center, uh, space, and so the data center provides the, uh, networking capability, it provides the energy, which is actually the largest constraint now, and it provide, like, physical engineers that take it off the truck, they install it. If it has an issue, they fix it. Um, and then it just runs.

    2. HS

      Does ElevenLabs do this? Is-

    3. CW

      Yeah. ElevenLabs is amazing at this. ElevenLabs-- I think Piotrek at ElevenLabs literally bought a bunch of GPUs early, early on and set them up in his house, and then they just kept building bigger and bigger and bigger clusters. The- they- they do the same thing that we do.

    4. HS

      How do you think about forecasting chip buying? It's incredibly difficult-

    5. CW

      Yeah

    6. HS

      ... to know, A, demand, but also, B, supply of chips.

    7. CW

      Yeah.

    8. HS

      How do you think about forecasting chip purchasing?

    9. CW

      Yeah. So number one, I wanna explain again, it's very different than buying an iPhone or buying a MacBook. I can only use one MacBook at a time, one iPhone at a time, but I can use all the chips I have at any given point in time, and I still will have more demand, especially when I have multiple teammates and 60 million users who are using inference on my Speechify software that's providing text-to-speech and helping them, you know, read their work and dictate their work and, you know, use Speechify Work, which is our, our newest product that's agentic, uh, kinda like JARVIS from "Iron Man." And so I go, "Okay, let's imagine I have 100 c- percent capacity that is the average usage per month that I need for GPUs for training of my AI models and for inference on my AI models." Inference is when you actually make a call to Speechify and you give me text and I give you back audio, like, there's math that happens in the background. That's inference. Training is I take a gigantic amount of data, I take all the architecture and software engineering that we're doing, and I go, "I think that this will give me a better model." I'm kind of baking that model in the oven and I'm gonna come out with a new black box, and then when I give you text, that black box is what calculates it and gives you back the audio. So those are the two usages. Let's say I have 100%, which is what I would have in, let's say, a month like November. In October, I'll have 140% because it's, like, a big month for us. In December, you know, everyone's at home, you know, they're not necessarily studying or working, so I might have 80% utilization. Okay. So then I go, "Cool. Well, I can take 20% of the usage that is normal and let me buy it because it's the best deal. I'll take another 25% of the usage and I do long-term contracts with hyperscalers. The rest I'll rent what's called spot instance from the hyperscalers." And then I'm still not even close to overcommitting myself. Um, and so that's kinda how we think about the math. And then we go, "Okay, well, also we have 45 engineers, but we want the team to be 150 engineers." And even inside of my 45 engineering person team, like, there's a couple people who are rock stars. They have dedicated, like, DGS racks just for that one person, and 25% of my team are almost, like, waiting, and I wanna double the size of the team. They just need-- Like, it's like you have a football team and you just need another field because they don't have enough field to practice on. Uh, and so that's how I think about how to allocate. And then in terms of depreciation of the asset over time, I go, "Okay, well, it-- these are still amazing GPUs." [chuckles] Like, even A100s you can run amazing experiments on. So it's completely valid to use that as long as it's hooked up and as long as it's not stopping to work. And so think about the mileage of a car, right? If a car gets to, like, 250 miles, you know it's kinda gonna break at this point. That's not necessarily true for a GPU because it doesn't have as much wear and tear. Yes, it's moving, and yes, all these things, but, like, it's in a very clean environment. It's very much cooled. It has constant maintenance 'cause it's not moving around. It's very expensive. Um, and NVIDIA just does a really good job. And so that asset is gonna stay for a very long time. And let's say it got so not good, so outdated that I no longer can run training on it. Cool. Now I'll use it for inference. There's one more thing that's very interesting that just happened. So, um, I believe earlier this month NVIDIA did a huge deal with Blackstone, BlackRock, Apollo, and Goldman Sachs, and they said, "Listen, we want more people to buy more GPUs. We're going to underwrite for you up to 25% the value of a GPU, that if you lend money to someone who buys a GPU, let's say Google or sm- or a startup, Core Weave, and that startup goes out of business and you have that GPU as collateral against that investment, we'll buy back the GPU for up to 25% of the value of the GPU." And so they're succeeding in creating a liquid secondary market for GPUs that they are underwriting, so now the large banks have an incentive to loan money at much better interest rates. This is actually exactly what Elon did in the beginning of SolarCity, he went to Morgan Stanley and Merrill Lynch and got them to amortize the price of a solar panel over 30 years. So the whole invention behind SolarCity was the fact that you could take a loan against the collateral of your solar panel. So NVIDIA has done an amazing job now in creating a clear floor for the value of the GPU over time.

    10. HS

      Do you think the circular economy fears that people often cast against NVIDIA are

  4. 12:2813:49

    Are NVIDIA’s Circular Economy Fears Overblown?

    1. HS

      justified or not? We saw their CFO push back on them and say, "Enough. Enough of this bullshit." Do you think that justified or not?

    2. CW

      Uh, I think that a lot of the things that about a year ago, like were going on between Oracle and OpenAI, like that was way too much. Like, that was ridiculous. I think the NVIDIA stuff is, is not, because you're talking about a real asset. So if you think, for example, about the logic behind the value of Bitcoin, right? Bitcoin, what is the intrinsic value of Bitcoin? I can't really tell you, right? What's the inj- intrinsic value of gold? Well, gold, you can use it for some medical stuff because it's a really amazing metal and, you know, it's jewelry, whatever. But a GPU, it has intrinsic value. Like, you can actually use that asset for something that's really, really valuable. And it doesn't matter where that GPU is. It could be in Iceland. It's still useful to anybody all over the world as long as it's networked. Uh, and so actually it has a pretty good store of value, even if new GPUs come on- online. Really my one question, and this is the math for everybody to come back to, is how many teraflops per second can this device do? And that is like essentially a token. That's the value. And so, like, there is intrinsic value. So yes, you can have all these like circular things, but at the end of the day, NVIDIA is making a project- a product that's real. It's not complete tool of mania. Like, there's a real, real intrinsic value here.

    3. HS

      What does no one know about buying chips-

    4. CW

      So much

    5. HS

      ... that they should know?

  5. 13:4918:26

    What No One Understands About Buying AI Chips

    1. HS

      What's the like, "Oh my God, people are so naive about this"?

    2. CW

      I mean, it's not that people are naive, it's just they haven't been in the space. So I'll give you an example. Imagine you're buying a GPU. Well, you're gonna buy it, you know, it's an, an NVIDIA-produced product. But NVIDIA is not gonna waste the time talking to Cliff Weitzman. So who do I buy it from? Well, one of the best-rated vendors is Dell. So everybody thinks Dell is a personal computer company. No. Dell is a GPU rack, uh, supplier at this point. And then, okay, I wanna buy it from Dell. Well, Dell has a constraint because there's not a lot of like, you know, r- Blackwells out there. Well, it happens to be that they have some in France. All right, well, I'm gonna order mine from France. Okay. Shoot, it was supposed to come a month ago and it's still not here, right? Because of whatever. Like, you know, there's demand. So then you need have to negotiate to make sure that you get it, which is why we're very willing to pay 100K per month extra to get them earlier.

    3. HS

      So, so you'll call up Pierre in France and say, you know, "Hey, we'll give you an extra 100K kicker if you get them here in a month"?

    4. CW

      Uh, even more than that. So in that France situation, which is something that happened to me, um, I was like, "Pierre, what the heck? We have a contract. You're not delivering on time." And so it is the case that we had a contract with another company beforehand, and they were, I don't know, a few weeks late. And I called them and I was like, "Listen, I've got a better deal. I'm canceling our contract 'cause you're, y- y- you're out. Like, you didn't deliver. So I'm gonna go with this other contract. But if you have a better price, like we'll go with you, but like, I just need the GPU now." And remember, I'm paying for the renting space of my data center. So the most expensive part of a delivery of a GPU is if it's late, I'm still paying rent for that data center space. Now that GPU... A- and so, you know, you put pressure on Pierre to send you the thing when he said he was going to send it to you, and then you go to NVIDIA or Dell or whatever, and you're like, "Well, it's, it's a market. Hey, can I pay more to get it earlier? Skip the queue?" "Yeah, you can." "Cool." Now, there's a truck somewhere in the United States with a GPU whose value is the value of a house that's coming to my data center, right? Well, I should have insurance on that, right? Because if that truck gets hit or there's too much humidity or the GPU gets flipped, like I lost multiple houses worth of GPUs. So okay, the insurance is really, really important. And then also the value of the amortization is really, really important. And it's like there's all these nuances of how to do the math through. Then there's the cooling, right? So like you're not only paying for the physical space and the networking and the energy, and the energy is the biggest constraint. We'll talk about it in a second. Well, how do you cool that thing? Because you have a thing that's just like moving and moving, moving, moving, moving. Um, well, and the thing that's like most new now is liquid cooling, because air is just not enough, and the, uh, the thermal load of water is much better, and there's other liquids that are even better than water. And so Rubins are, are liquid cooled. But most of these data centers don't have liquid cooling installations already approved. So we had to do a bunch of research and we find, okay, we could buy what's called a side cart of liquid cooling that you enter into the data center, then you pay someone at the data center to install it for you. Cool. Now you can have like the rack that you want. Uh, and so there's a big difference between running a purely software company and running a company that includes hardware. Um, and-

    5. HS

      But when I listen to all of this, I'm now more sure than ever that it is a mistake to price optimize and to spend the money to buy it versus to rent it. Because I get you on the optimization, but you're not saving 10 times more. It's 0.5X more per year.

    6. CW

      No, no. Per, per year. Exactly.

    7. HS

      Yeah, per year. But you have the flexibility to tailor it up and down. You don't have any of the logistical nightmares of insurance, transportation, security, water cooling, logistics, and then you can build your product, actually what matters most, against ElevenLabs who are fucking running fast. I don't wanna worry about water cooling and insurance for a freight truck.

    8. CW

      ElevenLabs worries about the same thing because for them to train excellent models, they need to have co-located GPUs with a lot of memory available.

    9. HS

      You can't do it if you rent it.

    10. CW

      You can. It just becomes, one, ridiculously expensive. Two, you need to commit for many, many years ahead of time 'cause you need to build a co-located cluster. Um, and then you don't have as much control 'cause you don't own it. So it's like difficult to... Like, you suddenly need an InfiniBand cable, which allows for the memory to flow from one DGX to the other one. And the answer then is, well, now my ability to train is so much bigger. I can have a larger AI team. Every person that in a AI team is leveraged, and I could just, I could shoot ahead of everybody so much faster. And let me just make one thing clear. If I want the Rubin, which is like these much faster GPUs, I'll get it faster if I buy it than if I wait for Google to buy it and then there's other people in front of me in line. So I'm, I'm gonna skip the queue by like a lot, and

  6. 18:2620:55

    Why Owning Compute Can Become a Competitive Advantage

    1. CW

      then I'm gonna have like a year of access to Rubins before everybody else does. The way I think about it is the following. How do you build an amazing company in a world where there's so much competition today? The team is the most important part, but the team is the most important part 'cause the team gets you the other resources. And so what are the missing, missing pieces? The missing pieces are data, compute, and architecture. In a world where intelligence is commodified and no one needs to hand write code at all anymore, right? Our engineers, really what I'm looking for is 10 really good decisions per day, which is very tiring, not like optimizing random parts of the code. And each one has like, you know, 5 to 18 agents running at any point in time doing long horizon tasks on these GPUs, coming up with theses, testing them, going back and forth, back and forth, back and forth. If they don't have the capacity to train, the team is limited. If they don't have the data to train, the team is limited. And by the way, a lot of data, it-- cleaning the data, right? You get this raw data in the beginning. Well, you need to organize it into data sets, and so you need the GPUs to also organize the data sets too. Like, I have one of my, my best engineers right now is not even writing models. He's making synthetic data sets to train models. And so, like, it really becomes a indispensable asset.

    2. HS

      Would you ever buy data?

    3. CW

      We have, but like small data sets.

    4. HS

      So I suggest you use Fireworks.

    5. CW

      Okay.

    6. HS

      But, uh, I mean, Fireworks is amazing. Lynda, the founder is-- she, one of the co-founders of PyTorch. But I, I had the very obvious realization that you'd have every company having their own specialized models of a certain size trained on their own data. Um, but you would need supplemental data-

    7. CW

      100%

    8. HS

      ... like the synthetic data or real-world data that you just don't have, and that you would buy that from data providers like Mercor-

    9. CW

      Yeah

    10. HS

      ... which is why I was so drawn.

    11. CW

      MicroOne, Surge, all these companies are amazing, and they shorten the cycle, by the way, to getting to revenue.

    12. HS

      100%.

    13. CW

      Because if you are Mercor, uh, shout out Brendan Footy, um, ElevenLabs or OpenAI or Anthropic is gonna make money for the next decade or two on the data that they bought from you. And so they're willing to pay a fraction of that 10 years of revenue to you today to supply them the data, and again, it's all a speed thing. Yes, ElevenLabs is-- or, you know, OpenAI can go and make a team that will get the data, but they don't wanna manage it. And so the data is key. All this-- Like, you need all three things. You need compute, you need data, and you need a team that writes great products, and ideally you need users that use you a lot and allow them to have a feedback loop of whether the stuff is good or not. So benchmarking.

    14. HS

      And so, and so for me, when I was doing the, as a venture investor, we do outcome scenario planning, which is the

  7. 20:5522:59

    How Big Can the AI Data Market Become?

    1. HS

      most bullshit exercise to pretend like you're smart predicting the future. We do it because, you know, it makes us feel important. Um, but, but, you know, they predominantly sell to frontier labs today, and that's where 90% of their revenue's from. With the rise of specialized models on a per-company basis with their own data, I believe that you move that customer base from purely frontier labs to every large-scale enterprise who needs supplemental data. If that is the case, how big a outcome is the data marketplace?

    2. CW

      So the first problem to understand about the data marketplace is it's not ARR, right? It's not annual recurring revenue. It's one-time deals every single time. So the buyer of the data is not required to buy it from you again. So it's a very risky business, and if you look at early days of companies like Mercor, they didn't raise significant comp funding off the bat because investors were very skittish about that fact. Let's put that aside. Well, very important that the T's are crossed and I's are dotted about how you got that data, right? So like, you know, like y- you need, you need to indemnify the companies who are using you, and that's part of why they buy it from you as opposed to sourcing it themselves, right? We've seen the lawsuits. Um, but it's a great business. And if you could do it well, but like, you need to be an ops monster. Like, you need to be really, really good at operations. You need to be very fast. And really, the key is the company training on your data needs to actually see improvements in their model at the end of the day. The thing that has always been challenging for Speechify compared to other companies is B2C comp- customers pay a lot less than B2B customers. So ElevenLabs, huge credit to them, leapfrogged us because they sell to B2B. Well, historically, we've only sold to B2C, and so our big constraint is we need to do this on a cost basis of it needed to cost us, you know, less than $10 per million characters. Eleven charges ten-- $100 per million characters. The OpenAI model on the benchmarks cost $196 per million characters. So ours, when we sell it to other B2B companies now, we just launched our API, Simba 3.2, it costs $10 per million characters.

    3. HS

      Dude, I am, [sighs] I'm too old to not ask the painful questions, and I think

  8. 22:5925:23

    How ElevenLabs Leapfrogged Speechify

    1. HS

      the joy is the more you ask them and kind of the less, you know, worried you are about asking. You said ElevenLabs kind of, uh, leapfrogged you. Is that on you for not doing B2B?

    2. CW

      Yeah, 100% it's on me. 100% it's on me. It's the biggest strategic mistake I made in the history of Speechify.

    3. HS

      How do you reflect on that?

    4. CW

      So I met Piotr and, uh, Mati. I was living in London at, at the time, in my house in London. I think it was 2022. And we were very impressed by them. And we wanted to use their model, by the way. It was just too expensive for us to use. And I looked at it, and my thought to myself was, "They're very smart. They're gonna do well, but I don't like their strategy because I think that an API for text-to-speech will become commoditized with time," right? You're gonna get to the point that you can run that API on your computer and then on your phone, and then like, what are they selling anymore? So I don't wanna go into that business. And I made a critical error. What I didn't understand is that the point of an AI lab like Speechify or like ElevenLabs is to continuously innovate, and the first product that you release is your wedge that gets other people to then later use your other technology. So for example, if you're into text-to-speech, you build the best text-to-speech model in the world for one specific voice. Cool. Well, now you can do other voices. Now you can add, add emotional prosody. Now you can add voice cloning. Now you can add speech-to-text. Now you bu- build duplex models where it makes the um, ah, laughter, uh, interruption handling, turn-taking. Um, you add a harness for voice conversations, then you optimize it for sales, and you optimize it for customer support, and you optimize it for all these things. And so what they did is they first built an amazing API They were great at launches. They built a really great, uh, product for creators. Then they built their best product ever, which was Agents. Agents is amazing because the buyer is no longer a software engineer. The buyer is a CTO, CIO, CEO, executive in the company. Uh, Sierra has this concept called outcome-based pricing. Brett Taylor is amazing. And so you can start fighting on the outcome, and having a AI agent is like having an AI coworker. But it was my mistake to think that an API product was a bad strategy because I thought it was something that would become commoditizable, and I forgot the central thesis about Silicon Valley, which is constantly innovate. Get the user to start using your product. I don't care if it's free. Then you sell them other things. And so that was my big, big, big, big mistake.

    5. HS

      How possible do you think it is? I think people underestimate the complexity of building out a B2B GTM.

  9. 25:2330:17

    Is Speechify Making a Mistake Going Into B2B?

    1. CW

      Super hard.

    2. HS

      I think it's a strategic mistake for Speechify to go to B2B.

    3. CW

      Hmm. A lot of people think that. Tell me your position.

    4. HS

      You are now competing against ElevenLabs and Sierra, really, and those two are competing. Whether they like to admit it or not, they absolutely are competing, and they will, I'm sure if you ask them off camera. Um, that's Brett Taylor.

    5. CW

      Yeah. You don't wanna compete against Brett Taylor.

    6. HS

      Motherfucker, I don't wanna compete against Brett Taylor. Uh, and that is the tidal wave of ElevenLabs. Now, ElevenLabs is an unstoppable machine at this point, to the point where it has government buy-in across all of the large major Western democracies. Actually, it's insane the government buy-in they have, and they started three months ago. I, I ju-

    7. CW

      You just said the key thing. They started three months ago.

    8. HS

      Yeah. And, and-

    9. CW

      And so you know the graph of OpenAI-

    10. HS

      And, and, but I think they've reached a tipping point where actually they've just taken the market. I think Sierra are running behind them chasing, and they're doing a decent job of it, but they've got Brett and they've got Si- Sequoia and Green Oaks and every royalty of Silicon Valley behind them, and they're still running behind chasing ElevenLabs with Sequoia kind of pretending to be neutral [laughs] 'cause they're in both of them, which is incredibly challenging. And I j- I just think being third, the Postmates effect is never a good market to be in when I could be the dominant consumer brand that leads with a really different and compelling story.

    11. CW

      So here's the two things to consider. The first one is if you go to the App Store and you search text-to-speech, Speechify has 98% of the installs in text-to-speech for B2C.

    12. HS

      Hmm.

    13. CW

      Speechify has served more than 770 billion words to users over the last few years, which in terms of times of listening, it's like 6,000 years of listening, right? If you go from today to 0 BC and back, you still have, like, thousands of years left. So we've, like, completely dominated that market, and it's still a business that's growing really, really fast. Uh, and we're constantly adding more features into that product. The thing is, we have a pretty big engineering team, and now everybody's capable of doing 10X what they did before. So I have extra staff. I have a huge AI engineering team with ability to make amazing models. Uh, so, like, where is the highest ROI for that to go? Well, it needs to go both B2C, but it should also go B2B. And one thing that I will never be is a person who doesn't learn, so I might as well just freaking learn B2B. Now, to your point about competing against giants like Sierra or ElevenLabs, hey, Anthropic came into the market as a second to OpenAI, and they were second for a very long time, and now they're not second. Facebook came as a second to Friendster and now, and MySpace, and now they're not second. And so the nice part is this space is not a monopolistic space. It's an oligopical space. And if you look at what happened with ElevenLabs, I'm gonna exclude Sierra because Brett Taylor effect is huge, it's just amazing to see how good of a business that is. And so it might very well be that for the core offering that they're currently winning on, I will not win. But what did I learn last time? It's fine if I offer my product essentially for free because I'm an AI research lab. And as long as people start to use me, with time, I'll be embedded in the system, and I'll keep coming out with more and more and more innovations that are useful to them. And so there's un- unbelievable demand from all these companies and governments and everybody else for great tools, whether they be AI agents or APIs or products. I just wanna be on your phone if you're a user or in your stack if you're a company and supply you with the best, uh, front deploy engineer experience and AI orchestration experience and API experience to give you an amazing experience, and there's room for everybody.

    14. HS

      I agree there's room for everybody. I think v- value accrues to top one player.

    15. CW

      I agree.

    16. HS

      I think, you know, it's kind of like the inference market where Fireworks will be a multi-hundred billion dollar company, and then, like, I think a genuine baseline will be a hundred billion dollar company, and then together and a load of the others will be 50, and which is amazing.

    17. CW

      It is completely true. Power law, yes.

    18. HS

      Huge, huge, hugely pow- amazing, valuable companies. [laughs]

    19. CW

      But you would then think that OpenAI would be the place where value accrues for voice AI, right? That's what you would have thought three years ago.

    20. HS

      Mm-hmm.

    21. CW

      And that's not what ended up happening. So you can't not go into the race because there's a big incumbent.

    22. HS

      Well, I think with all candor, that's because of incredibly poor management.

    23. CW

      I agree. But that's the thing. Every-

    24. HS

      And hiring. And, like, that, that was theirs to take, and they fumbled the bag across every spectrum.

    25. CW

      And for, and for every company in the world, no matter how exceptional the leadership team is, n- niches get fumbled, right? So voice AI was a niche for OpenAI, right? LLMs are the core, and by the way, they also fumbled AI coding. Now they're trying to catch because it's such a big space. All respect to Piotr and Mati, I think they're absolutely amazing, and w- I, I love working adjacently to them.

    26. HS

      I just don't think they're gonna fumble the bag. That's been my trouble is I don't think they're more-

    27. CW

      But, but, but they have so much in their net right now.

    28. HS

      That's true.

    29. CW

      And so much is getting added to the net constantly.

    30. HS

      That's true.

  10. 30:1731:10

    Why Every Winning AI Company Becomes a Compound Startup

    1. HS

      to be a compound startup. Can you talk to me-

    2. CW

      Yeah

    3. HS

      ... about that and how you think about that?

    4. CW

      It's not that every startup has to be a compound startup. It's at a certain point, you can't afford not to be that.

    5. HS

      Do you not think there are a few companies that are just absolutely fucking running rings around everyone else?

    6. CW

      Yeah, absolutely. Those are the winners, right? ElevenLabs is an example. Um- Uh, Anthropic is an example. Ramp is an example. Speechify is an example. Um, all the companies that have absolutely maniacal leadership teams and engineering teams, like, that's why people care about team more than almost anything else. Because a, the right team will iterate fast, get there, and then figure it out. And now, when everything can be turned into a reinforcement learning problem, where you can have long-horizon agents and orchestrating agents thinking about the problem for, like, two weeks at a time, if you set that up, of course you're gonna win.

    7. HS

      I got into a lot of trouble, as I always do with most of my social posts. Uh, I used to be quite a sweet little

  11. 31:1037:26

    Is This the Hardest Time Ever for Startups to Hire Great Talent?

    1. HS

      boy, actually. No, really, I used to be like the Harry Potter of venture capital, and now I'm more like-

    2. CW

      Yeah, you lost the glasses

    3. HS

      Lost the glasses and kind of became more like Piers Morgan, if you know Piers Morgan in the UK. A highly despised figure. Uh, very opinionated. Um, but, um, a question that I have is, like, I said if you're a startup, it's never been harder to hire great talent because OpenAI and Anthropic candidly have such a carrot reward mechanism in front of you, especially with impending IPOs, um, that the best talent just wants to go there, and talent follows talent. And you're seeing the fucking founder of Monzo, a multi-billion dollar bank in the UK, go there from YC as a partner. Matt Clifford, you know, the founder of EF, which is a multi-billion dollar company, I mean, he, he, he should be fucking prime minister, and he's going to join Anthropic. Am I wrong that this is the hardest time ever for startups to hire because the prizes of Anthropic and OpenAI are so great?

    4. CW

      My favorite type of person to hire is a CTO of another company. We have- When we were 21 people at Speechify, 18 of the folks at the company were previously either CEO, CTO, or VP of engineering at their last company. Anthropic, I have never seen a company like this, hires so many CTOs of publicly traded companies and other successful startups.

    5. HS

      Were they one of them?

    6. CW

      The reason is they build the, they build the best, most beloved product for engineers in the history of the world, so it's easy to hire CTOs. By the way, they hire much more CTOs than CEOs because CTOs are the ones who get the most excited about this product. And like, you're right, they're the fastest-growing company ever, especially at the scale that they are. So they're gonna keep growing. OpenAI's gonna keep growing. Uh, you had this, like, very condensed period, like fireworks of growth in both of those companies. Yeah, it's very hard to hire, but remember, they're hiring people that their annual compensation needs to be $15 million a year minimum. What startup is hiring someone and paying them $15 million a year? You're not. Like, you're seed founder. That was not some- one that you are going to hire, and so I will push back against it. The competition for growth stage companies hiring exceptional leadership talent is more difficult. For seed companies, I would say it's the easiest time ever because the impact of even just the founder on their own is bigger because they can orchestrate agents. But the same thing for hiring. So one thing that we have changed about our hiring in the last even six months is we really cared that you read a ton of textbooks about software engineering and that your handcrafted code was amazing. I still care that you read a lot of textbooks about software engineering and you understand it, but the thing I care about the most today is technical aptitude and just, like, raw technical intelligence because I know that we could teach you everything else, and in six months you could be a machine. And so we hire a lot of math Olympiads and LeetCoders and, like, Kaggle award winners and people who, like, studied physics and math. Like, they might have even not coded before. Because I just need the hunger and the work ethic and the intelligence, and anyone can become so good so fast now. And so the pool for hiring exceptional talent is bigger than ever before, and Duolingo did this really well. They love hiring college grads and then coaching them. And so I wouldn't say that it's harder to hire than ever before for seed companies. Seed companies now, almost anyone can be someone that you hire if they're smart and hardworking because you can teach them very fast. Um, what is, uh, more challenging to hire is for growth companies because you're fighting with just absolute juggernauts.

    7. HS

      Are you not a growth company?

    8. CW

      So it's challenging for us.

    9. HS

      [laughs]

    10. CW

      Why do you think it's hard to hire a really good salesperson?

    11. HS

      I totally get that-

    12. CW

      Yeah

    13. HS

      ... and I completely agree. I will see CR packages in the 50 million-plus range, by the way.

    14. CW

      Yeah, exactly. And so-

    15. HS

      15, 15 is like kids play.

    16. CW

      Hire, hiring-

    17. HS

      By the way, by the way, um, with the greatest of respects, I will even see $15 million on the table for comp packages for seed companies today. That, that is, that is the dislocation that I think, with the greatest of respects-

    18. CW

      Wait, wait, wait. Sorry, sorry. But it's, so this is a seed company that's, has, ha- has raised how much money at what valuation?

    19. HS

      Uh, well, I mean, you've got to understand a seed round today will be 150, 200 million.

    20. CW

      And-

    21. HS

      And there are several of them. I mean, there's 30, 40 companies that at seed have raised 100 to 300 million.

    22. CW

      And this is a company of, like, a guy who's, like, one year out of university?

    23. HS

      No, no, no, this is a guy who's probably spent four years at OpenAI-

    24. CW

      All right

    25. HS

      ... or spent four years at DeepMind.

    26. CW

      So then what about the company that's like, you know, the guy who's been at university for, like, two, three, four years and now they're starting a company? Or do you think that those people are out of the water now?

    27. HS

      No, I, I think that that's just a very different world.

    28. CW

      Mm-hmm.

    29. HS

      And so yeah, they'll, they'll raise $10 million seed rounds.

    30. CW

      Yeah.

  12. 37:2639:18

    How AI Is Completely Changing Software Engineering

    1. HS

      agents." What did you find? What did you learn-

    2. CW

      Mm

    3. HS

      ... in that discovery process around agent orchestration internally?

    4. CW

      Yeah. So inside of our AI research team, everybody's orchestrating agents. It's when you go lower, not lower, if you go then into the product-facing things that we build, for example, the platform team or the iOS team or the Mac team or the Chrome team or the web team or the Androids team, these are super smart folks who have been working in those domains for like 10 years, and they know iOS like the back of their hand. They know Kotlin, JetBra- Brains like the back of their hand. And so it's very easy for them to hand-code things because y- you're not dealing with something that's like super, super new. So why change? People, you know, it's hard to change, right? Um, and so you just need to force them to change. So one, the best thing is to inspire. So you do a Zoom screen share, and you show them how the best engineer in the team is orchestrating agent, and they're like, "Oh, wow, I didn't know you could even do that." And then you go, "Yeah, like please do it." You recommend, uh, blog posts for them to read, books for them to read, Twitter threads for them to read.

    5. HS

      What's the team using? Clawcode, Cursor, Codex?

    6. CW

      Cursor and Clawcode. Those are the two most popular. Yeah. It's a little bit of Codex usage, but it's not that big. I would say Clawcode is number one, then Cursor, and then Codex. We want you to use as many tokens as possible in whatever harness way is the best for you. Um, you mentioned Linear. Linear is amazing. Like automatically cutting tickets from Linear is fantastic and just like being able to go into your agents and be like, "Okay, I have these like six Linear tickets. Start on them." And then really a good engineer today is just an exceptional QA, right? The AI will make the feature. You will test the feature, see if it's good. You'll fi- figure out where the edge cases are. You'll prompt it to fix it, and then you try to make it as efficient as possible, which is hard to do, and then you need to make essentially like, you know, roughly 10 really good product and engineering architecture decisions a day.

    7. HS

      How do you think about token allocation internally? You know, we, we've seen leaderboards

  13. 39:1845:29

    How Should Companies Think About AI Token Spend?

    1. HS

      be used, which is I think the most fucked up form of incentive kind of playing. Uh, you don't wanna like prevent people from-

    2. CW

      There's a lot of people who are a lot of talk, and I'll ask for examples, and I'll read the examples and like, you know, "I'm doing this, I'm doing this, I'm doing this, I'm doing this." And then you look, and I'm like, "Eh." And so I think about it in terms of demos. Can we hop on a Zoom call, and you'll show me what you built? And then I use it myself, and I'm like, "Wow, that's amazing." Or you send me a screen recording of a feature or technology that you built, and I'm like, "Wow, that's so good." And so we give credit when things get shipped to production to users. So even inside of the AI team, if you build a really amazing-- And this is part of why Speechify ended up winning. You asked, "How did you build BiggerLabs?" The answer is we ship to production all the time. That's how we won. We are not in the theory space. We are an applied AI company. That's why we win. And so if you're an engineer at Speechify, the analogy I always give people is imagine that you are in the milk delivery business, and you make me a beautiful bottle of milk, and you leave it down the road. The milk will spoil. You have to get it to my door, knock. If you didn't do that, you get no credit. If you carry the football all the way to the line but you don't cross over to the end zone, if you don't kick it into the goal, you get no credit. If you bring the ball just to the rim, you don't put it in the rim, you get no credit. And in the rim means pushed to production with no bugs, and users are actually using it, and then we get feedback. How many companies do that iteration cycle fast? Almost no one. Definitely not with the user base number that Speechify has. And so in the AI team at Speechify, you make some amazing discovery. We're like, "Great. Push it to production." And then you go, "Oh, wait. There's this QA problem and this QA problem. And if you have this many people use it on the AI serving layer, then you have this other issue." Cool. You get no credit from me. It's not in production. I can't use it on my phone. When I can use it on my phone, I will give you credit. And so this morning-- Actually, not m- yesterday. Yesterday, I had a call with our AI engineering team, and I said, "Listen, the product that we have running for duplex models and for AI conversational harnesses is something I'm really excited about, and it's been moving fast. I want it to move faster. Here's like 14 notes that I want." And then what I do always is I'm on a Zoom call. I flip my computer around to face my phone, and I use the product in front of them, and we record it, and so then they see all the bugs. And then I send the recording in the chat. Someone on our team, he's 19 years old, sent me a demo this morning off of that conversation that solved all of my problems. And he was like, "Hey, I was waiting for like three training runs to finish, so I had a little bit of time while I was waiting, so I implemented everything that you asked." And it blew my mind. It was so good. That's using AI correctly. So it's not a token leaderboard. It's what did you show in production that was good.

    3. HS

      How many companies do you think are actually as token-pilled AI-centric as we think in terms of devs?

    4. CW

      Uh, there's a guy, Jason Jaeger, who used to work at Speechify, and now he has, uh, My Tech CEO on Instagram.

    5. HS

      Mm.

    6. CW

      He's super funny. And so he makes a lot of videos about like, you know, crazy CEOs who all, "Use tokens, use tokens." I think all founders in some way have that animal inside of them because you know that it's the right path. But there is a difference between reality and theory, and you need to make sure that you don't overdo it. So-

    7. HS

      Do you have any price sensitivity on tokens?

    8. CW

      Yeah, of course. Absolutely. I mean, I'll lose my mind if to implement a tiny feature, you use 15,000 tokens. Like, [chuckles] why did you do that? And, like, we will let people go if they just go bananas with something for no reason.

    9. HS

      Are you able to accurately budget tokens on a-

    10. CW

      Not accurately, but within bounds.

    11. HS

      Yeah.

    12. CW

      Um, the other thing is, like, a lot of engineers are... Look, y- you go into engineering because you like optimization. Most engineers are not blind, and it physically hurts them to overspend tokens. Um, and a- I, again, I always think that the best way to interact with AI is you are chatting in the chat or actually doing it verbally, and you're essentially pseudocoding with your words constantly, and you're explaining architecture. And a great example would be, um, uh, I know someone who, uh, has no engineering background, and they wanted to build an app. And they built exactly what they wanted. It took them two hours. Uh, but they needed an API call, and they needed to scrape this website, and they basically scraped every single page of the website, every single part of the website. And so the bill that they got for the scraping was gigantic. And then I was like, "Why are you doing like that, like that? Why aren't you going into the database to this exact URL and then scraping that from the URL?" So the amount of nodes they needed to hit became, like, 20 instead of 25,000. And so an engineer will spend their time making sure that the thing is optimized like that. So that's how you build, like, a good database or a good architecture system, whatever. You do the same thing when you're interfacing with the agent. You want the agent to take the path of least resistance, not the path of most resistance.

    13. HS

      I think one of the biggest problems is the agents are goal-seeking. And so they are-

    14. CW

      Yeah

    15. HS

      ... like-

    16. CW

      It's all about the target. You need to be good at picking the right target. And, uh, I, I think Anthropic published this paper, um, when Fable 1 came out about long horizon tasks with Fable. So the first thing is it was much better at, like, running a two-week task. And it could burn $12,500 worth of tokens in two weeks and basically make a better model with that. That's a perfect, amazing way of using tokens. That's exactly what you want. And what you don't want is burning 12,000 tokens in the span of five hours doing something that's, like, just totally unnecessary and doesn't make any sense. You need the loops to happen, and then you need to check the result. So what you wanna build, and Boris, uh, who's the investor, uh, inventor of Claude Code, talks about this all the time, it's all about the loop. You say, "Here is the target. Here's how you measure the target. Now iterate against the target over and over and over again until you get it."

    17. HS

      What did you not know about building an AI-centric dev team that you wish you'd had known?

    18. CW

      How useful is it to own your own GPUs? They said on GPU-

    19. HS

      What was that realization moment? Just did you see-

    20. CW

      Um, yeah, yeah

    21. HS

      ... a bill one day?

    22. CW

      The realization moment was when we realized that we had really talented engineers who were essentially moving at one-seventh of the speed they could have if they had the compute, um, one-to-one with their creativity and ideas.

    23. HS

      How, if you're a founder listening to this, how should I change my hiring

  14. 45:2947:03

    How Your Hiring Process Needs to Change in the AI Era

    1. HS

      process in a new AI world?

    2. CW

      Number one, functional interviews. Build this, and then you see if they can build the thing, and then you run it through unit tests. The second one is give them a large code base, uh, even an open source repository, and have them understand the code base, make changes, and then check what they broke. And then, yeah, like, they have to be able to orchestrate agents well. And if they're not doing that, it's kind of not worth to have the person. And then the next thing I'll say is, it is more fun to have a smaller team. Like, having a big team is great as long as everyone's carrying their weight. Um, but the way I kind of think about it is, yes, I can have multiple agents running on my computer, or I can have several Slack chats with really smart people who are m- bigger domain experts than I am, and basically that human being is the outcome owner for that task. And they have the agents. And so I can run, as a founder, multiple projects at the same time to a really amazing level of granularity. And so I think about moments earlier in the year where my brother Tyler would literally have an alarm to wake up at 3:00 in the morning 'cause he needed to check what the agent was doing at 3:00 in the morning, and then he'd wake up, make sure it's good, go back to sleep. Like, y- you want to babysit your agent basically every three hours, and the beautiful thing now is you can go work out and the agent will tell you the answer, and then you, like, voice note back with Speechify what you want it to do next, and, like, that'll happen. And so you want people who are essentially that level of addicted. Obviously, that creates massive AI fatigue, so make sure your teams don't burn out. Um, but you want someone who is that level of excited. And so I think hiring for slope more than intercept is more important today than ever before. Said another way, I look for the potential the person has more than I look for where they are today.

    3. HS

      When I look at Whisper Flow and Willow, and I did this tweet,

  15. 47:0350:18

    Is Voice AI Becoming Completely Commoditized?

    1. HS

      and I deleted it because I don't ever wanna be sulky and miserable, and it's an amazing thing to build a company, and you should be incredibly, you know, credited for doing so as an entrepreneur. But I found Whisper Flow's product was just getting worse. Um, and I said it on Twitter just 'cause I honestly just wanted alternatives. I, I, I-

    2. CW

      Yeah

    3. HS

      ... really need this product, and I wanted alternatives. I got 500 different alternatives, and I was like, "Motherfucker, we'll talk about the commoditization of a market. That is not one that I wanna be in." Can you help me understand? Are we seeing the complete commoditization of that Whisper Flow, Willow speech-to-text for productivity?

    4. CW

      What they came out to the market with first was not necessarily their own model. Part of the reason they got worse is they switched to their own model 'cause it's a lot more affordable. Um, and so they had a harness that ties together a bunch of other things. Um, probably it was DeepL under the, or Deepgram under the hood, uh, with a bunch of optimizations and more, more products.

    5. HS

      Well, now they're trying to do notes, and they're trying to move more into productivity.

    6. CW

      And I think they're being successful with it, yeah.

    7. HS

      100%.

    8. CW

      Yeah. So that, that, like, that's to your point of the compound startup. Uh, one of my, uh, biggest mentors-

    9. HS

      When you look at them, do you not reflect on your, we said before, not announcing fundraisers, not announcing anything.

    10. CW

      They're the opposite of me.

    11. HS

      They've announced everything. They announced-

    12. CW

      They're the opposite of me

    13. HS

      ... going to the bathroom. Um-

    14. CW

      Correct

    15. HS

      ... and hence they have a, I would say, a bigger brand.

    16. CW

      W- uh, not in terms of users. Like, w- if you walk down the street in New York City, way more people will know Speechify than know Whisper Flow just by virtue of the fact we have way more users. Um, but in the tech world, way bigger brand, right? Investors know who Whisper Flow is because they announce. We intentionally don't announce. But we don't have any competitors Who are you gonna use instead of Speechify to do text-to-speech for your models? Like, the closest thing is ElevenLabs, and we're, like, so much bigger than ElevenLabs for B2C.

    17. HS

      What about Gradium?

    18. CW

      What's Gradium?

    19. HS

      Brazilian company that does text-to-speech. Competitors to ElevenLabs.

    20. CW

      Okay, but they do B2C?

    21. HS

      Yeah, no.

    22. CW

      Yeah. So that's, there's unlimited numbers of companies doing B2B-

    23. HS

      Yeah

    24. CW

      ... text-to-speech.

    25. HS

      Right.

    26. CW

      But, like, we are unique in our market. So because, uh, uh, WhisperFlow was so public about it, they now have a lot of competition, and so Peter Thiel, right? It... Only losers compete. Try to not compete. And so, yes, they-

    27. HS

      What happens to that market? WhisperFlow take majority, and then there's thousands of ankle biters?

    28. CW

      [sighs] I don't know. I mean, I want them in that market, right? I, I think that market becomes oligophetical as well. Um, obviously, and this, by the way, I think is another mistake that I made. I built my own speech-to-text, uh, experience that I've been using on my computer for the last, like, seven years, um, side loaded on my iPhone and on my computer. But I figured it's a commoditized product, right? Apple's gonna release it instead of the button. It'll be great, and, like, there you go. But Apple keeps not doing it. If you remember two years ago, Apple announced a partnership with ChatGPT that will improve Siri. Nothing happened. And so that's also the reason why I never went after Siri. And so now we've launched a product to compete with Siri, and we've launched a product to compete with WhisperFlow, and we've launched a product to compete with Open, uh, with ElevenLabs, because I learned a lesson that I should've learned before, which is the same lesson from ElevenLabs. The way that you win is you offer an excel- excellent product for free, and then you have a wedge, and then you add more and more and more things. And so I don't know what happens with the WhisperFlow space. I just know that if you're a founder, you should also always try.

    29. HS

      Final one. I, another one that I get in trouble for but I, I stand by strongly is I just think the customer

  16. 50:1853:34

    Is Customer Support the Wrong AI Market to Bet On?

    1. HS

      support market's a challenging market to really get behind.

    2. CW

      Yeah.

    3. HS

      You have Sierra and Decagon out in front with the majority of funding and attention. But to say that, uh, there are 18 companies that have now raised over 100 million in the last 18 months. Uh, there is a w- kind of what I would call, like, the mid-tier, which is like your Intercoms and your Talkdesks and your Crescendos and all these ones, whereas, like, you know, they're not old, but they're old enough.

    4. CW

      Yeah.

    5. HS

      8 to 10 years old, and they're pretty good.

    6. CW

      Yeah.

    7. HS

      And then you've got Salesforce, Atlassian, and the much older ones. And then the worst thing about this market is that for any sophisticated buyer, an Airwallex, a Klarna, a Niovan, a technology-facing company, everyone has built their own-

    8. CW

      Yes

    9. HS

      ... because they need a sophisticated-

    10. CW

      Of course. Then why would you pay a tax for it?

    11. HS

      So why... What am I missing?

    12. CW

      Yeah. So the first thing you're missing is the core product that we're offering B2B is the API, not the agents, right? So Sierra doesn't have their own, uh, model team. They use other people's models, right? Because the value of Sierra is the go-to-market. It's Brett Taylor. Um, and so that's why if you talk to Monty and Piotr, they'll tell you we're not competitive with Sierra because their main business historically has been the API. So that's the first thing. In the API business, you have Speechify, ElevenLabs, Gemini, uh, Groq, so SpaceX is now in the race, and, uh, Kurtisja, and that's kind of it. Um, and so that's not that competitive a space compared to, yeah, the B2B, um, uh, customer support thing. Everybody's in that space, Finn, everybody. And so I'm not building that product. The, uh, what can I offer you that's 10X better than the next person? Not much. And so in the core API side, I can offer you better quality, faster speed, and 10X cheaper. Good offering. But then I have to also offer agents because there are so many pockets of value that have not been unlocked. And unless I am... Again, I have this model for leadership. You don't want to be a fat manager who's like a general sitting in the back saying, "Take that hill." You want to be the warrior who runs up with their sword and engages the enemy first. You need to be the same thing with your product. You need to be the number one user of your B2C product, and you need to help your customers use your product better. And if you do that, you will learn their problems, and then you will figure out what the next product is that you need to offer them. So unless I have front deployed engineers working with my B2B customers, building agents for them using our technology, I will not figure out what the really amazing next innovation across the hill is. And so you mentioned the right thing, which is ElevenLabs now has all these partnerships with governments. Governments is not exactly customer support. They would have never gotten to governments had they not done a great job on the private sector first. I agree. ElevenLabs is, in, in addition to OpenAI, is the most integrated company right now, AI company with governments. That means they figured something out, but you gotta start in something like customer support. Now, we support the models, so if you're a comp- person building a customer support product and your CFO is frustrated with the size of your ElevenLabs bill and you wanna cut it by 10X, go to speechify.ai, um, and or hit me up cliff@speechify.com. I'll give you some discounts. Um, but you j- you... Again, if you're a founder, you need to try. You cannot not try. You cannot give up before you're even in the race.

    13. HS

      What will be a bigger company in five years, Sierra or ElevenLabs?

    14. CW

      Brett Taylor has the best resume,

  17. 53:3455:18

    Sierra vs ElevenLabs: Who Becomes Bigger?

    1. CW

      I think, of anyone in the world, right?

    2. HS

      Mm-hmm.

    3. CW

      I think he s- started Google Maps, then he was CTO of Meta, then he was co-CEO of Salesforce. He's on the board of OpenAI, and now he founded Sierra. I would never f- try to fight Brett Taylor, and I think the field is so large. Like, no, like, we don't understand how big the space for AI agents is, like, not even, like, AI voice agents, not even close, in the same way that people didn't understand how big the field was for LLMs in two- 2019 and the same way people didn't understand how big the space was for AI coding agents in 2021. Like, this is the next huge space, and so both those companies are going to be massive.

    4. HS

      I think they're playing very different games.

    5. CW

      All right.

    6. HS

      I think, I think Brett Taylor's actually trying to recreate a next generation of Salesforce. He, he is absolutely not playing the customer support game. He's moving-

    7. CW

      Yeah

    8. HS

      ... into, into pre-sales.

    9. CW

      Everything.

    10. HS

      He's moving post-sales.

    11. CW

      But neither is ElevenLabs. ElevenLabs has a product that also does customer support, but they do everything else too. That's why I call it AI agents, not customer support.

    12. HS

      But I think Monty's building-

    13. CW

      Like ElevenLabs is not Finn

    14. HS

      ... a p- a, a very opinionated voice-centric company.

    15. CW

      Correct.

    16. HS

      It's voice cen-

    17. CW

      Oh, you think Brett Taylor is doing all of it

    18. HS

      ... and I think Brett Taylor is doing all of it.

    19. CW

      Put another way, if you use a tool like Sierra, the wedge right now is voice, but the important part is tool calling ElevenLabs lets you do some tool calling, but that's not the bread and butter. There was a really good presentation that Brett Taylor did a screen share of him building a guitar store on Shopify, and how he uses Sierra to do customer support and sales and everything else. It was extremely impressive. If you haven't searched this, you should search this. Brett Taylor is a big guitar guy. Um, that is a very different product than what ElevenLabs is doing. And so they're both going to crush. I agree with you on the Sierra, uh, uh, conclusion.

    20. HS

      What crazy thing today will be incredibly...

  18. 55:181:04:24

    Quick-Fire Round

    1. HS

      And this is quick-fire, my friend, 'cause I could talk to you all day. What crazy thing today will be very common in five years' time? You know, before it was like, find your partner online. Duh. No. Weird. Put your credit card online. Fuck no. What today is a no, and in five years' time we'll be like, "Yeah, of course"?

    2. CW

      Human computer interface is going to become primarily voice as opposed to a screen. So part of the reason why Google succeeded, it's, is a very simple interface. There's a text box and a button. That's it. Anyone can learn how to use it. The reason why ChatGPT worked as opposed to GPT-3 is because it was also a very simple interface. Just chat. There's a text box and a button. You get a response. That's it. The simpler version of that is just having a conversation. I say something, I hear something in response. If you use Voice AI from ChatGPT right now, it sucks. It's too slow. The LLM is much dumber than the core LLM. The escalation to the higher quality LLM is pretty weak. I think what will happen, and Meta has the right idea, by the way, so go Chris Cox, is people are going to be talking to their computer and phone and some wearable constantly throughout the day and using screens a lot less.

    3. HS

      You can buy one, SpaceX or Meta. Which do you buy?

    4. CW

      Meta.

    5. HS

      Why?

    6. CW

      Elon's distracted.

    7. HS

      Is he distracted or is he building full stack? 'Cause actually I think he's never been more strategically positioned, and he has an outlet for each of the different products that he's built, and each one feeds the next. When you look at Zuck and Meta, you know, bluntly the compute spend that he's producing, the outlet is increased conversion on an ads business, which is the biggest ads business in the world. So 7% on $240 billion is a lot of fucking money.

    8. CW

      Yeah.

    9. HS

      But it's actually not in the same quantum league as doing space data centers.

    10. CW

      Yeah. So let's take the space data centers out for a second. I think the space data centers is a very interesting idea, and what it does really well is it lets me underwrite a gigantic TAM for my expectation for SpaceX.

    11. HS

      It, it ruins all estimates. [laughs]

    12. CW

      Right? And so that's like, that was a great, uh, rabbit out of the hat by Elon-

    13. HS

      Yeah

    14. CW

      ... in order to pitch investors really well. Let's take that out for a second, and I'm going to talk to you about SpaceX and Tesla like they're one company, because really I'm assessing Elon, I'm not assessing every- like, you know, SpaceX as an individual stock. Um, you know, for data centers, the biggest constraint right now, right now, is memory cards, and then very soon it's going to be energy, and it's energy a lot of the times. So what do you need for energy? You need energy supply and you need energy storage. And so the best energy storage right now actually comes from Tesla. Tesla also has a chip manufacturer that they're doing, basically competing with everyone else, and like that's gonna do really well. And if you saw that, uh, Joe Rogan interview with Elon maybe two years ago, he was explaining that the hard part is not building the product. The hard part is building the manufacturing for the physical product, so Elon is number one in the world for manufacturing complex items like that. So like that's very exciting. Um, and so the TAM for Elon's companies are bigger, however, I think that Meta trades... What does Meta trade at right now? Less than SpaceX.

    15. HS

      Uh, it's less than SpaceX.

    16. CW

      Yeah.

    17. HS

      It's, it's fucking dumps.

    18. CW

      And, and so I, I think Meta has more data than anybody else in the world. I think Meta is actually super hampered by laws like GDPR. Like if GDPR didn't exist and the other laws in the US didn't exist, Meta would be ripping. They just can't train on their data properly. And so they'll figure that out at some point in some way. I don't know how, but I believe in Zuck. And at the end of the day, I'm a huge believer in founder-led companies. And so both-- We're talking about two of the best founders in the world. And the last thing I'll say, look at Zuck's age and look at Elon's age, and Zuck's not gonna stop and Elon's not gonna stop, but at a certain point, one of them will expire. Um, and so Zuck has like 20 extra years. And so depending on how long you're investing, I'm younger than Zuck, let's see what happens.

    19. HS

      If Zuck expired, Meta's stock price would increase.

    20. CW

      What? I disagree completely.

    21. HS

      Uh, no, because you'd have a CEO who comes in and understands that... And this may be a short term-

    22. CW

      Yeah

    23. HS

      ... but that's saying we're gonna invest more and more and more and more and more and more and more in CapEx-

    24. CW

      Yeah

    25. HS

      ... when we don't have an outlook for it.

    26. CW

      Yeah.

    27. HS

      You'd actually see a stock price appreciation in the short term. Every time Zuck steps out on the podium and says, "CapEx, CapEx, CapEx," he's like fucking hammered for it. Say I'm gonna-

    28. CW

      But that's why Meta is a good investment right now, because what Meta doesn't have is what Palantir has, which Palantir has the Alex Karp effect. Alex is really good at pumping up the PE ratio of the stock, and Zuck, y- I agree, is the opposite.

    29. HS

      It's the same as Elon. It's the Elon problem.

    30. CW

      Same as Elon. Exactly. And so-

Episode duration: 1:04:34

Install uListen for AI-powered chat & search across the full episode — Get Full Transcript

Transcript of episode hj5oRzAnp2M

Get more out of YouTube videos.

High quality summaries for YouTube videos. Accurate transcripts to search & find moments. Powered by ChatGPT & Claude AI.