Skip to content
Lenny's PodcastLenny's Podcast

How we built Grok Bot in a month | Roman Ugarte (SpaceXAI)

Roman Ugarte helped incubate and build Grok Bot, the popular new knowledge-work agent from SpaceXAI. A small, isolated team took it from first line of code to a working internal product in four weeks, and to a hugely successful public launch just three weeks later. Before Grok Bot, Roman led Growth at Cursor, where he helped scale the company from 15 people to over 1,000 before its acquisition by SpaceX. *In our in-depth conversation, we discuss:* 1. The origin story of Grok Bot 2. The key decision to build it from scratch instead of adding it to Cursor 3. Why the team personally onboarded nearly 300 of its first users 4. The two early product decisions that made Grok Bot so successful 5. Their “colleague-pilled” product philosophy 6. Roman’s advice on moats, and what has allowed Cursor to keep winning in the most competitive market in the world *Brought to you by:* WorkOS—Make your app enterprise-ready, with SSO, SCIM, RBAC, and more: https://workos.com/lenny Mercury—Radically different banking, now with Command: https://mercury.com/ *Episode transcript:* https://www.lennysnewsletter.com/p/how-we-built-grok-bot-in-a-month *Archive of all Lenny's Podcast transcripts:* https://www.dropbox.com/scl/fo/yxi4s2w998p1gvtpu4193/AMdNPR8AOw0lMklwtnC0TrQ?rlkey=j06x0nipoti519e0xgm23zsn9&st=ahz0fj11&dl=0 *Where to find Roman Ugarte:* • X: https://x.com/romanugarte_ • LinkedIn: https://www.linkedin.com/in/romanugarte • Website: https://x.ai *Where to find Lenny:* • Newsletter: https://www.lennysnewsletter.com • X: https://twitter.com/lennysan • LinkedIn: https://www.linkedin.com/in/lennyrachitsky/ *In this episode, we cover:* (00:00) Introduction (02:09) The origin story: building from scratch in one month (08:40) Why Grok Bot was built as a separate product (11:20) Manually onboarding a couple hundred people (14:29) Hiding internal mechanics from users (18:41) Timeline from beta to public launch (19:14) Unshipping features and simplifying (23:50) Early use cases and feedback (26:50) Product philosophy: “Grok Bot can now” (30:02) Cloud-first architecture (33:12) The fresh-start advantage (35:54) The vision: a true team of AI colleagues (39:20) The “colleague-pilled” framework (42:36) Work versus personal: one product or two? (47:14) Long-lived agents, persistent memory, and the computer abstraction (51:04) Grok Bot as an always-on infovore and chief of staff (53:35) How fast the team moves and what preserves the startup feeling (58:20) SpaceXAI pillars (1:00:44) The first 90% vs. the last 10% (1:03:30) Moving fast at scale (1:06:40) How Cursor kept winning in the most competitive market in the world (1:10:04) Company values: “deleting the product” and “just do the thing” (1:11:45) Moats: discovered, not planned (1:15:11) Tips for new users and power users (1:18:00) Lightning round and final thoughts *Referenced:* • Grok Bot for iOS: https://apps.apple.com/us/app/grok-bot/id6794501026 • Grok Bot for Android: https://play.google.com/store/apps/details?id=ai.x.grok.bot&hl=en_US • How I AI: Grok Bot + Grok 4.6—what’s great (and what’s still hype) & Lessons from spending $20,000 on Devin in one month: https://www.lennysnewsletter.com/p/how-i-ai-grok-bot-grok-46whats-great • Codex: https://chatgpt.com/codex • ChatGPT Work: https://chatgpt.com • Shopify: https://www.shopify.com • Salesforce: https://www.salesforce.com • The playbook for building high-talent-density teams | Adam Ward, Head of Talent at Cursor: https://www.lennysnewsletter.com/p/the-playbook-for-building-high-talent • Superhuman: https://superhuman.com • Stripe: https://stripe.com • OpenClaw: https://openclaw.ai • Listen: OpenClaw: A power user’s guide to the most powerful personal AI tool since ChatGPT: https://www.lennysnewsletter.com/p/listen-openclaw-a-power-users-guide • Hermes: https://hermes-agent.nousresearch.com • From skeptic to true believer: How OpenClaw changed my life | Claire Vo: https://www.lennysnewsletter.com/p/how-openclaw-changed-my-life-claire-vo • Cursor: https://cursor.com • Grok Build: https://x.ai/build • Roman’s post on X, “An AI that does 100% of the job feels categorically different from one that gets you 90% there”: https://x.com/romanugarte_/status/2087344044435505175 • Notion: https://www.notion.com • Casablanca: https://www.imdb.com/title/tt0034583 • Monk: https://www.imdb.com/title/tt0312172 • Exa: https://exa.ai • Desiderata - Words for Life: https://allpoetry.com/desiderata---words-for-life *Recommended books:* • Cat’s Cradle: https://www.amazon.com/dp/038533348X • The War of Art: Break Through the Blocks and Win Your Inner Creative Battles: https://www.amazon.com/War-Art-Through-Creative-Battles/dp/1936891026 _Production and marketing by https://penname.co/._ _For inquiries about sponsoring the podcast, email podcast@lennyrachitsky.com._ Lenny may be an investor in the companies discussed.

Roman UgarteguestLenny Rachitskyhost
Sep 8, 20261h 22mWatch on YouTube ↗

EVERY SPOKEN WORD

  1. 0:002:09

    Introduction

    1. RU

      The ultimate vision of Grok Bot is incredibly simple. You should have a team of AI bots that help you with your job and help you with your life.

    2. LR

      Grok Bot is the hottest AI product in the world right now. That is a very high bar. [chuckles] There's a lot of competition for that slot.

    3. RU

      We wanted to build something that wasn't just a great product for developers and engineers. We decided to create this very small team internally to go off into a cave for about a month with the sole objective of build an amazing knowledge work product that brings agents to the rest of the company.

    4. LR

      It's been only three weeks since launch. I went to a Grok Bot meetup. There were hundreds of people there, standing room only. It's very clear to me that you guys have built something very special.

    5. RU

      Once you start breaking out of, "This is AI chat with a set of connections," instead to, "This is a colleague with a computer," it just raises the ceiling of what you would think to give to AI.

    6. LR

      You have this tweet, "An AI that does 100% of the job feels categorically different from one that gets you 90% there."

    7. RU

      What made me so excited to work on Grok Bot is it was the first time for non-coding tasks that I felt like I could truly delegate work to AI, not have to think about it, and I would come back and it's done.

    8. LR

      What is it that you think you did that is so different, that made Grok Bot so successful?

    9. RU

      It was two early decisions that at the time definitely did not feel obvious, but in hindsight I think are critical to what makes Grok Bot work.

    10. LR

      Today my guest is Roman Ugarte. I'm gonna keep this intro very short so we can get right into it. Roman was employee number 15 at Cursor. He was head of growth for the last two years. Most recently, he helped incubate Grok Bot, a product that I am obsessed with. It has changed my life. I use it 100 times a day for all kinds of things, and I think it's safe to say it is the hottest and most exciting new AI product in the world right now. Roman leads product for Grok Bot. He's been part of the core team from early prototype until today, and we get into how it all started, where it's all going, and all the things that he and his team have learned since it launched just a few weeks ago. With that, I bring you Roman Ugarte.

  2. 2:098:40

    The origin story: building from scratch in one month

    1. LR

      [gentle music] Roman, thank you so much for being here, and welcome to the podcast.

    2. RU

      Thank you. It is great to be here.

    3. LR

      I am so excited to have you here. I am so hooked on Grok Bot. I have it over here in my window. I have, like, 15 bots that I use every day, all the time. Uh, I went to a meetup the other day, a Grok Bot meetup. There were hundreds of people there, standing room only, people sharing all the ways they're using Grok Bot. It's very clear to me that you guys have built something very special. It's very hard to break through the noise in the AI world. Uh, Grok Bot is the hottest AI product in the world right now. That is a very high bar. [chuckles] There's a lot of competition for that slot. I personally noticed I've moved a lot of my use cases from CoWork and Codex into Grok Bot just, like, just, like, very quickly, which, again, feels like a really big deal and a very special moment. Uh, and so I'm excited to talk about so many things. I wanna understand how you guys did this, where this came from, uh, where this is going, what you've learned about the journey so far. Um, first of all, just nice job. Nice, nice work. This is very hard, what you've done.

    4. RU

      Thank you. I mean, I remember onboarding you, uh, by hand about a month ago, and I think you were skeptical at first. Uh, but we're very glad that you've been using it, and it's been great to see so many people, um, really take advantage of Grok Bot.

    5. LR

      I am gonna talk about that onboarding. Uh, that was a very interesting, uh, element of how this worked. Uh, I actually remember in that onboarding, I tried it. You asked me to do, like, a, "Let's try something," and I was like, "Okay, try to t- come up with a tweet to promote my latest podcast episode." And so I'm just like, uh, "Come up with a tweet to promote my last episode." That's it. And it was actually very good. It figured out what the hell the last episode was, how to promote it. So I actually remember in the moment being like, "Wow, this is really good." So let's actually start with the origin story. Where did this start? What was kind of the original idea, and when did the, the work on this begin?

    6. RU

      Hmm. Yeah, it started really as a blank page, completely from scratch, build from zero exercise, where I think we'd been feeling for a long time that we wanted to build something that wasn't just a great product for developers and engineers, which is really where we started, and I think we've gained a lot of intuition about how to build great agents and useful products that way. Uh, but what would that product look like for knowledge work? And we decided to kind of create this very small team internally, it was really just a handful of people, uh, to go off into a cave for about, for about a month [chuckles] with the sole objective of build an amazing knowledge work product that brings agents to the rest of the company. And I think from the first line of code to when we released this prototype internally, it was only about a month. It was, like, a very quick, um, you know, scrappy prototype that was pulled together. And I think in hindsight, this would not have been possible if it had been, I think, a much bigger group. I think it took a small focused group that was completely isolated from the rest of the company, and I mean that literally. It was, like, a separate part of the office where this team sat, um, private Slack channels. And the goal, and I think in hindsight it was a lot of what allowed us to move so quickly on this, was we needed to make a lot of micro decisions every day, uh, some things that maybe we'll talk about a bit later, that were not obvious, were not really things that we'd done on other product surfaces before. And I think if it had been a very big group of people and we were kind of thinking about this six to 12 month-long vision, we just wouldn't have really gotten to the place that we ended up landing at. And so that was about a month from first line of code to here's a functional, useful product that the core team is excited about. And so then there was a moment of rolling this out across the company, rolling this out across all of SpaceX AI. Uh, and so there was an all-hands where we kind of shared the progress that had been made so far. There's this brand-new product. We would love for you to use it. And I think what was most encouraging, because at that point we were excited about it, I think we were using it constantly, but it's easy to use the thing that you built, and you kind of understand the mechanics and what it's good for. And so this was a real pressure test with reality, like are people actually going to switch from other internal tools, other external tools, to use GroqBot as their primary agent surface? Um, and in that first week after that all-hands, um, I mean, I can't even tell you just the outpour of, of love for GroqBot from people that maybe you wouldn't expect, or from groups of the company that maybe you wouldn't expect. Uh, people who, you know, were daily driving ChatGPT or like a chat interface, switching all of their, you know, day-to-day agentic tasks over to GroqBot as like their primary surface for doing work. And so the internal reception was, was really extraordinary. Um, can tell some kind of funny stories from that, that week or two period. Um, and then once we saw, I think, the internal reception, we immediately switched into, "Let's get this ready for the world. There's a lot of work to do to scale this out to millions of users." Uh, and then that led to the, the GA launch that we had a few weeks ago.

    7. LR

      This episode is brought to you by our season's presenting sponsor, WorkOS. What do OpenAI, Anthropic, Cursor, Replit, Sierra, Clay, and hundreds of other winning companies all have in common? They are all powered by WorkOS. If you're building a product for the enterprise, you've felt the pain of integrating single sign-on, SCIM, RBAC, audit logs, and other features required by large companies. WorkOS turns those deal blockers into drop-in APIs with a modern developer platform built specifically for B2B SaaS. Literally every startup that I'm an investor in that starts to expand upmarket ends up working with WorkOS, and that's because they are the best. Whether you are a seed stage startup trying to land your first enterprise customer, or a unicorn expanding globally, WorkOS is the fastest path to becoming enterprise ready and unblocking growth. It's essentially Stripe for enterprise features. Visit workos.com to get started or just hit up their Slack where they have actual engineers waiting to answer your questions. WorkOS allows you to build faster with delightful APIs, comprehensive docs, and a smooth developer experience. Go to workos.com to make your app enterprise ready

  3. 8:4011:20

    Why Grok Bot was built as a separate product

    1. LR

      today. Okay, so many questions. One that is really interesting here. So obviously there's Anthropic OpenAI. They went from, they had this coding agent that they're like, "Holy shit, this is a big opportunity." And then they're like, "Okay, people are using this for knowledge work. Let's build a knowledge work component." So there's Cowork evolved out of that within the product, and then Codex they've invested in, "Let's make this useful for all kinds of things." Interestingly, you guys decided, "Okay, Cur- we're not gonna build this into Cursor. We're gonna start something fresh." Was that, was that just like obvious from the beginning, "Okay, this is not gonna work inside Cursor, the product. We need to start fresh"? How, how, like, controversial was that decision?

    2. RU

      It was not obvious at all.

    3. LR

      Mm.

    4. RU

      I think you're completely right that that was one of those original decisions that, uh, at the time we had a lot of discussions about, and I'm very glad with where we landed. And I think to your point, it being a brand-new product that you control every pixel of the experience, and you have this consistent vision about where knowledge work is going, and it's all contained in this, in this new thing, I think has, has contributed a lot to the success. Uh, but there were a lot of discussions about, you know, Cursor, for example, and some of our coding products. People use it for non-coding tasks all the time. Um, and it, you know, these coding agents are really excellent at some of these things. But you run into small paper cuts. Sometimes the product itself is kind of intimidating to non-technical users. There's a brand association with these things. And so I think we evaluated that path, and I think we saw what maybe some of our competitors have been doing of this is all just one surface. You add new tabs for each new form factor, and it feels a little cluttered. And I think for users, they can feel that, that this was not a single, consistent, uh, vision of the way that work should work. And instead it's three different visions that all kind of share a screen, and you can hop between. But it is kind of a shipping your org chart style thing that I think users are reacting negatively to. And so we decided, let's just start completely from scratch. Let's see where we can get from there. There might be some really amazing opportunities to bring people from other surfaces into this more bot native experience, but it's really important for people to just have, like, an amazingly simple and amazingly powerful experience.

    5. LR

      That is a really valuable lesson for people to take away here, just that that is, might be the solution instead of adding in complicated and existing AI product. So interestingly, Codex went the other direction [chuckles] and it, and it's like a different path and a different product, but it's interestingly, they're like, "No, we're gonna help make it one thing." So there, you know, there's many ways to make it work, and it also feels like the path you take there will kind of lead you. But maybe, maybe we'll look back and be like, "That was not maybe

  4. 11:2014:29

    Manually onboarding a couple hundred people

    1. LR

      the best idea." Something you mentioned that you did that is also really unique is this onboarding of early users. Uh, I heard you and your team onboarded two to three hundred people manually, including me. Uh, talk about why you thought that was necessary and what you learned from that experience, and just, like, how long that period was of this kind of manual onboarding.

    2. RU

      I mean, you just learn so much. And, um, the first few onboardings were pretty painful. I'm glad you got a good one, Lenny [chuckles] . Uh, but there were some that were kind of rough. And we learned a lot, and I think it was important for the core team to be in the room for those and to just sit on a call for 20 minutes when the computer isn't spinning up or when someone's in onboarding, and they're just incredibly confused. So that immediately after you're like, "That can never happen again. We need to solve this tomorrow, because tomorrow I'm onboarding this person, and it needs to go better." And so there was about a- ... two-week period, uh, where we were in that mode and onboarded a couple hundred people. Um, and not only did we learn a lot about the product, I think we didn't really know, uh, I think sometimes with these products there's some groupthink of ways to use them. And I think internally, because people inside of SpaceX AI were just constantly sharing tips and tricks for how to use Grok Bot, some patterns were starting to emerge that we thought would be useful to the world, but we weren't really sure, and we definitely didn't want to bias the world.

    3. LR

      Is there an example of that?

    4. RU

      So when we rolled out Grok Bot internally, there was about a week or two where the common pattern of the way people would interact with the product was you would have five to 10 bots, and each bot you would give a different scope, a different domain, and it was kind of shorthand for different lanes of work. And then at around the end of week two, we started to see these messages internally in Slack of people promoting one of their bots who was a bit of a standout performer. [chuckles] And it was, like, their kind of primary personal assistant promoting that to their chief of staff. And then they would actually mostly talk to their chief of staff, and the chief of staff would fan out all of these tasks to the other bots and would kind of manage the team. And y- there are some funny screenshots of people, like, actually telling their, uh, you know, the bot they're promoting that they're, they're promoted, and the bot is asking if they get a raise and, you know, is their token budget higher, all of these things. Um, and we kind of took note of that, and I think more of the company started to slightly shift in that direction, but it was not the majority of the company. People use this product in very different ways. And so in some of the onboarding sessions, and just from the early access program in general, we really did not wanna lead the witness and say, you know, "Create a chief of staff bot. Here's exactly the way that, uh, you know, that chief of staff should manage all of the other bots," and see if early access users would get there themselves. And we actually did see that many of them did. And so then we had a bit more of an opinionated take in the product of this feels like a pattern that's working, this feels like a pattern that we should slightly encourage, but it shouldn't be a one-way door. And there are a few other examples of internal theses that we really wanted to make sure, you know, would bear out in actual external usage without us imposing

  5. 14:2918:41

    Hiding internal mechanics from users

    1. RU

      that in the product.

    2. LR

      Is there anything else there? Any, uh, examples come to mind?

    3. RU

      Yeah. I think another thing we really tried to pay attention to in the early onboardings, um, was just how much users wanted to see. And I think it is something a bit shocking or just different about Grok Bot when you first start using it versus some of the other products that you mentioned, um, where a lot of the internal mechanics of how Grok Bot works are not shown to the user. And the reason for that is we think as these models get smarter, the same way that your teammate, you know, you wouldn't ask for second-by-second updates of exactly all the buttons they're pressing and websites they're going to, I think it's too much to ask your bots to do that, too. And I think it's honestly just overwhelming, uh, and can create more harm than good. And so we moved completely in the other direction of you send a message, you tell your bot to do something, it just starts doing it. It sends you progressive updates as it sees fit, and you just see that, like, typing indicator [chuckles] and the little green, you know, active, uh, active, um, circle kind of Slack-like that it's, it's active, it's doing work, it'll get back to you soon. But you don't see the internal mechanics. You don't see the tool calls. You don't see exactly every little click it's making on its own computer. And we really wanted to take a strong stance that users did not need to see all of those mechanics, and so that's where we started. And we did get some feedback that's like, "I would love to see my bot's to-do list. I would love to see roughly how it's prioritizing tasks and what it's doing." And that's great feedback. But it was useful to hear that nobody wanted the, like, long stream of just text streaming out and chain of thought sequences. So that also gave us more, uh, confirmation that that was the right direction.

    4. LR

      The fact that you did 200 to 300 onboarding calls on a, with a small team, I knew that, I know the team grew, uh, over time, but just that is a huge time commitment, and then you could argue a distraction from the building. Clearly not a distraction. Clearly a core part of the success. Uh, do you feel like that's, like, that's the volume people need to do to figure out what actually needs to happen?

    5. RU

      Well, one thing to emphasize is the early access group is not necessarily just people that are highly influential tastemakers. You know, you're in this category, Lenny, and we certainly wanted to get a lot of your feedback, um, just from being very close to many other products on the market and just being a power user of these things. But we also wanted to get, uh, early access to kind of more unconventional profiles that we as a company had never really interacted with. So one example is, uh, there's a coffee shop owner that was a friend of a friend through the company who had heard about Grok Bot and, you know, uh, one day somebody on the core team had kind of shown them a, a demo of the test flight and got very excited. And this coffee shop owner ended up being not only an amazingly, uh, an amazing power user of Grok Bot, but also a rich source of feedback for us. [chuckles] Uh, we have a very lively thread with, with many, many, uh, bugs that get identified or feature requests, and it's a completely different use case of running a small business. And so, for example, if, you know, the Shopify, uh, integration was a bit flaky, or if it wasn't writing copy, uh, for products in, in a particular way, we would get really rich feedback on that, which is pretty different from the type of feedback we'd get from dogfooding this internally. And so I think it was an, an important exercise for us to check our blind spots and say, A, this is going to be a very general product that is not just a thing that developers use. In fact, it is likely that this is most powerful for non-developers. We need to understand that group much better. And then B is we absolutely live in this kind of Silicon Valley AI bubble, uh, which I think is, is a useful place to be to kind of push the, the frontier and push the future of, of how these products are evolving. But we need to actively get out of that because I think a product like this has the chance of really being the way that the mainstream user and the mainstream, uh, kind of business customer can interact with AI in a way that's useful.

  6. 18:4119:14

    Timeline from beta to public launch

    1. LR

      Let's go back to the timelines real quick just to kind of- ... understand that. So it was a month from first line of code to internal beta, and then what happened after that?

    2. RU

      It was about three weeks from internal beta to public launch, and then I think we're about three weeks out from public launch as of recording this.

    3. LR

      Wow. Okay. So a month of building the first thing, three weeks only of iterating, and that, and then it's been only three weeks since launch. It feels like it's [chuckles] changed the world in my, from my vantage point. So,

  7. 19:1423:50

    Unshipping features and simplifying

    1. LR

      wow. Okay. What most changed in those, I don't know, in those three weeks of internal beta, let's say?

    2. RU

      We un-shipped a lot. Um, I wish I could have shown you what things looked like maybe two weeks out from launch, where I think we'd realized that the core team, we had a lot of experimental features that we wanted to get internal feedback on, uh, which was useful. We also were kind of putting pseudo developer-y, like, visibility tools into GroqBot instead of having, like, a separate observability pane, uh, for those things. So for example, we actually did expose sometimes a lot of the internal thinking of the models and the specific memories it would store and all of these things, uh, which was useful to debug issues, and if you're building the product, you didn't wanna go somewhere else to maybe pull that context. But we had to really aggressively trim what we think the user absolutely needs to see in the surface versus what they don't. I think there's even more room, uh, there to run, which is something the team's focused on right now, is how can we just ruthlessly simplify this product and abstract away anything the user doesn't need to actively be thinking about? Uh, so that was one big push, uh, was un-ship a lot of, a lot of jank. Um, and then I think the second big push those few weeks was making it just work. And I think a lot of what people want from AI is not this thing with lots of dropdown menus and bells and whistles, but just a thing that you describe a task, a task that's meaningful to you, and it goes and it does it, and it comes back with complete work, or it comes back with something for you to react to and then steer it in its kind of next cycle. Um, and so in order to actually deliver on that promise, it's actually not a lot of feature product roadmap style stuff. It's like hill climbing five really important problems in the back end that many users don't directly experience, but you totally feel, uh, when, you know, your bot is going off and doing something and can't click the right button, or your bot is off and doing something and can't log into a website, and it just completely stalls your ab- your ability to make progress on that task. And so those few weeks, we collected a really rich set of what are tasks that people actually are giving bot? How can we quantify these things? And how can we see week over week that across those categories of tasks that we're hill climbing on very important dimensions to making that just work behind the scenes?

    3. LR

      Is there an example, one of those, uh, hills you were climbing that was kind of a technical breakthrough or technical challenge you overcame that really helped it just work?

    4. RU

      Yeah, one example was from rolling out GroqBot, one group inside of the company that was actually incredibly bot pilled, so to speak, um, was our go-to-market team, was sales. And there are a bunch of tools that sales uses, um, that do not have well-supported MCPs or APIs. And I think that's a lot of what made bot so before and after powerful for this group, was these were things that they just could not give another AI tool reliably. We would get stuck at some part in the process. And then bot kind of felt like they had an assistant, or kind of felt like they onboarded someone to their personal team. They gave it a laptop, and it could just run. And so there were a bunch of small things and, you know, probably a list of, of 10 or 20 of them, of places where for whatever reason, you know, the mouse would just not have fine enough control to click on exactly that part of the Salesforce dashboard or something like that, um, that we would have to take back to the, the core team working on really the infrastructure to say, "Here's a very concrete case of where, uh, the agent not having this visibility into the browser or this visibility into the pixels on the screen is making it impossible for this task to be done." And that was just a lot more tangible than seeing a number on a dashboard slowly creep up. It was kind of like new chunks of work getting unlocked, and you would immediately feel the feedback, where you would ship an improvement, uh, that was kind of behind the scenes, kind of infrastructure-y, and then the next day you would just get this outpour of, you know, love and appreciation from the sales team that now this workflow that was failing the last seven days finally works. And it's just a constant exercise of finding those next tasks to unlock and then solving them.

    5. LR

      So s- uh, computer use, uh, basically, uh, improvements seems like a big unlock.

  8. 23:5026:50

    Early use cases and feedback

    1. LR

      Uh, I heard also the, um... I had Adam Ward on the podcast-

    2. RU

      Mm-hmm

    3. LR

      ... uh, who's head of recruiting, head of hiring, basically head of talent. Uh, I heard his team was, like, one of the top users of, of, uh, GroqBot.

    4. RU

      Yes. The recruiting team, um, gave us a lot of great feedback. Anytime the product, especially in the early days, if there was a little bug, we'd get a, get a ping from, um, some folks on the recruiting team. Um, yeah, I think the main use cases for recruiting that were particularly interesting, um, was first it was incredibly valuable as a sourcing tool, and I think one thing Adam talked about on the podcast with you, uh, and it's a big part of our hiring philosophy internally, is be looking for a job, being on the market is not a precondition for us trying to hire you. [chuckles] And in a lot of ways, um, you know, the best way to hire is really just look at- The biggest problems at the company that need someone to, to own it or take it to the next level, find out of the total universe of people in the world who would be best, and then ruthlessly go after them and try to convince them to join. And this is a lot of the philosophy from the very beginning of the company. And so if that's your mindset, really the best recruiting work stream is not, or, or workflows to automate are not, you know, here are a bunch of resumes, read through them, help sort them. The most useful thing is here's a, an entire universe of potential people. Help match that to this very concrete business problem or this very concrete role that we're recruiting for, and help me get in touch with them, help me get coffee with them. Let's just throw everything at it. And so there have been some cases of, um, really kind of, uh, unexpected ways of finding top talent that is, is beyond even just, you know, looking on LinkedIn and trying to find, uh, interesting people. But who are the co-authors of this paper, and the PDF doesn't exist on Google Scholar, it just exists on this conference website. I want you every morning to go to this conference website, download the PDFs. If there are any new ones, you should find every new name that we've not yet tracked. You should add that name to a spreadsheet. You should do research. You should look at everybody at SpaceX AI, see if there's anyone directly connected. If so, you should send them a Slack message asking for an introduction. Like, it's those types of always on sourcing, um, use cases that I think in the past were incredibly manual, and now it's the type of thing AI is superhuman at, and our team can focus on closing great candidates and getting conversations with great candidates and not pulling these huge lists.

    5. LR

      Wow, that is such a cool example. Uh, first of all, someone's about to take the transcript of what you just said, put it into a bot, and create their version of this, which is great. On, on, on the other hand, I think you guys could sell a template of this bot for a billion dollars.

    6. RU

      [laughs]

    7. LR

      If we could basically use, uh, Adam's, uh, team strategy for finding the best people and just turn that into a bot. Holy moly. [laughs] Democratizing, uh, hiring.

  9. 26:5030:02

    Product philosophy: “Grok Bot can now”

    1. LR

      I wanna come back to a few things. Okay. So one is you made this point about unshipping. Such a important point, I think, that people can overlook because AI is not good at telling you what to take out. It's very good at, okay, here's more, more ideas, here's more stuff. And something that's come up a number of times on the podcast is that's a big opportun- that's a big space for humans to continue to be very important and valuable, is knowing what not to ship and what to cut and what not to do. And so it's so interesting to hear that that's been a big part of the internal evolution of from prototype to launch is deciding, okay, we need to cut a bunch of stuff. Anything more there?

    2. RU

      Yeah. So one thing we talk about internally is for anything that we're working on for GroqBot, what is the launch post? Like, what is the thing that we would actually tell users? And if it's not interesting-

    3. LR

      A launch tweet, I imagine

    4. RU

      ... launch tweet. And if it's not compelling, maybe we shouldn't be working on it, um, if, if it's not something that users will directly feel in the product. And to take that even one step further, um, I think there is this old school software tendency to say things like, "Grok Bot now has..." And when you think of completing that sentence, it would be like a new button to press, or it'd be a new dropdown, or it'd be a new integration that you can press plus and add. And instead to reframe it as, "Grok Bot can now..." Which is, is I think a much more human way of kind of describing these capabilities. And I think it's forced us to think more in the frame of what are tools and, and what are capabilities that we can give GroqBot, not what are new things we can add to the product. Like, adding things to the product is not the goal. That's not the thing that's gonna push this, this product forward and make it more useful to more people. Making your bots reliably do really impactful work for you behind the scenes in a way that just works, and giving them the capabilities to do that, like, that's what users actually care about. Uh, and so I think in the context of unshipping, there have been a lot of GroqBot now has things, uh, that we've realized are actually just capabilities that don't need pixels. You know, let's kill as many pixels as we can. Those can just be things that your bot manipulates behind the scenes for you, and you don't need to directly control. And I think one example of this is the way that many of our competitors, you set up automations or routines, is you go into a sidebar, you press plus, you select, you know, what the trigger event is. Uh, you then select, uh, you know, what action it should take after that. You might describe it in natural language. And it's just really clunky, and it means that people don't set up many automations for, for many things. We certainly have seen this in, in the coding realm. And so I think what GroqBot did in response to that was actually you should just define automations in natural language, and you should tell your bot, "Remind me that at 8:00 AM every day, please." And then it should just do it, and you should never, ever have to see that interface of creating an automation. And so that's kind of the decision that, that we've made, and now that's how 99% of automations on the platform get built. And I think there are a bunch of other places

  10. 30:0233:12

    Cloud-first architecture

    1. RU

      where we can do things like that.

    2. LR

      So I tweeted about how much I love GroqBot when it launched, and a lot of people replied, they're like, "Wait, can't you just do all this with Codex and CoWork?" And you can. Technically everything, as far as I know, you can do with GroqBot, you can do with the other foundational models, the coding assistants. So let me just ask you this big question. What is it that you think you did that is so different that allow- that made GroqBot so successful?

    3. RU

      I think it was two early decisions that at the time definitely did not feel obvious, but in hindsight I think are critical to what makes GroqBot work for people. And the first is- You should never have to think about local and cloud and where are these workflows running? Does my computer have to be awake? If I kick it off from my phone, does it need to be tethered to my computer back at home? Like, there's so much jank happening right now when people are trying to conceptualize where this runtime lives, and we made a really early decision that this should just all be in the cloud. And if it's in the cloud and it's this persistent colleague that has its own computer, it can do its own work, it has the same state everywhere you interact with it, it opens up a lot of really amazing opportunities to text your bot, kick it off from your phone. In the future, you should be able to call your bot from anywhere, and it should be able to do real work. Like, this is its own entity, and it lives separately from your device. And I think that was a very important decision that current products, uh, I, I think haven't made that same decision, and I think it, it has a bunch of paper cuts as a result of it that users feel every day. I think the second decision was kind of to take that one step further of not only should this be an agent loop that kind of runs in the cloud and you can interact with in various ways, it's actually really important that these bots have their own computer. And part of it is what I described earlier of there are a bunch of tasks that don't have well-supported MCPs and APIs. We as humans don't do our jobs via MCPs and APIs. Like, we use a computer and we click on pixels and we kind of type things in input boxes, and it's very important that your bot has those baseline capabilities as well. Uh, but even to go one step further, I think we're in a really weird moment right now that I think we're gonna look back on and be like, "I'm surprised that this is the way that a lot of people worked with AI," where you're onboarding these super intelligent new colleagues, these AI bots, and you're asking them to share the same computer that you have. It's crazy. Like, if you were onboarding someone to your team and you said, "It's your first day. I'm gonna onboard you. Uh, you don't have your own laptop. You're gonna sit next to me. We're gonna share this laptop forever and constantly trip over each other. You're gonna have access to my credentials. I'm gonna have access to your credentials." Like, that's just th- th- there's a good reason why that's not the way people operate, and I think bots and these kind of AI colleagues of the future will also need, um, you know, a way to onboard them that's, that's somewhat similar, and we've tried to push the product in that direction.

    4. LR

      That is so funny. [chuckles]

  11. 33:1235:54

    The fresh-start advantage

    1. LR

      Um, why do you think the other companies didn't do this? My guess is they w- were building off of their existing coding assistant platform and approach, and this was, like, a pretty big shift.

    2. RU

      I think a lot of it comes down to starting from scratch and how freeing that is, and we felt that ourselves, where a lot of the primitives in Grok Bot we had attempted or we had built in other ways. Um, you know, cloud infrastructure we'd built for coding agents. Um, the way that you could maybe name agents and talk to them as discrete entities, it's a pattern that we're also seeing for developers, uh, kind of bringing specific named agents into Slack. But instead of kind of trying to retrofit those concepts into some new surface or into an existing surface, which I think would've been the strong default, I think, for many companies, we decided to start from scratch and we decided to just try to get these things really right for general knowledge work, which is a new audience. And then second, for the point in time that we're at now where the models are very capable, and if you give them the right tools and the right infrastructure, they can do a lot. But a lot of these ideas, these aren't strokes of genius on our part, and I, I think for good reason. I think these are primitives that had already been getting attention and product market fit by other products, you know, the OpenClaws of the world. Um, and I think we took a lot of inspiration from that and tried to productize it into a bit of a tighter surface, something that required a little bit less setup and I think was more accessible to more people. And so I think our competitors and other tools that have been trying to solve these types of problems, I think we're all seeing the same opportunity. I think we're seeing a lot of those, the feedback from the market. But I think it's just been hard to act on if you're stuck in the existing paradigm, and if you have a lot of kind of sunk cost in that existing paradigm, it's very painful to create a new thing from scratch. Uh, and I think a lot of that is what allowed the product to just work and click for so many people.

    3. LR

      So what I'm hearing here is the keys to success of what made this break out, a cloud-based i- computer for every, uh, bot instead of locally, uh, a name kind of like specific bot. And by the way, there's this like we're all moving from agents to bots now. Nice job. Like, [chuckles] that was feels like you guys have pushed the, pushed it over now. Okay, we're all bots now. So a bot per kind of task use case, very unique, versus like a thread conversation or something or, like, a s- one-off job. And then it feels like there's, it just works was a core part of this, and you talked about how long it took to get to that place of like, okay, now it actually works really well.

  12. 35:5439:20

    The vision: a true team of AI colleagues

    1. LR

      Um, you mentioned OpenClaw. Obviously, this is inspired by that, which to me, when I first used OpenClaw, I'm like, "Holy shit, this is the future. How could we not have this?" Uh, and then Hermes came out and everyone's been trying to build the OpenClaw that works very easily for everybody. Can you say more about just, like, how OpenClaw and that story informed the way you guys thought about it?

    2. RU

      Yeah. So I think OpenClaw got two major things right that when we were seeing the way the market was reacting to OpenClaw and ourselves using the product, we found quite exciting. I think the first thing was the models are really smart, and they're gonna continue to get smarter. But even at current capability levels, if you can just give your bot access to the tools that you do, that you use to do your job- Um, it can get a lot of the way there. In a lot of the places where people think AI is, is dumb or, uh, may- maybe not as impactful, um, as it's been promised, a lot of that I think is downstream of it just being harnessed in the wrong way. And so if you kind of give access to a much larger set of things, if it has access to its own computer, how far can you go? And I think OpenClaw, you know, really kind of forced that question for many people. And then I think the second way OpenClaw changed the mental model of, of AI, uh, was really viewing these things much more as colleagues and teammates and people, and personifying it a bit more, and it being this helper entity that has access to your life and can kind of extend you even further. And so we took a lot of that, and I think what Grok Bot maybe extended was, A, it needs to be really easy to set up. And the hacky, you know, you have a VPN at home and a Mac Mini set up, clearly it was not going to scale to millions of users. Clearly it's not gonna be the way, importantly, that businesses take advantage of this technology. And so we really wanted to build an amazing product with that in mind. And then second is, I think there are a lot of places, a lot of rough edges to sand down and just make a delightful product experience, and make these things just work, and try to remove some of the abstractions that power users of AI are very familiar with, things like skills, for example. Um, how can we make a Grok Bot user not even have to know what a skill is? They should never have to type a slash command. These things should be created in the background as a useful primitive that the bots have access to. But something that users, you know, it's not incumbent on them to always be on the, on the cutting edge of AI. And so that's really where we tried to, to innovate, and I think there's still more room to go there.

    3. LR

      Uh, as you say that, I have my Mac Mini with my formerly live OpenClaw on there, and, uh, that was an era. And, uh, it's so awesome the work that it has inspired. I know it continues. I know there's still a lot of value to OpenClaw, but when I saw Claire Vo, who's been, like, the biggest proponent of OpenClaw and has... It's, like, become a core part of the way she lives and works with her kids and does all her work. She just switched all of her OpenClaws. She shut them all down and switched to Grok Bot. That's a huge... Like, it sounds funny, but that's actually a huge, uh, mi- milestone of just how much things have shifted. Um,

  13. 39:2042:36

    The “colleague-pilled” framework

    1. LR

      what's kind of the, the vision for Grok Bot? What's like... Where does this go? What does this look like in the future? What's, like, the ideal platonic version of Grok Bot?

    2. RU

      The ultimate vision of Grok Bot is incredibly simple, which is you should have a team of AI bots that help you with your job and help you with your life, and it should really feel like a team. It should really feel like teammates that are autonomous are helping you. You can steer them in various ways. You don't have to micromanage them. They have access to the tools necessary to do great, ambitious work. And one thing we really use as a North Star on the product side, uh, in building this is as we kind of get closer to this teammate future, how can we, in every product decision we make, think about this less from the perspective of a SaaS product and more from the perspective of we're trying to build useful AI teammates? And so there have been a bunch of examples where we kind of have to push ourselves to be more, like, colleague-pilled in a way. We sometimes use that term where we're having a product debate about something. There are good arguments on one side. There are good arguments on another side. Both paths feel sensible. Like, in product land, this maybe doesn't feel, uh, like there's a clear-cut answer. And then you zoom out a little bit and you remove yourself from the, you know, tech company-ness of it all, and you start thinking, how would a human do this? Like, what would you want from your teammate in this exact situation? And oftentimes the answer is really clarifying and pretty unanimous. There's oftentimes not a lot of disagreement among the room of, like, how a human teammate you would prefer to work with in a certain way. And then once that answer is there, well, then we just need to build it. And there are product implications, there are model implications. There's a lot of things that need to go right to actually deliver on that experience. But in some ways, it's not rocket science. It doesn't require you being a genius. You just need to ask the question of what would you want from a human teammate, and can we push AI to behave in a similar way? And so to give some examples of that, I mean, we've been thinking about what the right voice experience with, uh, these bots should be. And I think in the context of a human, for example, we have a really good analog of a lot of times I'm Slacking back and forth with a teammate. We're sharing context. And a lot of times it's just much simpler to get on a five-minute huddle with them and just press huddle, talk back and forth. I share my screen. I show exactly what's on my mind. They share their screen. We hop off, and then we continue async from there. And that's not really an experience that any AI product has gotten right right now, and it is deeply integral to the way that I think humans collaborate. And so we want to build something like that. And there are a bunch of other examples of these, like, very clear patterns that just work, uh, that I think you should also feel when working with AI.

    3. LR

      I love this term colleague-pilled. Uh, such... A- and it's come up so many times over the course of this chat already how that is, uh, kind of a through line to making these decisions. For example, the computer example you gave is so good. Uh, obviously people would have their own computer. The naming piece is also a very important part of that.

  14. 42:3647:14

    Work versus personal: one product or two?

    1. LR

      A big question on my mind in this space, and I'm so curious to get your take, is the separation between work and personal. Do you think people will have two different assistants, a work and a personal, or do you think it'll be one?

    2. RU

      When people think about a work product versus a consumer product, I think there's just a lot of baggage that comes from the last decade or two of horrible B2B software [chuckles] um, that leads to people seeing a product that is very simple. In some ways, ChatGPT was like this. I think GroqBot has many of these properties. And assuming that it's not a work product, or assuming that it's not a power tool. And when you think of a power tool in this kind of last generation, I, in my head, picture something a bit like Photoshop, for example, where there are all of these different dials to turn very precisely. You know, the user of the tool is this kind of, um, you know, ultimate, um, cockpit flyer who knows exactly what all the knobs do and can, like, use them perfectly. And I think power tools of the future will actually be very different from that, uh, where it is mostly just intent being expressed and good steering on the part of the human, and these AI tools abstract away all of the knobs. You should never see them unless you need to directly manipulate it, which might happen, and there should be a great affordance for that. But ultimately, it really is just working with a teammate, and so the interface for that is quite conversational. And so in a lot of ways, GroqBot, when you look at it, like when I walk by someone's desk and I see GroqBot up on their computer, for me, for a split second, I'm like, "Oh, are they on, like, a messaging app?" And it's like, no, they're... You know, this is actually the primary tool that they're using to do much of their work. And so I think to your question of are you going to have a different set of bots for your personal life and a different set of bots for your work life, um, I do think there will be a separation. For many people, they want a separation between personal and, and work life, and I think that's, I think that's great. I think that's important, and I think there are a lot of common sense reasons why those things should be separate, even from the perspective of, of an enterprise. But I think our goal and the thing we're trying to build towards is GroqBot should be the way that a large portion of the things you do day-to-day in your work, you should be able to delegate a lot of that to GroqBot and focus on the higher leverage things. And then it should similarly be the way that you delegate a lot of the low leverage parts of your personal life. And those two things actually are not different problem sets. In a lot of ways, the product form factor and the ways of solving those problems is pretty much the same. And so my instinct is that I think one product will be the best form factor for both of those things, and that's really what we wanna build.

    3. LR

      Bam. That's a big tam right there. Uh, I love, I love to hear. It makes so much sense. Obviously, the question is how do you avoid cross-contamination, you know, personal stuff somehow, uh, infiltrating, exfiltrating stuff from work. Uh, but it feels like that's kind of okay. So what I'm hearing is that's the direction. The question is just how to do that and make people feel super safe, have kind of like the SOC 2 stuff in place, and also just feel really fun. This episode is brought to you by Mercury, radically different banking now with Spend. I've been a Mercury customer for so many years now. I switched all my business banking to Mercury, and honestly, I could not be happier. It's what online banking feels like when it's built by product people, not by bankers. And now with Spend, you can give your team individual cards, set spending limits per person or per team, and have expense receipts automatically pulled in from Gmail or over text. You can even give your AI agents their own cards with their own limits and policies. Most founders start out the same way, one card used by everybody at the company. It works until it stops working. Someone goes over, a receipt disappears. You spend two days trying to figure out who spent what and why. Spend is expense management built directly into Mercury. All your team's cards, budgets, and reimbursements all live in the same place as your business banking. No chasing, no manual reviews, no end of month scramble. The result is a team that can move fast and a founder who is no longer the bottleneck. Learn more and get signed up at mercury.com. Mercury is a fintech company, not an FDIC-insured bank. Banking services provided through Choice Financial Group and Column A Members FDIC. The IO card is issued by Patriot Bank and a member FDIC pursuant to a license from MasterCard International Incorporated.

  15. 47:1451:04

    Long-lived agents, persistent memory, and the computer abstraction

    1. LR

      Let me ask a couple technical questions. Uh, on the computer side, how do, how do, how-- What's the simplest way to think about what you get as a part of your GroqBot account? Is it like AVM that is running in the cloud with multiple logins? Is it like a separate VM instance per bot? How do we understand that as much as you can share?

    2. RU

      Mm-hmm. Yeah. I think to go back to the teammate frame of the product, um, to kind of extend the analogy even further, if we were on a team together, you and me, um, you know, I think the number of times that you would have to manually take over my computer and start clicking on things and like, you know, "You're doing this wrong. You should go here instead," and type in manually, hopefully is pretty close to zero. [chuckles] Hopefully, that is not something, uh, you have to really do with, with a colleague or a teammate. And so similarly, I think right now we're in a place where computer use is good. It's getting much better. And in very short order, I think the computer concept will be completely abstracted away from the user. You should never be clicking into a remote virtual machine. You should never have to take control, uh, you know, of something if there's like a wasteful path and you have to kind of steer it into the correct path. Um, so in the medium term, I think the computer concept will be an important concept for users to have, but will not actually be something that they're interacting with. So I think the right way of thinking about GroqBot is it's a team of bots. It's a team of agents that are, that do work for you. And in terms of what they have access to, they have access to a very long memory set of your past interactions with them. And so I think there's a current paradigm if you create a new chat for each discrete unit of work, uh, I think there are a lot of problems with that. I find myself copying and pasting between chats all the time. Um, I think it's just not a great way of grouping categories of work. Instead, the same way on a team you have a good way of grouping, uh, categories of work of kind of roles, you should have roles of kind of different swim lanes of, of work that you do, and it should learn from you, and it should get smarter over time. So I think that's one very critical thing is these are long-lived agents. These are not individual one-off sessions, and these agents get smarter over time. And then the second thing is those agents have access to all of the tools that you would expect a human colleague to have, which is the APIs, the MCPs, that's great, but then access to its own computer, which it can freely manipulate the way that you would.

    3. LR

      Uh, at this meetup that I went to, uh, Shub, who's on the, I think, growth market team, uh, demoed something that blew everyone's mind. Because you have a computer within each agent, you can run a lot of different things on a computer. He was running GroqBot within GroqBot. Like, the bot can run its own GroqBots. [chuckles] And I know he was using it for testing and watching regressions and things like that, but that's just like a mind-expanding idea, and I'm curious how many levels you can go before the universe collapses on itself.

    4. RU

      I do that one too. That one's actually a very useful thing to do is you download GroqBot for one of your bots. Mine is like a QA tester bot. Uh, and that way if there's ever bug report or if we're kind of testing out a new, a new build, for exam- example of the desktop app, I can just say, "Hey, here are 10 workflows that we need to make sure are getting better release after release. I want you to test it. I want you to write it to this Notion document that has, like, an extensive list of all of the past tests that we've done of past client versions, and compare them." And so I think once you start breaking out of this is AI chat with a set of connections, which is I think where most people are conceptually now, instead to this is a colleague with a computer, and anything I would ask a colleague to do on a computer, I can ask GroqBot to do. It just raises the ceiling, I think, of, of what you would think

  16. 51:0453:35

    Grok Bot as an always-on infovore and chief of staff

    1. RU

      to give to AI.

    2. LR

      Are there any other mind-expanding use cases or ways to use GroqBot that, uh, you've seen that or you use?

    3. RU

      Mm-hmm. One pattern that I've seen from many users, um, that is simple but I think there's a lot of depth, uh, if you keep investing and making it better, and this is kind of where I can get kind of nerdy about optimizing my setup, um, is GroqBot as an infovore in some ways of just consuming huge quantities of information, removing that from your own cognitive load, giving you peace, and then coming to you with the stuff that's important. And I think the V1 implementation of that, which many people do, is GroqBot sits on top of Slack, and it sits on top of email, and I tell it high level, "Here's my role at the company. Here's kind of what I care about. Uh, I want you to notify me in these cases. In these cases you don't need to ping me directly, but you should include this in your daily roundup that I read every day." That's like the V1 implementation. I'm not sure what the V10 implementation is, but, like, maybe I'm at V3 or 4, which is you can give these bots a complete fire hose of information. So I have mine hooked up to, like, every mention of GroqBot ever on X, and it's interacting with our internal context. It's interacting with the QA tester to, like, see if it can repro any bugs or, or feedback that we're getting. I've hooked up to my own kind of messaging services to, like, quickly act on feedback and reach out to people. Um, and I think there's this just, like, always on kind of chief of staff entity that can preserve your focus on the things that actually matter, but is always, like, s- kind of surveilling to see if there's anything that should get your attention. And we've seen some funny cases of people actually giving their GroqBots, which I have not done this yet, but maybe soon, giving their GroqBots access, uh, the ability to page them. And so if something, like, super urgent happens and they're at a coffee or whatever, they get paged by GroqBot, which is the type of thing that you only wanna do if it's urgent and you really wanna, you wanna trust, uh, that GroqBot, you know, does not have false positives. So far those people have reported, uh, that it, that it's been very helpful and successful. But I think we're gonna see more of that type of stuff, um, where the agent or the bot should actually be more proactive to you than you reaching out to it, and I think that will be the next

  17. 53:3558:20

    How fast the team moves and what preserves the startup feeling

    1. RU

      shift in AI.

    2. LR

      This touches on-- There's a number of things that I've been very impressed with watching your team operate. Uh, one is speed, which I wanna talk about, but the other is how you're... Like, there's awareness that this is a moment in time to capture a lot of, uh, market share and really, uh, take as much of the market as you can before somebody comes around and they're like, "Okay, now we got something awesome," especially one of the foundation labs. So watching just how many free accounts you guys are giving out. Also the, uh, focus on use case is so smart because it's such a novel thing, and you open it up and it's like, "What do I do with this?" And there's such a focus on, "Okay, here's a bunch of things people do with it," and then there's all this talk on Twitter and just, like, all the ways people are using a template. Makes so much sense, just these two kind of like focuses from what I can tell. Get as many people on it as possible as fast as possible until somebody's like, "Okay," you know, 'cause someone's gonna come around and be like, "All right, here's the next thing." Uh, super smart, and also the use case focus. Uh, I know you're a go-to-market person at Cursor before this. Anything you wanna share there about just the approach to the, to go to market right now for getting this out there?

    3. RU

      I think the pattern we saw for coding will be somewhat similar to what we see for general knowledge work, and I think we've learned a lot from that on the go-to-market side and more- Generally, um, just building practical AI that people use. And I think we as a company have culturally really cared about not building demoware, uh, like building actually useful stuff in the world, uh, and kind of obsessing over that. And there are so many shiny objects and, like, fun prototypes to build, but ultimately that's a very different problem than getting this in the hands of millions of people and having it transform companies. So that's really, I think, culturally where we've, we're, we've always been focused. And so I think on the go-to-market side, what we saw for coding, um, was a very simple pattern, which was there was an early adopter crowd. The early adopter crowd would use these coding tools and really push them to the limits, and they would mostly push them to the limits on individual projects. They would, on nights and weekends, I'm thinking like 2023, you know, kind of earlier, um, people would kind of go home from work. At, at work they were using a basic IDE, this is pre-AI, and then at home they'd work on a side project and they'd be using Cursor or they'd be using, you know, the latest and greatest AI coding tool. And that would give them an extreme amount of acceleration. It would feel like they were experiencing the future, and then they would come back to work and they would demand it. [chuckles] They would say, "I cannot picture working any other way than this. I feel like I'm completely walking through molasses right now. This needs to change." And I think for knowledge work, we're gonna see a similar pattern of people really feeling the aha moment, sometimes in a personal capacity, and I think we're certainly seeing a lot of this, like, on X right now. You see all of these examples of, uh, Grok Bot controlling their home robot computer, or home robot vacuum cleaner, uh, or Grok Bot, you know, helping them save money on their Tesla charger negotiation. Like, all of these fun use cases. But I think the next step is going to be this is not a consumer product. We think this is gonna transform businesses. We think this is gonna transform teams, and it will be bots coming into teams and contributing really economically valuable work, especially as they get much smarter. And so on the go-to-market side, we're certainly, um, making a big push on prioritizing businesses and thinking about not just the single player use case of working with a single bot, but how does a bot work inside of a broader team? How does a bot work inside of real company systems, uh, that are complicated and there's a lot of context and a lot of history to understand? What does memory look like in kind of a broader, uh, organization versus kind of a single individual you're catering to? And I think there are a lot of unanswered questions there, but I do think Grok Bot is the right primitive to create this switch to agents, uh, for the rest of the company outside of coding, and that's a place where we're quite focused right now.

    4. LR

      And along those lines, it's very clear you all understand the power of distribution and how y- you need to find both an amazing product and get distribution right, because, you know, Grok Bot's amazing, but the combination of how smart you guys have been with getting it out there in all these different ways is really impressive, and I think that shows you what it takes these days to build something that's really successful.

  18. 58:201:00:44

    SpaceXAI pillars

    1. LR

      I want to ask about the brand of, the different brands around, uh, this product and the company, just so people can try to understand, 'cause I know you're going through a transition, acquisition, SpaceX, all these things. So there's Grok Bot, there's Cursor. Is that... So talk about, like, the products and the way to think about these different brands today, and I know it'll probably continue to evolve, just so we could, uh, communicate a- about it correctly.

    2. RU

      Definitely. Yeah. Um, I think there are three big pillars right now of SpaceX AI. So the first pillar is the coding product and set of products, and right now that's Cursor and Grok Build. And I think we're big believers that having a professional work surface for developers and for the engineering part of the organization is going to be really critical. And right now people use Grok Bot sometimes to kick off cloud agents or to kind of merge PRs or to do QA, a bunch of engineering adjacent tasks. But ultimately, when you're shipping production software, we're big believers that that is gonna require, you know, a product where every pixel is optimized for that end user. So we're making big investments there. The second category is general knowledge work, and we think Bot is a really exciting step in that direction. I think there's a lot more work to do of making it more useful, extending it to new surfaces, it really feeling like an AI teammate that you can delegate work to, especially inside of companies and businesses. So that's kind of the second pillar. And then third is the general mo- model effort. Um, we wanna train the smartest models in the world, um, that are really capable. And I think one thing that somewhat distinguishes SpaceX AI from other AI labs, um, is I think our goal is less to build, um, you know, chase super intelligence or some kind of vague aspirational ideal. Uh, and the goal is actually very practical, which is to build useful AI, and we do that on the product side, we do that on the model side. And I think part of that is also just cultural of, like, the group of people contributing to these models are engineers and people who, um, kind of came into the, the model training effort from, like, a very applied mindset. And I think that's what gets this company going, and I think is, is actually a slightly different direction from some of the other competitors out there.

    3. LR

      Super interesting.

  19. 1:00:441:03:30

    The first 90% vs. the last 10%

    1. LR

      Okay, there's a couple directions I wanna go. One is you have this tweet that is, uh, I think you pinned it, or maybe it's your last tweet. It's up there in your timeline if people check you out. So the tweet is, "An AI that does 100% of the job feels categorically different from one that gets you 90% there. I've significantly updated what I think AI is capable of." Say more about that

    2. RU

      I think for me, what made me so excited to work on Grok Bot and contribute to it is it was the first time for non-coding tasks that I felt like I could truly delegate work to AI and not have to think about it, and I would come back and it's done. And I think engineers have been feeling this for quite some time. For maybe a year, a year and a half, things have been like that. I mean, the job of a developer has completely transformed. Uh, it is unrecognizable from what it was two years ago, and, and many, many words have been said on that topic. Uh, but I think it's underrated how different that experience is from what most people are feeling about AI right now and the way that AI has changed their lives. And it looks quite similar to the way that people would use AI, like, two years ago, where you create a new thread for a task, you type it into a, to a, uh, you know, an input box, you hit enter, you watch all of these steps happen, you get an output. It's not quite right, you keep working on it. And Grok Bot, I think, short-circuits a lot of that. When you first kind of s- lay eyes on the first screen, you're like, "Whoa, this is clearly different. Let's see if it actually works, but this is different." And then you give it something, and it kind of works, um, you know, to a sup- surprising extent. Uh, and I think we're gonna do a lot to, to make it work much better. And so I think what I was expressing in that was when you have a teammate that you only 90% trust, uh, and you give something to, which luckily I do not have the experience of here 'cause I work with great people. Uh, but if you delegate something to someone and you're like, "I know I'm gonna have to be thinking about this while you're doing it, and I know it, like, probably is not gonna be quite there and I'm gonna have to intervene and kind of steer it slightly," you're not... That's not 90% task completion. You're still doing the thing, and it feels that way and it's, it's weighing on you in the same way, versus, like, truly throwing a no-look pass to a colleague and being like, "You got this. Here's the context. Go off and run. I'm excited to see what you do." Like, that's a different category, and I think that's the type of thing that people feel with Grok Bot every day are these no-look passes, and you just trust that it can get it done, and then it does, and it's just a very magical experience.

    3. LR

      Yeah. I've had that experience consistently. Um,

  20. 1:03:301:06:40

    Moving fast at scale

    1. LR

      okay, so another element of how y'all operate that has really impressed me, and I've not seen this before, is how fast y'all move. So I got added to this, like, Slack as you... I, I was giving feedback with some folks, and it's just like, "Okay, how about, okay, tomorrow we're gonna give you some free codes to give out. You could do it tomorrow. We'll do this tomorrow." Or, uh, "We're gonna launch a marketplace with templates. We're gonna launch this in two days." I was just like, "What? I don't f- have [chuckles] I don't have time for this. How do you guys... With all the things going on, all these things constantly shipping, and also staying consistent and high quality and feeling clear that it's towards a specific vision." So there's kind of like two parts to this question, just what, what's the secret to how fast you all have been moving, and how do you stay aligned moving that fast towards a vision of, that you all want, that you all believe in and where you want it to go versus just like, you know, Band-Aiding it along the way?

    2. RU

      Mm-hmm. Yeah. One thing I've been really happy has never changed is that startup feeling inside of the company. And for context, when I joined Cursor originally, we were about 15 people. We scaled to about, to over 1,000, um, and then now we're a part of, of SpaceX AI, which is kind of an even bigger organization. And it's something that is just so fun to be a part of when you're around this group of incredibly talented people. Everyone's moving 100 miles an hour. You trust, you deeply trust everybody to execute on their part of the equation, and there's a clear vision, uh, that everyone is fired up about and, like, and knows that, that they need to execute on. Um, and you know, as companies grow, and we've had the fortune of hiring really great people from other companies that have gone through hypergrowth, things slow down and you kind of keep telling yourself, "We're still a startup. We still move quickly," but you really don't, and everyone knows that you don't, and it's just, you know, it's easier, uh, to say than to actually be. And you know, fingers crossed this, this continues to be true. I think it's really critical for our success if this continues to be true. Uh, but even as we've scaled, it has always felt like that startup that kind of I, I first joined. Um, and I think if you define a startup by number of people or by, like, the funding round, like, none of those things really make any sense. The core thing that defines a startup is exactly what you're describing, which is this kind of scramble energy of things are kind of chaotic and kind of disorganized, and, like, for a lot of people, that's not a pleasant working environment to be in, but it has these amazing properties of you can make extreme impact in a particular direction in a short amount of time, and you really do get out of a system what you put in. And so I think as a culture, I think as an organization and the way we construct ourselves, uh, it's really been to enable that property in a way that I think some of our competitors and, and other AI labs have gotten much bigger, and you can feel it. Um, and I think us, even as we scale, there is that, that startup-y impulse that is, is quite important to move quickly on these things.

    3. LR

      Let me pull on this thread, and let me ask you this big question that I've been looking forward to asking you.

  21. 1:06:401:10:04

    How Cursor kept winning in the most competitive market in the world

    1. LR

      If you, if you were to look at Cursor from the outside, you, it shouldn't have worked. It shouldn't have, uh, lasted because, one, it's in the most competitive market in the world, competing against the fastest growing companies in history, OpenAI and Anthropic. So that's one. It's like the com- competition is unlike anything anyone's ever experienced. Two, it sits on top of those platforms [chuckles] to power it. And what I've seen As an outsider is what has allowed Cursor to win and have this massive exit and continue to succeed, is how quickly you all adjust to the reality of the market. Started as autocomplete, and then things moved on to just talking to agents, and then into the cloud, and now GroqBot. To me, that feels like a core part of the success is, uh, quickly adjusting to reality and also building the best-in-class experience for a thing that also exists other places. GroqBot's a great example. You could do this other places, but it's the best-in-class experience. Cursor, the ID, the best way to code. Uh, so, so that's my question. Maybe I answered it, but what do you think has been core to Cursor's ability to not just survive in this crazy competitive market, but, uh, do so incredibly well consistently for so long?

    2. RU

      I mean, a lot of this, and it's a fuzzy answer, a lot of this I think is downstream from culture, and the culture that you set, and the people that you bring in, and the way that they approach these problems. And I think for us, exactly as you said, we have never been complacent. Uh, we've never felt like we've won, and it's always been about the next thing. And I think there's been a really deep belief across the company that AI is moving incredibly quickly. Our goal is to translate those capabilities into amazing products for customers, but those products are going to change, and they need to meet the moment as the capabilities get stronger. And what met the moment two years ago is completely different than what's meeting the moment today. And if we as a company can't completely reinvent ourselves every six months, which recently it's felt even shorter than that, of kind of complete, like, very significant reinventions of our priorities, the core product, what users feel, uh, we're gonna lose. And I think it's that spirit of always pushing to be on the frontier, never thinking it's over or that we've won or that we've gotten it right, uh, and just constantly updating our beliefs that has gotten us to where we are now. And to your point on the competitiveness of this space, I mean, one thing that I think is important to point out is AI coding has always been competitive from when Cursor first, first kind of came to be. And at the time, the competitors were Microsoft and others, and a handful of maybe 10 or 20 companies. Um, and I think it's notable that none of those competitors are at the forefront of AI coding right now, in large part not because of any incorrect decisions that they made or any, uh, lack of resources on their part, but this cultural inability to move quickly and to change to meet the moment as the moment's changing. And so I think that's exactly, uh, what has led us to invest in things like GroqBot, for example.

  22. 1:10:041:11:45

    Company values: “deleting the product” and “just do the thing”

    1. LR

      Are there any, uh, core values, just like specific ways you phrase this to kind of remind everyone of this is how we work?

    2. RU

      Yeah. Two values that I find myself coming back to quite a bit. The first one is this idea of deleting the product, and I think it exactly ties back to what you're saying right now, where when you look at every past iteration of Cursor, for example, but even I think when you look at every past iteration of GroqBot, I think we will feel the same thing, uh, is things going away, not new things getting added. But, you know, these scaffolding product overhang style things that get built in because the models have not yet gotten smart enough to just do it themselves. Those things will get moved away over time, and we need to feel comfortable making kind of hard decisions that might upset a small set of users or a small set of us internally to do the bigger thing of make the product simple, make the product powerful, and adapt to where the future is going. So I think that's been very core. And then the second thing, which ties back to the kind of pace of execution, um, is just do the thing, uh, which I find myself kind of repeating a lot, um, even as we've kind of grown as a company, is it's on you. You know, we're all in this boat together. We wanna win, and if you see something that you think needs to happen, you know, this is not an ask for permission culture. You go out and you fix the thing, and you pull in the resources that you need to make it happen. And I think that has made many people very successful here before, and I think it's something we really share with, with SpaceX AI as well.

    3. LR

      Hmm. Agency, as, uh, you may have heard.

    4. RU

      Yeah.

    5. LR

      Um, so interesting.

  23. 1:11:451:15:11

    Moats: discovered, not planned

    1. LR

      One of the big questions that comes up and, and, uh, and this might be my final question, is around moats. And a lot of people look at Cursor as a really interesting example of they're in a market with, and technically maybe no moats, but they've continued to win and succeed. The two moats you think about with Cursor is the data feedback loop of people auto-completing, learning what they're doing, and training models based on that. So that's, that's unique. The other is just best-in-class experience and being like a high da- daily active user product and finding over time what works and what people need. What have you just learned? And I guess any thoughts on moats in the space that might be helpful for folks that are trying to figure this out for themselves.

    2. RU

      Yeah. There's a lot of talk about moats, and it is a pretty interesting moment in time to be starting a company, so I can understand why so many founders are kind of asking themselves that and trying to project out 12 months from now, 24 months from now. It just feels like an eternity. I will say that I think if Cursor and many other successful companies of this kind of vintage, I think if they had thought about moats/ kind of tried to work backwards from some strategy diagram or like, you know, a maybe more abstract notion of how a company should work, I don't think that would have created this outcome or this product. I think what really created the magic of Cursor, uh, was an obsession with building a useful thing today. And I think it was constantly this exercise of you can kind of see where the world is going three months from now, six months from now. Models are gonna get smarter. A thing that isn't solvable now is finally gonna be solvable. And I think Cursor was a little bit this recurring prompt of how could we pull that stuff to today, even if it requires a little bit of, of engineering on top to make it work, or a lot of engineering on top to make it work. Even it requires changing the product in a specific way so a user can interact with this new capability. How can we bring that forward? And then three months from now, we should delete all that stuff because it'll just be good and, like, basic, you know, common bare minimum of the product, and then we'll build the thing for three months from then. And then it was constantly just doing that over and over again that I think led to users really trusting us and placing, you know, their, their time inside of our product and trusting that we were kind of bringing things to this next frontier and to the next future. And so I would really encourage many founders or people starting out today to be more grounded in that perspective of how can I make something that is not possible now possible? Users are gonna come to me to use that thing. I'm gonna pull them to the next impossible frontier, and then through all of that, I'm gonna gain a lot of distribution advantages, I'm gonna gain data advantages. There will be value there. Uh, but I think that's really the place to play.

    3. LR

      I love that answer. Essentially, uh, s- like the way I'm thinking about it is just build something people are obsessed with, don't overthink the moats piece, and if you can continue to do that, you'll find something, which in Cursor's case ended up being a few things.

    4. RU

      Completely.

    5. LR

      Uh, and that, that came up actually recently on another podcast I did. I don't know if it'll come out before or after this, this idea that moats are discovered, not planned ahead of time a lot of times.

  24. 1:15:111:18:00

    Tips for new users and power users

    1. LR

      Okay, let me actually ask you for GroqBot tips, uh, as an actual last question. Uh, some people are gonna be like, "Oh, shit, I gotta try this thing. What are all these, what's all this excitement all about?" Um, what would be some advice for folks that are trying out, let's say for people that are new to it, just like, "Here are some keys to success," and maybe some power tip for someone that's already with it and just like, "Oh, wow, I didn't know that."

    2. RU

      Yeah. I wanna stay away from the, you know, super hacky pro tip stuff because I think our philosophy as a team and as a company is that those things really shouldn't exist. There shouldn't be all these crazy knobs. You should be able to delegate something to GroqBot and it should do it. Um, and so what I would encourage for someone new who's, like, just downloading the app, you're looking at this, this screen, um, I think the first thing is give GroqBot the context it needs to be successful. So in a similar way as if you were onboarding someone to your team, it'd be really helpful for them to have access to, you know, Slack and your email and the company records that you use every single day. So I'd give it access to the tools, and then I would actually ask GroqBot what it can do for you, and let it kind of go through those connections that you've initially set up. In my case, it might be I give it my email, I give it Slack. And I was really surprised when I was first onboarding. This is-- We didn't have any onboarding screens at this time, so this was kind of the first task I gave it, was, "Go through my Slack, go through my email, and suggest, like, five things that you can take off of my plate, and what it would take for you to do that." And it suggested five, and, like, two of them were actually really helpful, and I just immediately spun off two bots to solve those two. Um, and that was my big wow moment of feeling like no other AI tool in the past could have done those two things. It was not like draft an email. It was like do a chunk of work. And so I'd encourage people who are brand new to do it that way. And then for people who are not brand new, um, I kind of am constantly finding new patterns for ways that my bots can interact with each other and can collaborate with each other. And so I've been creating a bit more of a scaffold of kind of where these artifacts that GroqBots create should live and how it can write to a place that's very legible to me. So I have, like, these frequent digests that I read every day, and it kind of pushes to a database, and I can just read it very easily. Um, and so I would encourage power users to think about ways that GroqBot can actually write to, like, a, a single store where you can organize a lot of its outputs much easier.

    3. LR

      Damn. Uh, we need another episode of going deep on Groq- Roman's GroqBot setup, which probably has way too much private, sensitive information we couldn't show it. But that's okay. That's an amazing tip.

  25. 1:18:001:22:40

    Lightning round and final thoughts

    1. LR

      Roman, is there anything that you wanted to share or anything else you wanted to touch on before we get to our very exciting lightning round?

    2. RU

      Nothing. Nothing else on my side.

    3. LR

      We covered so much ground. That was-- I can't believe that was only an hour and a half-ish. Uh, I felt like we've [chuckles] been talking for ages and covered everything I was hoping to cover. With that, we've reached our very exciting lightning round. I've got four questions for you. Are you ready?

    4. RU

      I am ready.

    5. LR

      What are two or three books that you find yourself recommending most to other people?

    6. RU

      Yeah, so two, two books for you. One is I love Kurt Vonnegut, uh, so Cat’s Cradle has been a fun, fun recommendation that, and a copy that I've bought many friends before. Um, and then second is The War of Art by Steven Pressfield, uh, that I find myself frequently coming back to, even if it's just a page or two at a time, um, and I'd recommend for anybody.

    7. LR

      War of Art, incredible. It's, like, such a short book, and it's like once you read it... And it's not The Art of War, which is what people might think they're hearing.

    8. RU

      Yes.

    9. LR

      But it's The War of Art. It's a play on that, and it's about the, uh, the challenge of being creative and doing, creating something new and how to overcome the resistance. Uh, I love that recommendation. Uh, next question. Favorite recent movie or TV show if you've had any time to watch any of these things?

    10. RU

      Yes. Um, recently, so every year I do a watch of Casablanca, which is one of my favorite movies, and it incidentally also has a character with my last name, Ugarte, which is, like, the only example of, I think, a Ugarte in the media. Um, so Casablanca, always a great rewatch. Um, and then on the TV side, um, I sometimes sneak in an episode of Monk The detective show, uh, which was one that I watched kind of as a kid with my family and I've come back to now that I live in San Francisco. And it's just a great moment in time snapshot of San Francisco in the, you know, late '90s, early 2000s when it was shot, um, that I, I really enjoy.

    11. LR

      First Monk reference on the podcast. Okay. Uh, favorite or most interesting AI product right now. You can say Grok Bot if you want, but if there's anything else, you get bonus points.

    12. RU

      Um, I've always been a big AI semantic search nerd. I love any SemSearch product, especially the kind of out of the ordinary ones. [chuckles] Um, so I w- I was like a very early user of, of Metaphor at the time, which became Exa, and I love kind of using Exa to do all of these maybe more strange queries over the internet. But I see a lot of examples of people building like semantically search over, you know, an embedded image store of the MoMA or kind of things like that, and I always have so much fun playing with those. So anything semantic search engine over like a weird dataset, I love.

    13. LR

      And Exa in particular is one you'd recommend?

    14. RU

      I love Exa, yeah.

    15. LR

      Mm-hmm. Very cool. Okay. Favorite life motto that you often come back to in work or in life?

    16. RU

      Not as short as a single motto, um, but I love the Desiderata, uh, which I don't know if you've read.

    17. LR

      Mm-hmm.

    18. RU

      But, um, I have it on my, on my door, uh, and I've, I've had it since I was a teenager, and everywhere I move, I kind of paste it there. And it's, it's a very short poem, but each line I just, I find myself finding something new in it every time I read it, and I find it really grounding.

    19. LR

      Roman, this was amazing. Uh, what a, what a point in time. We're here right now at this moment in time of Grok Bot, of AI in general. Uh, it's gonna be really fun to revisit this, I don't know, in a year and be like, "Wow, we were so right and so wrong about so much." Uh, thank you so much for doing this. I know it's a very busy time [chuckles] on your team right now, so really appreciate you carving out a couple hours to chat. Um, is there any place you wanna point people to? Anything you wanna plug other than check out Grok Bot? Is that-

    20. RU

      Check out Grok Bot, of course. Um, and yeah, main thing would be please send feedback. I think we're in the very early innings of this still. I mean, we released a beta three weeks ago, um, and a lot of the feedback that we've been getting from this early set of users is directly translating to what we built and how we build it. And so really appreciate, um, all the input that people are giving.

    21. LR

      Nice job, Roman, and team. I know there's a whole team-

    22. RU

      Yeah

    23. LR

      ... behind all this. Uh, Roman, thank you so much for being here.

    24. RU

      Awesome. Thanks, Lenny.

    25. LR

      Bye, everyone. [upbeat music] Thank you so much for listening. If you found this valuable, you can subscribe to the show on Apple Podcasts, Spotify, or your favorite podcast app. Also, please consider giving us a rating or leaving a review, as that really helps other listeners find the podcast. You can find all past episodes or learn more about the show at lennyspodcast.com. See you in the next episode.

Episode duration: 1:22:42

Install uListen for AI-powered chat & search across the full episode — Get Full Transcript

Transcript of episode maSdsTLaMuU

Get more out of YouTube videos.

High quality summaries for YouTube videos. Accurate transcripts to search & find moments. Powered by ChatGPT & Claude AI.