Go back

How we built Grok Bot in a month | Roman Ugarte (SpaceXAI)

82m 43s

How we built Grok Bot in a month | Roman Ugarte (SpaceXAI)

Brockbot was conceived as a revolutionary AI product that shifts the paradigm from AI chat with limited tools to a team of persistent, autonomous bots acting as human colleagues. Built by a small, isolated team in just one month, it prioritizes simplicity, reliability, and real-world functionality over complex features. The core innovation lies in giving each bot its own cloud-based computer, enabling it to perform tasks that require pixel-level interaction—like clicking in dashboards or navigating software—without user intervention. Early onboarding of hundreds of users, including non-technical profiles, revealed natural work patterns, such as bots being promoted to manage teams and handle daily tasks. This validated the product’s potential beyond coding, proving its value in knowledge work. Key decisions—like removing local device dependencies and hiding internal mechanics—reduced user friction and made interactions feel more natural. The product’s success stems from its focus on user experience, not feature bloat, and its vision of AI as a true teammate. In the future, bots will become proactive, monitoring workflows, surfacing urgent alerts, and operating autonomously—offering a seamless blend of work and personal life management. This approach positions Brockbot not just as a tool, but as a new standard for how AI will integrate into everyday human work.

Transcription

15610 Words, 82908 Characters

English
The ultimate vision of Brockbot is incredibly simple. You should have a team of AI bots that help you with your job and help you with your life. Brockbot is the hottest AI product in the world right now. That is a very high bar. There's a lot of competition for that slot. We wanted to build something that wasn't just a great product for developers and engineers. We decided to create this very small team internally to go off into a cave. For about a month with the sole objective of build an amazing knowledge work product that brings agents to the rest of the company. It's been only three weeks since launch. I went to a Brockbot meetup. There were hundreds of people there standing room only. It's very clear to me that you guys have built something very special. Once you start breaking out of this is AI chat with a set of connections. Instead to this is a colleague with a computer. It just raises the ceiling of what you would think to give to AI. You have this tweet and AI that does 100% of the job feels categorically different from one that gets you 90% there. What made me so excited to work on Brockbot is it was the first time for non-coding tasks that I felt like I could truly delegate work to AI. I have to think about it and I would come back and it's done. What is it that you think you did that is so different? That made Brockbot so successful. It was two early decisions that at the time definitely did not feel obvious, but in hindsight I think our critical to what makes Brockbot work. Today my guest is Roman Ugarte. I'm going to keep this intro very short so we can get right into it. Roman was employee number 15 at cursor. He was at a growth for the last two years. Most recently he helped incubate Brockbot, a product that I am obsessed with. It has changed my life. I use it a hundred times a day for all kinds of things. And I think it's safe to say it is the hottest and most exciting new AI product in the world right now. Roman leads product for Brockbot. He's been part of the core team from early prototype until today. And we get into how it all started, where it's all going, and all the things that he and his team have learned since it launched just a few weeks ago. With that I bring you Roman Ugarte. Roman thank you so much for being here and welcome to the podcast. Thank you. It is great to be here. I am so excited to have you here. I am so hooked on Brockbot. I have it over here in my window. I have like 15 bots that I use every day all the time. I went to a meetup the other day at Brockbot meetup. People sharing all the ways they're using Brockbot. It's very hard to break through the noise in the AI world. I personally noticed I've moved a lot of my use cases from co-work and codex into Brockbot just like very quickly. Which again feels like a really big deal in a very special moment. And so I'm excited to talk about so many things. I want to understand how you guys did this. Where this came from? Where this is going? What you've learned about the journey so far. First of all, just nice job. Nice work. This is very hard, which you've done. Thank you. I mean, I remember onboarding you by hand about a month ago. And I think you were skeptical at first. But we're very glad that you've been using Anna. It's been great to see so many people really take advantage of Brockbot. I'm going to talk about that onboarding. That was a very interesting element of how this worked. I actually remember in that onboarding, I tried it. You asked me to do like this try something. And I was like, okay, try to come up with a tweet to promote my latest podcast episode. So I'm just like, come up with a tweet to promote my last episode. That's it. And it was actually very good. It figured out what the hell the last episode was. I had to promote it. So I actually remember in the moment being like, wow, this is really good. So let's actually start with the origin story. Where did the start? What was kind of the original idea? And when did the work on this begin? Yeah, it started really as a blank page completely from scratch, build from zero exercise. Where I think we'd been feeling for a long time that we wanted to build something that wasn't just a great product for developers and engineers, which is really where we started. And I think we've gained a lot of intuition about how to build great agents and useful products that way. But what would that product look like for knowledge work? And we decided to kind of create this very small team internally. It was really just a handful of people to go off into a cave for about, for about a month with the sole objective of build an amazing knowledge work product that brings agents to the rest of the company. And I think from the first line of code to when we released this prototype internally, it was only about a month. It was like a very quick, you know, scrappy prototype that was pulled together. And I think in hindsight, this would not have been possible if it had been, I think a much bigger group. I think it took a small focus group that was completely isolated from the rest of the company. And I mean that literally. There's like a separate part of the office where this team sat, private Slack channels, and the goal, and I think in hindsight, it was a lot of what allowed us to move so quickly on this, was we needed to make a lot of micro decisions every day. Some things that maybe we'll talk about a bit later, that we're not obvious. We're not really things that we'd done on other product surfaces before. And I think if it had been a very big group of people and we were kind of thinking about this six to 12 month long vision, we just wouldn't have really gotten to the place that we ended up landing at. And so that was about a month from first line of code to here's a functional, useful product that the core team is excited about. And so then there is a moment of rolling this out across the company, rolling this out across all the SpaceX Act. And so there was an all hands where we kind of shared the progress that had been made so far. There's this brand new product. We would love for you to use it. And I think what was most encouraging, because at that point, we were excited about it. I think we were using it constantly. But it's easy to use the thing that you built and you kind of understand the mechanics and what it's good for. And so this was a real pressure test with reality. Like, are people actually going to switch from other internal tools, other external tools to use Grogbot as their primary agent surface? And in that first week, after that all hands, I mean, I can't even tell you just the out core of love for Grogbot, from people that maybe you wouldn't expect or from groups of the company that maybe you wouldn't expect. People who, you know, were daily driving Cheshire PT or like a chat interface, switching all of their, you know, day-to-day, agentic tasks over to Grogbot as like their primary surface for doing work. And so the internal reception was really extraordinary. Can tell some kind of funny stories from that, that week or two period. And then once we saw, I think the internal reception, we immediately switched into, let's get this ready for the world. There's a lot of work to do to scale this out to millions of users. And then that led to the GA launch that we had a few weeks ago. This episode is brought to you by our season's presenting sponsor, WorkOS. What to open AI andthropic cursor, Replic, Sierra, Clay, and hundreds of other winning companies all have in common? They are all powered by WorkOS. If you're building a product for the enterprise, you felt the pain of integrating single sign-on, skim, R-back, audit logs, and other features required by large companies. WorkOS turns those deal blockers into drop-in APIs with a modern developer platform built specifically for B2B SaaS. Literally every startup that I'm an investor in, that starts to expand up market, ends up working with WorkOS. And that's because they are the best, whether you are a seed stage startup trying to land your first enterprise customer or a unicorn expanding globally. WorkOS is the fastest path to becoming enterprise-ready and unblocking growth. It's essentially striped for enterprise features. Visit workos.com to get started or just hit up their slack where they have actual engineers waiting to answer your questions. WorkOS allows you to build faster with delightful APIs, comprehensive docs, and a smooth developer experience. Go to workos.com to make your app enterprise-ready today. Okay, so many questions. One that is really interesting here. So obviously there's Anthropic Open AI. They went from this coding agent that they're like, "Holy shit, this is a big opportunity." And then they're like, "Okay, people are using this for knowledge work. Let's build a knowledge work component." So there's co-work involved that of that within the product. And then Codex, they've invested in. Let's make this useful for all kinds of things. Interestingly, you guys decided, "Okay, we're not going to build this into cursor. We're going to start something fresh." Was that just like obvious from the beginning? Okay, this is not going to work inside cursor of the product when you just start fresh. How like controversial was that decision? It was not obvious at all. I think you're completely right that that was one of those original decisions that at the time we had a lot of discussions about. And I'm very glad with where we landed. And I think to your point, it being a brand new product that you control every pixel of the experience. And you have this consistent vision about where knowledge work is going. And it's all contained in this new thing. I think it has contributed a lot to the success. But there were a lot of discussions about cursor, for example, and some of our coding products. People use it for non-coding tasks all the time. And these coding agents are really excellent at some of these things. But you run into small paper cuts sometimes. the product itself is kind of intimidating to non-technical users, there's a brand association with these things. And so I think we evaluated that path, and I think we saw maybe some of our competitors have been doing of, this is all just one surface, you add new tabs for each new form factor. And it feels a little cluttered, and I think for users they can feel that, that this was not a single consistent vision of the way that work should work, and instead it's three different visions that all kind of share a screen and you can hop between, but it is kind of a shipping your org chart style thing that I think users are reacting negatively to. And so we decided, let's just start completely from scratch, let's see where we can get from there. There might be some really amazing opportunities to bring people from other surfaces into this more bot native experience, but it's really important for people to just have like an amazingly simple and amazingly powerful experience. That is a really valuable lesson for people to take away here, just that that might be the solution instead of adding and complicated and existing product. So interestingly, codex went the other direction. And it's like a different path than a different product, but it's interesting, like they're like, now we're going to make it one thing. So there's, you know, there's many ways to make it work, and it also feels like the path you take there will kind of lead you, but maybe, maybe we'll look back and be like, that was not maybe the best idea. Something you mentioned that you did that is also really neat is this onboarding of early users. I heard you and your team onboarded two to three hundred people manually, including me. Talk about why you thought that was necessary and what you learned from that experience and just like how long that period was of this kind of manual onboarding. I mean, you just learned so much. And the first few onboarding are pretty painful. I'm glad you got a good one, Lenny. But there was some that were kind of rough. And we learned a lot and I think it was important for the core team to be in the room for those and to just sit on a call for 20 minutes when the computer isn't spinning up or when someone's in onboarding and they're just incredibly confused. So that immediately after you're like, that can never happen again. We need to solve this tomorrow because tomorrow I'm onboarding this person and it needs to go better. And so there was about a two week period where we were in that mode and onboarded a couple hundred people. And not only did we learn a lot about the product, I think we didn't really know. I think sometimes with these products, there's some groupthink of ways to use them. And I think internally because people inside of SpaceX AI were just constantly sharing tips and tricks for how to use GROKBOT. Some patterns were starting to emerge that we thought would be useful to the world. But we weren't really sure and we definitely didn't want to buy us the world. Is there an example of that? So when we rolled out GROKBOT internally, there was about a week or two where the common pattern of the way people would interact with the products was you would have five to ten bots. In each bot, you would give a different scope, a different domain, and it was kind of shorthand for different lanes of work. And then at around the end of week two, we started to see these messages internally in Slack of people promoting one of their bots who was a bit of a standout performer. And it was like their kind of primary personal assistant promoting that to their chief of staff. And then they would actually mostly talk to their chief of staff and the chief of staff would fan out all of these tasks to the other bots and would kind of manage the team. And there are some funny screenshots of people actually telling their, you know, the bot they're promoting, that they're promoted and the bot is asking if they get a raise and, you know, is their token budget higher, all of these things. And we kind of took note of that and I think more of the companies started to slightly shift in that direction, but it was not the majority of the company. People use this product in very different ways. And so in some of the onboarding sessions and just from the early access program in general, we really did not want to lead the witness and say, you know, create a chief of staff bot. Here's exactly the way that, you know, that chief of staff should manage all of the other bots and see if early access users would get there themselves. And we actually did see that many of them did. And so then we had a bit more of an opinion they had to take in the product of this feels like a pattern that's working. This feels like a pattern that we should slightly encourage, but it shouldn't be a one way door. And there are a few other examples of internal feces that we really wanted to make sure, you know, would bear out an actual external usage without us imposing that into product. Is there anything else there? Any examples come to mind? Yeah, I think another thing we really tried to pay attention to in the early onboarding was just how much users wanted to see. And I think it is something a bit shocking or just different about Gropbot when you first start using it versus some of the other products that you mentioned. Where a lot of the internal mechanics of how Gropbot works are not shown to the user. And the reason for that is we think as these models get smarter, the same way that your teammate, you know, you wouldn't ask for second by second updates of exactly all the buttons they're pressing and websites they're going to. I think it's too much to ask your bots to do that too. And I think it's honestly just overwhelming and can create more harm than good. And so we moved completely in the other direction of you send a message, you tell your bot to do something. It just starts doing it. It sends you progressive updates as it sees fit and you just see that like typing indicator and the little green, you know, active, active circle kind of slack like that it's active, it's doing work, it'll get back to you soon. But you don't see the internal mechanics, you don't see the tool calls, you don't see exactly every little click it's making on its own computer. And we really wanted to take a strong stance that users did not need to see all of those mechanics. And so that's where we started. And we did get some feedback that's like I would love to see my bots to do list. I would love to see roughly how it's prioritizing tasks and what it's doing. And that's great feedback. But it was useful to hear that nobody wanted the like long stream of just text streaming out and chain of thought sequences. So that also gave us more confirmation that that was the right direction. The fact that you did two to three hundred onboarding calls on us with a small team. I knew that I know the team grew over time, but just that is a huge time commitment. And then you could argue a distraction from the building clearly not a distraction, clearly a core part of the success. Do you feel like that's like that's the volume people need to do to figure out what actually needs to happen? Well, one thing to emphasize is the early access group is not necessarily just people that are highly influential taste makers. You know, you're in this category, Lenny, and we certainly wanted to get a lot of your feedback just from being very close to many other products on the market and just being a power user of these things. But we also wanted to get early access to kind of more unconventional profiles that we as a company had never really interacted with. So one example is there's a coffee shop owner that was a friend of a friend through the company who had heard about Gropbot and you know, one day somebody on the core team had kind of shown them a demo of the test flight and got very excited. And this coffee shop owner ended up being not only an amazingly an amazing power user of Gropbot, but also a rich source of feedback for us. We have a very lively thread with with many, many bugs that could identify it or feature request and it's a completely different use case of running a small business. And so for example, if you know the Shopify integration was a bit flaky or if it wasn't writing copy for products in a particular way, we would get really rich feedback on that, which is pretty different from the type of feedback we'd get from dog fooding this internally. And so I think it was an important exercise for us to check our blind spots and say, A, this is going to be a very general product that is not just a thing that developers use. In fact, it is likely that this is most powerful for non-developers. We need to understand that group much better. And then B is we absolutely live in this kind of Silicon Valley AI bubble, which I think is a useful place to be to kind of push the frontier and push the future of how these products are evolving. But we need to actively get out of that because I think a product like this has the chance of really being the way that the mainstream user and the mainstream kind of business customer can interact with AI in a way that's useful. Let's go back to the timelines real quick just to kind of understand that. So it was a month from first line of code to internal beta. And then what happened after that? It was about three weeks from internal beta to public launch. And then I think we're about three weeks out from public launch as of recording this. Wow. Okay. So month of building the first thing, three weeks only of iterating. And then it's been only three weeks since launch. It feels like it's changed the world from my vantage point. So wow. Okay. What most changed in those, I don't know, in those three weeks of internal beta, let's say, we unshipped a lot. I wish I could have shown you what things looked like maybe two weeks out from launch, where I think we'd realize that the core team, we had a lot of experimental features that we wanted to get internal feedback on, which was useful. We also were kind of putting pseudo-developary visibility tools into Grockbot, instead of having a separate observability pane for those things. So for example, we actually did expose sometimes a lot of the internal thinking of the models in this specific memory. it would store and all of these things, which was useful to debug issues and if you're building the product, you didn't want to go somewhere else to maybe pull that context. But we had to really aggressively trim what we think the user absolutely needs to see in the surface versus what they don't. I think there's even more room there to run, which is something that teams focused on right now is how can we just ruthlessly simplify this product and abstract away anything the user doesn't need to actively be thinking about. So that was one big push. It was unshipped, a lot of junk. And then I think the second big push those few weeks was making it just work. And I think a lot of what people want from AI is not this thing with lots of drop-down menus and bells and whistles, but just a thing that you describe a task, a task that's meaningful to you. And it does it and it comes back with complete work or it comes back with something for you to react to and then steer it and it's kind of next cycle. And so in order to actually deliver on that promise, it's actually not a lot of feature product roadmap style stuff. It's like hill climbing, five really important problems in the back end that many users don't directly experience, but you totally feel when your bot is going off and doing something and can't click the right button or your bot is often doing something and can't log into a website and it just completely stalls your ability to make progress on that task. And so those few weeks, we collected a really rich set of what are tasks that people actually are giving bot? How can we quantify these things? And how can we see week over week that across those categories of tasks that were hill climbing on very important dimensions to making that just work behind the scenes? Is there an example, one of those hills you were climbing that was kind of a technical breakthrough or technical challenge you ever came that really helped it just work? Yeah, one example was from rolling out crock bot, one group inside of the company that was actually incredibly bot-pills, so to speak, was our go-to-market team was sales. And there are a bunch of tools that sales uses that do not have well-supported MCPs or APIs. And I think that's a lot of what made bot so before and after powerful for this group was these were things that they just could not give another AI tool reliably. We would get stuck at some part in the process. And then bot kind of felt like they had an assistant or kind of felt like they onboarded someone to their personal team. They gave it a laptop and it could just run. And so there were a bunch of small things, and probably a list of 10 or 20 of them, of places where for whatever reason, the mouse would just not have fine enough control to click on exactly that part of the Salesforce dashboard or something like that, that we would have to take back to the core team working on really the infrastructure to say, "Here's a very concrete case of where the agent not having this visibility into the browser or this visibility into the pixels on the screen is making it impossible for this task to be done." And that was just a lot more tangible than seeing a number on a dashboard slowly creep up. It was kind of like new chunks of work getting unlocked, and you would immediately feel the feedback where you would ship an improvement. It was kind of behind the scenes, kind of infrastructure-y. And then the next day, you would just get this outpour of, you know, love and appreciation from the sales team, that now this workflow that was failing the last seven days finally works. And it's just a constant exercise of finding those next tasks to unlock and then solving them. So as a computer use, basically improvements seems like a big unlock. I heard also the, I had Adam Ward on the podcast, who's head of recruiting, how to firing, basically go ahead of talent. I heard his team was like one of the top users of Grocbot. Yes, the recruiting team gave us a lot of great feedback anytime the product, especially in the early days, if there was a little bug we'd get a ping from some folks on the recruiting team. Yeah, I think the main use cases for recruiting that were particularly interesting was first, it was incredibly valuable as a sourcing tool. And I think one thing Adam talked about on the podcast with you. And it's a big part of our hiring philosophy internally is be looking for a job. Being on the market is not a precondition for us trying to hire you. And in a lot of ways, you know, the best way to hire is really just look at the biggest problems at the company that need someone to own it or take it to the next level, find out of the total universe that people in the world who would be best and then ruthlessly go after them and try to convince them to join. And this is a lot of the philosophy from the very beginning of the company. And so if that's your mindset, really the best recruiting work stream is not or workflows to automate are not, you know, here are a bunch of resumes read through them, help sort them. The most useful thing is here's an entire universe of potential people help match that to this very concrete business problem or this very concrete role that we're recruiting for and help me get in touch with them, help me get coffee with them, let's just throw everything at it. And so there have been some cases of really kind of unexpected ways of finding top talent that is beyond even just, you know, looking on LinkedIn and trying to find interesting people. But who are the co-authors of this paper and the PDF doesn't exist on Google Scholar? It just exists on this conference website. I want you every morning to go to the conference website, download the PDFs. If there are any new ones, you should find every new name that we've not yet tracked. You should add that name to a spreadsheet, should do research. You should look at everybody at SpaceX AI, see if there's anyone directly connected. If so, you should send them a slack message asking for an introduction. Like it's those types of always-on sourcing use cases that I think in the past were incredibly manual and now it's the type of thing AI is superhuman at and our team can focus on closing great candidates and getting conversations with great candidates and not pulling these huge lists. Wow, that is such a cool example. First of all, someone's about to take the transcript of what you just said put into a bot and create their version of this, which is great. On the other hand, I think you guys could sell a template of the spot for a billion dollars. I think it basically used Adam's team strategy for finding the best people and just turned it into a bot, holy moly. Democritizing, hiring. I want to come back to a few things. One is, you made this point about unshipping. Such an important point, I think that people can overlook because AI is not good at telling you what to take out. It's very good at, okay, here's more ideas, here's more stuff. And something that's come up a number of times on the podcast is that's a big opportunity, that's a big space for humans to continue to be very important and valuable is knowing what not to ship and what to cut and what not to do. And so it's so interesting to hear that that's been a big part of the internal evolution of from prototype to launches deciding, okay, we need to cut a bunch of stuff. Anything more there? Yes, so one thing we talk about internally is for anything that we're working on for Grunkbot, what is the launch post? What is the thing that we would actually tell users? And if it's not complete, I imagine, launch tweet. And if it's not compelling, maybe we shouldn't be working on it. If it's not something that users will directly feel in the product. And to take that even one step further, I think there is this old school software tendency to say things like Grunkbot now has. And when you think of completing that sentence, it would be like a new button to press, or it would be a new dropdown, or it would be a new integration that you can press plus and add. And instead to reframe it as Grunkbot can now, which is I think a much more human way of kind of describing these capabilities. And I think it's forced us to think more in the frame of what are tools and what are capabilities that we can give Grunkbot? Not what are new things we can add to the product? Like adding things to the product is not the goal. That's not the thing that's going to push this product forward and make it more useful to more people. Making your bots reliably do really impactful work for you behind the scenes in a way that just works. And giving them the capabilities to do that, that's what users actually care about. And so I think in the context of on shipping, there have been a lot of Grunkbot now has things that we've realized are actually just capabilities that don't need pixels. Let's kill as many pixels as we can. Those can just be things that your bot manipulates behind the scenes for you, and you don't need to directly control. And I think one example of this is the way that many of our competitors, you set up automations or routines, is you go into a sidebar, you press plus, you select what the trigger event is, you then select what action it should take after that, and you might describe it in natural language. And it's just really conky, and it means that people don't set up many automations for many things. We certainly have seen this in the coding realm. And so I think what Grunkbot did in response to that was actually you should just define automations in natural language, and you should tell your bot, remind me that at 8 a.m. every day, please. And then it should just do it, and you should never ever have to see that interface of creating an automation. And so that's kind of the decision that we've made, and now that's how 99% of automations on the platform get built. And I think there are a bunch of other places where we can do things like that. So what I tweeted about how much I love Grockbot when it launched and a lot of people replied they're like, "Wait, can't you just do all this with Codex and Codework?" And you can. Technically everything, as far as I know, you can do with Grockbot. You can do with the other foundational models, the coding assistance. So let me just ask you this big question. What is it that you think you did that is so different, that allowed, that made Grockbot so successful? I think it was two early decisions that at the time definitely did not feel obvious. But in hindsight, I think our critical to what makes Grockbot work for people. And the first is you should never have to think about local and cloud and where are these workflows running? Does my computer have to be awake? If I kick it off from my phone, does it need to be tethered to my computer back at home? Like there's so much junk happening right now when people are trying to conceptualize where this runtime lives? And we made a really early decision that this should just all be in the cloud. And if it's in the cloud and it's this persistent colleague that has its own computer, it can do its own work. It has the same state everywhere you interact with it. It opens up a lot of really amazing opportunities to text your bot, kick it off from your phone. And in the future, you should be able to call your bot from anywhere. And it should be able to do real work. Like this is its own entity and it lives separately from your device. And I think that was a very important decision that current products I think haven't made that same decision. And I think it has a bunch of paper cuts as a result of it that users feel every day. I think the second decision was to kind of take that one step further of not only should this be an agent loop that kind of runs in the cloud and you can interact with in various ways. It's actually really important that these bots have their own computer. And part of it is what I described earlier of there are a bunch of tasks that don't have well supported MCPs and APIs. We as humans don't do our jobs via MCPs and APIs. Like we use a computer and we click on pixels and we kind of type things in input boxes. And it's very important that your bot has those baseline capabilities as well. But even to go on stuff further, I think we're in a really weird moment right now that I think we're going to look back on and be like, I'm surprised that this is the way that a lot of people worked with AI where you're onboarding these super intelligent new colleagues, these AI bots. And you're asking them to share the same computer that you have. It's crazy. Like if you were onboarding someone to your team and you said, it's your first day, I'm going to onboard you. You don't have your own laptop. You're going to sit next to me. We're going to share this laptop forever and constantly trip over each other. You're going to have access to my credentials. I'm going to have access to your credentials. Like that's just there's a good reason why that's not the way people operate. And I think bots and these kind of AI colleagues of the future will also need a way to onboard them. That's that's somewhat similar and we've tried to push the product in that direction. That is so funny. Why do you think the other companies didn't do this? My guess is they were building off of their existing coding assistant platform and approach and this was like a pretty big shift. I think a lot of it comes down to starting from scratch and how freeing that is. And we felt that ourselves where a lot of the primitives in Gartbad, we had attempted or we had built in other ways. You know cloud infrastructure we built for coding agents. The way that you could maybe name agents and talk to them is discrete entities. It's a pattern that we're also seeing for developers kind of bringing specific named agents into Slack. But instead of kind of trying to retrofit those concepts into some new surface or into an existing surface, which I think would have been the strong default I think for many companies. We decided to start from scratch and we decided to just try to get these things really right for general knowledge work, which is a new audience. And then second for the point in time that we're at now where the models are very capable. And if you give them the right tools and the right infrastructure, they can do a lot. A lot of these ideas, these aren't strokes of genius on our part. And I think for good reason, I think these are primitives that had already been getting attention and product market fit by other products, you know, the open cause of the world. And I think we took a lot of inspiration from that and tried to productize it into the tighter surface, something that required a little bit less setup and I think was more accessible to more people. And so I think our competitors and other tools that have been trying to solve these types of problems, I think we're all seeing the same opportunity, I think we're seeing a lot of the feedback from the market. But I think it's just been hard to act on if you're stuck in the existing paradigm. And if you have a lot of kind of some cost in that existing paradigm, it's very painful to create a new thing from scratch. And I think a lot of that is what allowed the product to just work and click for so many people. So what I'm hearing here is the keys to success of what made this breakout, a cloud-based computer for every bot instead of locally, a name kind of like specific bot. And by the way, there's this like we're all moving from agents to bots now, nice job, it feels like you guys have pushed it over now, okay, we're all bots now. So a bot per kind of task use case, very unique versus like a thread conversation or something or like one off job. And then I feel like there's just works was a core part of this and you talked about how long it took to get to that place of like, okay, now it actually works really well. You mentioned OpenClaw, obviously this is inspired by that, which to me when I first use OpenClaw, like holy shit, this is the future. How can we not have this? And then Hermes came out and everyone's been trying to build the OpenClaw that works very easily for everybody. Can you say more about just like how OpenClaw and that story informed the way you guys thought about it? Yeah, so I think OpenClaw got two major things right that when we were seeing the way the market was reacting to OpenClaw and ourselves using the product, we found quite exciting. I think the first thing was the models are really smart and they're going to continue to get smarter, but even at current capability levels, if you can just give your bot access to the tools that you do that you use to do your job, you can get a lot of the way there. And a lot of the places where people think AI is down or maybe not as impactful as it's been promised, a lot of that I think is downstream of it just being harnessed in the wrong way. And so if you kind of give access to a much larger set of things that has access to its own computer, how far can you go? And I think OpenClaw really kind of forced that question for many people. And then I think the second way OpenClaw changed the mental model of the AI was really viewing these things much more as colleagues and teammates and people and personifying it a bit more and it being this helper entity that has access to your life and can kind of extend you even further. And so we took a lot of that and I think what Crockbot maybe extended was a it needs to be really easy to set up and the hacky you know you have a piano at home and Mac mini setup clearly was not going to scale to millions of users clearly is not going to be the way importantly that businesses take advantage of this technology. And so we really wanted to build an amazing product with that in mind. And then second is I think there are a lot of places, a lot of rough edges to sand down and just make a delightful product experience and make these things just work and try to remove some of the abstractions that power users of AI are very familiar with things like skills, for example. How can we make a Crockbot user not even have to know what a skill is they should never have to type a slash command these things should be created in the background is a useful primitive that the bots have access to. But something that users you know it's not incumbent on them to always be on the cutting edge of the eye. So that's really where we tried to innovate and I think there's still more room to go there. As you say that I have my Mac mini with my formerly alive open claw on there and that was an era and it's so awesome the work that it has inspired. I know it continues and others still a lot of value to open cloud but when I saw Claire row who's been like the biggest proponent of open claw and has it's like become a core part of the way she lives and works with her kids and does all her work she just switched all of her open claws. She showed them all down and switched to Crockbot that's a huge like it sounds funny but that's actually a huge milestone it just how much things have shifted. What's kind of the vision for Rockbot what's like where does this go what does this look like in the future what's like the ideal platonic version of Rockbot. So the ultimate vision of Rockbot is incredibly simple which is you should have a team of AI bots that help you with your job and help you with your life. And it should really feel like a team it should really feel like teammates that are autonomous are helping you you can steer them in various ways you don't have to micromanage them they have access to the tools necessary to do great ambitious work. And one thing we really use as a North Star on the product side in building this is we kind of get closer to this teammate future. How can we, in every product decision we make, think about this less from the perspective of a SaaS product and more from the perspective of we're trying to build useful AI teammates. And so there have been a bunch of examples where we kind of have to push ourselves to be more like colleague-pilled in a way, we sometimes use that term, where we're having a product debate about something, there are good arguments on one side, there are good arguments on another side, both paths feel sensible, like in product land, this maybe it doesn't feel like there's a clear cut answer. And then you zoom out a little bit and you remove yourself from the, you know, tech companyness of it all and you start thinking, how would a human do this? Like what would you want from your teammate in this exact situation? And oftentimes the answer is really clarifying and pretty unanimous, if there's oftentimes not a lot of disagreement among the room of like how a human teammate you would prefer to work with in a certain way. And then once that answer is there, well then we just need to build it. And there are product implications, there are model implications, there's a lot of things that need to go right to actually deliver on that experience. But in some ways, it's not a rocket science, it doesn't require you being a genius. You just need to ask the question of what would you want from a human teammate and can we push AI to behave in a similar way? And so to give some examples of that, I mean, we've been thinking about what the right voice experience with these bots should be. And I think in the context of a human, for example, we have a really good analog of a lot of times I'm slacking back and forth with a teammate, resharing context, and a lot of times it's just much simpler to get on a five minute huddle with them, and just press huddle, talk back and forth, I share my screen, and I show exactly what's on my mind, they share their screen, we hop off, and then we continue async from there. And that's not really an experience that any AI product has gotten right right now. And it is deeply integral to the way that I think humans collaborate, and so we want to build something like that. And there are a bunch of other examples of these very clear patterns that just work that I think you should also feel when working with AI. I love this term colleague build, and it's come up so many times over the course of this chat already, how that is kind of a through line to making these decisions. For example, the computer, hopefully you can have it so good, obviously people would have their own computer. The naming piece is also very important part of that. A big question on my mind in this space, and I'm so curious to get your take is this separation between work and personal, do you think people will have two different assistants of work in a personal, or do you think it'll be one? When people think about a work product versus a consumer product, I think there's just a lot of baggage that comes from the last decade or two of horrible B2B software that leads to people seeing a product that is very simple, and some ways chat to be to you is like this, I think Grockbot has many of these properties, and assuming that it's not a work product or assuming that it's not a power tool. When you think of a power tool in this kind of last generation, I, in my head, picture something a bit like Photoshop, for example, where there are all of these different dials to turn very precisely, the user of the tool is this kind of ultimate cockpit flyer who knows exactly what all the knobs do and can use them perfectly. I think power tools of the future will actually be very different from that, where it is mostly just intent being expressed in good steering on the part of the human, and these AI tools abstract away all of the knobs. You should never see them unless you need to directly manipulate it, which might happen, and there should be a great affordance for that. But ultimately, it really is just working with a teammate, and so the interface for that is quite conversational. And a lot of ways Grockbot, when you look at it, when I walk by someone's desk and I see Grockbot up on their computer, for me for a split second, I'm like, oh, are they on like a messaging app? And it's like, no, this is actually the primary tool that they're using to do much of their work. So I think to your question of, are you going to have a different set of bots for your personal life and a different set of bots for your work life? I do think there will be a separation for many people. They want a separation between personal and work life, and I think that's great. I think that's important. And I think there are a lot of common sense reasons why those things should be separate, even from the perspective of an enterprise. But I think our goal and the thing we're trying to build towards is Grockbot should be the way that a large portion of the things you do day to day in your work, you should be able to delegate a lot of that to Grockbot and focus on the higher leverage things. And then it should similarly be the way that you delegate a lot of the low leverage parts of your personal life. And those two things actually are not different problem sets. And a lot of ways the product form factor and the ways of solving this problem is pretty much the same. And so my instinct is that I think one product will be the best form factor for both of those things. And that's really what we want to build. Bam, that's a big tam right there. I love to hear. It makes so much sense. Obviously, the question is how do you avoid cross-contamination, you know, personal stuff, somehow infiltrating, exfiltrating stuff from work. But it feels like that's kind of, okay, so what I'm hearing is that's the direction the question is just how to do that, it makes people feel super safe, have kind of like the sock tube, stuff in place, and also just feel really fun. This episode is brought to you by Mercury, radically different banking, now with spend. I've been a Mercury customer for so many years now. I switched all my business banking to Mercury, and honestly, I could not be happier. It's what online banking feels like when it's built by product people, not by bankers. And now with spend, you can give your team individual cards, set spending limits per person or for team, and have expense receipts automatically pulled in from Gmail or over text. You can even give your AI agents their own cards with their own limits and policies. Most founders start out the same way. One card used by everybody at the company. It works until it stops working. Someone goes over, a receipt disappears. You spend two days trying to figure out who spent what and why. One is expense management built directly into Mercury, all your team's cards, budgets and reimbursements, all live in the same place as your business banking, no chasing, no manual reviews, no end-of-month scramble. The result is a team that can move fast and a founder who is no longer the bottleneck. Learn more and get signed up at mercury.com. Mercury is a Fintech company, not an FDIC-insured bank. Banking services provided to choice financial group and column in a member's FDIC. The I/O card is issued by Patriot Bank and a member FDIC pursuant to a license for MasterCard International Incorporated. Let me ask a couple technical questions. On the computer side, what's the simplest way to think about what you get as a part of your GROC bot account? Is it like AVM that is running in the cloud with multiple logins? Is it like a separate VM instance per bot? How do we understand that as much as you can share? Yeah, I think to go back to the teammate frame of the product to kind of extend the analogy even further, if we were on a team together, you and me, I think the number of times that you would have to manually take over my computer and start clicking on things and like, you're doing this wrong, you should go here instead and type in manually. Hopefully, it's pretty close to zero. But hopefully, that is not something you have to really do with a colleague or a teammate. And so similarly, I think right now we're in a place where a computer use is good. It's getting much better. And in very short order, I think the computer concept will be completely abstracted away from the user. You should never be clicking into a remote virtual machine. You should never have to take control, you know, if there's like a wasteful path and you have to kind of steer it into the correct path. So in the medium term, I think the computer concept will be an important concept for users to have, but will not actually be something that they're interacting with. So I think the right way of thinking about Garkbot is it's a team of bots, it's a team of agents that do work for you. And in terms of what they have access to, they have access to a very long memory set of your past interactions with them. And so I think there's a current paradigm if you create a new chat for each discrete unit of work. I think there are a lot of problems with that. I find myself copying and pasting between chats all the time. I think it's just not a great way of grouping categories of work. Instead, the same way on a team, you have a good way of grouping categories of work of kind of roles. You should have roles of kind of different swim lanes of work that you do. And it should learn from you. And it should get smarter over time. So I think that's one very critical thing, is these are long lived agents. These are not individual one-off sessions, and these agents get smarter over time. And then the second thing is those agents have access to all of the tools that you would expect a human colleague to have, which is the APIs, the MCPs, that's great, but then access to its own computer, which it can freely manipulate the way that you would. At this meetup that I went to, Shab, who's on the, I think, Gurd Market team, demoed something that blew everyone's mind because you have a computer within each agent, you can run a lot of different things on a computer. He was running Rockbot within Rockbot. like the bot can run its own grog bots. And I know he was using it for testing and watching regressions and things like that, but that's just like a mind-expanding idea. And I'm curious how many levels you can go before the universe collapses on itself? - I do that one too. That one's actually a very useful thing to do is you down the grog bot for one of your bots. Mine is like a QA tester bot. And that way if there's ever bug reports or if we're kind of testing out a new build for example of the desktop app, I can just say, hey, here are 10 workflows that we need to make sure are getting better release after release. I want you to test it. I want you to write it to this notion document that has like an extensive list of all of the past tests that we've done of past client versions and compare them. And so I think once you start breaking out of, this is AI chat with a set of connections, which is I think where most people are conceptually now, instead to, this is a colleague with a computer. And anything I would ask a colleague to do on a computer, I can ask grog bot to do. It just raises the ceiling, I think of what you would think to give to AI. - Are there any other mind expanding use cases or ways to use grog bot that you've seen or you use? - One pattern that I've seen from any users that is simple, but I think there's a lot of depth if you keep investing and making it better. And this is kind of where I can get kind of nerdy about optimizing my setup, is grog bot as an info vor in some ways of just consuming huge quantities of information, removing that from your own cognitive load, giving you peace, and then coming to you with the stuff that's important. And I think the V1 implementation of that, which many people do, is grog bot sits on top of Slack and it sits on top of email. And I tell it high level, here's my role at the company, here's kind of what I care about. I want you to notify me in these cases, in these cases you don't need to ping me directly, but you should include this in your daily roundup that I read every day. That's like the V1 implementation. I'm not sure what the V10 implementation is, but like maybe I'm at V3 or 4, which is you can give these bots a complete fire hose of information. So I have mine hooked up to like every mention of grog bot ever on X, and it's interacting with our internal context, it's interacting with the QA tester to like see if it can repro any bugs or feedback that we're getting, have hooked up to my own kind of messaging services, to like quickly act on feedback and reach out to people. And I think there's this just like always on kind of chief of staff entity that can preserve your focus on the things that actually matter, but it is always like kind of surveilling to see if there's anything that should get your attention. And we've seen some funny cases of people actually giving their grog bots, which I have not done this yet, that maybe soon, giving their grog bots access the ability to page them. And so if something like super urgent happens and they're at a coffee or whatever, they get paged by grog bot, which is the type of thing that you only wanna do if it's urgent and you really wanna, if you wanna trust that grog bot, does not have false positives. So far, those people have reported that it's been very helpful and successful. But I think we're gonna see more of that type of stuff where the agent or the bot should actually be more proactive to you than you reaching out to it. And I think that will be the next shift in AI. - This touches on, there's a number of things that I've been very impressed with watching your team operate. One is speed, which I wanna talk about, but the other is how you were, like there's awareness that this is a moment in time to capture a lot of market share and really take as much of the market as you can before somebody comes around and they're like, okay, now we got something awesome, especially one of the foundation labs. So watching just how many free accounts you guys are giving out, also the focus on use cases so smart, because it's such a novel thing. And you open it up and it's like, what do I do with this? And there's such a focus on, okay, here's a bunch of things people do with it. And then there's all this talk on Twitter and just like all the ways people are using it, template makes so much sense. Just these two kind of like focuses for what I can tell, get as many people on as fast as possible until somebody's like, okay, you know, 'cause someone's gonna come around, be like, all right, here's the next thing. Super smart and also the use case focus. I know you're a go-to-market person at Cursor before this, anything you wanna share there by just the approach to the, to go to market right now for getting this out there. I think the pattern we saw for coding will be somewhat similar to what we see for general knowledge work. And I think we've learned a lot from that on the go-to-market side and more generally, just building practical AI that people use. And I think we as a company have culturally really cared about not building demoware, like building actually useful stuff in the world and kind of obsessing over that. And there are so many shiny objects and like fun prototypes to build, but ultimately that's a very different problem than getting this in the hands of millions of people and having it transform companies. So that's really, I think culturally, where we've always been focused. And so I think on the go-to-market side, what we saw for coding was a very simple pattern, which was there was an early adopter crowd. The early adopter crowd would use these coding tools and really push them to the limits. And they would mostly push them to the limits on individual projects. They would, on nights and weekends, I'm thinking like 2023, you know, kind of earlier, people would kind of go home from work at work. They were using a basic IDE, this is pre-AI. And then at home, they'd work on a side project and they'd be using cursor or they'd be using, you know, the latest and greatest AI coding tool. And that would give them an extreme amount of acceleration. It would feel like they were experiencing the future and then they would come back to work and they would demand it. And they would say, I cannot picture working any other way than this. I feel like I'm completely walking through molasses right now, this needs to change. And I think for knowledge work, we're gonna see a similar pattern of people really feeling the aha moment. Sometimes in a personal capacity, and I think we're certainly seeing a lot of this, like on X right now, you see all of these examples of Grockbot controlling their home robot computer, home robot vacuum cleaner, or Grockbot, you know, helping them save money on their Tesla charger negotiation, like all of these fun use cases. But I think the next step is going to be, this is not a consumer product. We think this is gonna transform businesses. We think this is gonna transform teams. And it will be bots coming into teams and contributing really economically valuable work, especially as they get much smarter. And so on the good and market side, we're certainly making a big push on prioritizing businesses and thinking about not just the single player use case, they're working with a single bot, but how does a bot work inside of a broader team? How does a bot work inside of real company systems that are complicated and there's a lot of context and a lot of history to understand? What does memory look like in kind of a broader organization versus kind of a single individual you're catering to? And I think there are a lot of unanswered questions there, but I do think Grockbot is the right primitive to create this switch to agents for the rest of the company outside of coding. And that's a place where we're quite focused right now. And along those lines, it's very clear you all understand the power of distribution and how you need to find both an amazing product and get distribution, right? Because you know, Grockbot's amazing, but the combination of how smart you guys have been with getting it out there in all these different ways is really impressive. And I think that shows you what it takes these days to build something that's really successful. I'm gonna ask about the brand of the different brands around this product and the company, just so people can try to understand. 'Cause I know you're going through a transition, acquisition, SpaceX, all these things. So there's Grockbot, there's cursors, is that, so talk about like the products and the way to think about these different brands today. And I know it'll probably continue to evolve just so we can communicate about it correctly. - Definitely. - Yeah. I think there are three big pillars right now of SpaceX AI. So the first pillar is the coding product and set of products and right now that's cursor and Grockbilt. And I think we're big believers that having a professional work surface for developers and for the engineering part of the organization is going to be really critical. And right now people use Grockbot sometimes to kick off cloud agents or to kind of merge PRs or to do QA, a bunch of engineering adjacent tasks. But ultimately when you're shipping production software, we're big believers that that is gonna require, you know, a product where every pixel is optimized for that end user. So we're making big investments there. The second category is general knowledge work. And we think Bot is a really exciting step in that direction. There's a lot more work to do of making it more useful, extending it to new surfaces. It really feeling like an AI teammate that you can delegate work to, especially inside of companies and businesses. So that's kind of the second pillar. And then third is the general model effort. We wanna train the smartest models in the world that are really capable. And I think one thing that somewhat distinguishes SpaceX AI from other AI labs is I think our goal is less to build, you know, chase super intelligence or some kind of vague aspirational ideal. And the goal is actually very practical, which is to build useful AI. And we do that on the product side, we do that on the model side. And I think part of that is also just cultural of like the group of people contributing to these models are engineers and people who kind of came into the model training effort from like a very applied mindset. And I think that's what gets this company going and I think is actually a slightly different direction from some of the other competitors out there. Super interesting. Okay, there's a couple of directions I want to go. One is you have this tweet that is I think you pinned it or maybe it's your last tweet it's up there in your timeline if people check you out. So the tweet is an AI that does 100% of the job feels categorically different from one that gets you 90% there. I've significantly updated what I think AI is capable of. Say more about that. I think for me what made me so excited to work on Grockbot and contribute to it is it was the first time for non-coding tasks that I felt like I could truly delegate work to AI and not have to think about it and I would come back and it's done. And I think engineers have been feeling this for quite some time for maybe a year a year and a half things have been like that. I mean the job of the developer has completely transformed. It is unrecognizable from what it was two years ago and many, many words have been set on that topic. But I think it's underrated how different that experience is from what most people are feeling about AI right now and the way that AI has changed their lives. And it looks quite similar to the way that people would use AI like two years ago where you create a new thread for a task, you type it into an input box, you hit enter, you watch all of these steps happen, you get an output, it's not quite right, you keep working on it. And Grockbot I think short circuits a lot of that. When you first kind of lay eyes on the first screen, you're like, well, this is clearly different. Let's see if it actually works, but this is different. And then you give it something and it kind of works, you know, it's a surprising extent. And I think we're going to do a lot to make it work much better. And so I think what I was expressing in that was when you have a teammate that you only 90% trust and you give something to which luckily I do not have the experience up here because I work with great people. But if you delegate something to someone and you're like, I know I'm going to have to be thinking about this while you're doing it. And I know it like, probably is not going to be quite there. And I'm going to have to intervene and kind of steer it slightly. You're not that's not 90% task completion. You're still doing the thing and it feels that way and it's it's weighing on you in the same way versus like truly throwing a no look pass to a colleague and being like, you got this. Here's the context. Go off and run. I'm excited to see what you do. Like that's a different category. And I think that's the type of thing that people feel with rock bottom every day are these no look passes and you just trust that it can get it done. And then it does. And it's just a very magical experience. Yeah, I have had that experience consistently. Okay, so another element of how you all operate that has really impressed me and I've not seen this before is how fast you all move. So I got added to this like slack as you're as giving feedback with some folks. And it's just like, okay, how about? Okay, tomorrow we're going to give you some free codes to give out. You could do it tomorrow. We'll do this tomorrow. Or we're going to launch a marketplace with templates. We're going to launch this in two days. I was just like, what? I don't know if I don't have time for this. How do you guys with all the things going on? All these things constantly shipping and also staying consistent and high quality and feeling clear that it's towards a specific vision. So there's kind of like two parts of this question. Just what what's the secret to how fast you all have been moving? And how do you stay aligned moving that fast towards a vision of that you all want? They all believe in it and where you want it to go versus just like, you know, bandating it along the way. Yeah, one thing I've been really happy has never changed is that startup feeling inside of the company. And for context, when I joined Chris originally, we were about 15 people. We scaled to about to over a thousand. And then now we're a part of space X AI, which is kind of an even bigger organization. And it's something that is just so fun to be a part of when you're around this group of incredibly talented people, everyone's moving 100 miles an hour. You trust, you deeply trust everybody to execute on their part of the equation. And there's a clear vision that everyone is fired up about. And like it knows that they need to execute on. And you know, as companies grow, and we've had the fortune of hiring really great people from other companies gone through hyper growth, things slow down. And you kind of keep telling yourself, we're still a startup. We still move quickly. But you really don't. And everyone knows that you don't. And it's just, you know, it's easier to say than to actually be. And you know, fingers crossed, this continues to be true. I think it's really critical for our success if this continues to be true. But even as we scaled, it has always felt like that startup that kind of I first joined. And I think if you define a startup by number of people or by like the funding round, like none of those things really make any sense, the core thing that defines a startup is exactly what you're describing, which is this kind of scramble energy of things are kind of chaotic and kind of disorganized. And like for a lot of people, that's not a pleasant working environment to be in. But it has these amazing properties of you can make extreme impact in a particular direction and a short amount of time. And you really do get out of the system what you put in. And so I think as a culture, I think as an organization and the way we construct ourselves, it's really been to enable that property in a way that I think some of our competitors and other AI labs have gotten much bigger. And you can feel it. And I think us, even as we scale, there is that start up the impulse that is quite important to me quickly on these things. Let me pull on this thread. And let me ask you this big question that I've been looking forward to asking you. If you if you were to look at cursor from the outside, you, it shouldn't have worked. It shouldn't have lasted. Because one, it's in the most competitive market in the world, competing against the fastest growing companies in history, open AI and Anthropic. So that's one, it's like the competition is unlike anything anyone's ever experienced. Two, it sits on top of those platforms to power it. And what I've seen as an outsider is what has allowed cursor to win and have this massive exit and continue to succeed is how quickly you all adjust to the reality of the market started as autocomplete and then things moved on to just talking to agents and then into the cloud and now rockpot. To me, that feels like a core part of the success is quickly adjusting to reality and also building the best in class experience for a thing that also exists other places. Rockpot's a great example. You could do this other places, but it's the best in class experience, cursor the ID, the best way to code. So so that's my question. Maybe I answered it. But what do you think has been core to cursor's ability to not just survive in this crazy competitive market, but do so incredibly well consistently for so long? I mean, a lot of this, and it's a fuzzy answer. A lot of this, I think, is downstream from culture and the culture that you set and the people that you bring in and the way that they approach these problems. And I think for us, exactly as you said, we have never been complacent. We've never felt like we've won. And it's always been about the next thing. And I think there's been a really deep belief across the company that AI is moving incredibly quickly. Our goal is to translate those capabilities into amazing products for customers. But those products are going to change and they need to meet the moment as the capabilities get stronger. And what met the moment two years ago is completely different than what's meeting the moment today. And if we as a company can't completely reinvent ourselves every six months, which recently it's well to even shorter than that of kind of completes like very significant reinventions of our priorities, the core product, what users feel, we're going to lose. And I think it's that spirit of always pushing to be on the frontier, never thinking it's over or that we've won or that we've gotten it right. And just constantly updating our beliefs that has gotten us to where we are now. And to your point on the competitiveness of this space, one thing that I think is important to point out is AI coding has always been competitive from when cursor first first kind of came to be. And at the time, the competitors were Microsoft and others and handful of maybe 10 or 20 companies. And I think it's notable that none of those competitors are at the forefront of AI coding right now in large part, not because of any incorrect decisions that they made or any lack of resources on their part, but this cultural inability to move quickly and to change to meet the moment as the moment's changing. And so I think that's exactly what we're going to do. exactly what has led us to invest in things like Rockbot, for example. Are there any core values, just like specific ways you phrase this to kind of remind everyone of this is how we work? Yeah, two values that I find myself coming back to quite a bit. The first one is this idea of deleting the product. And I think it exactly ties back to what you're saying right now, where when you look at every past iteration of cursor, for example, but even I think when you look at every past iteration of Rockbot, I think we will feel the same thing is things going away. Not new things getting added, but these scaffolding product overhangs style things that get built in because the models have not yet gotten smart enough to do it themselves. Those things will get moved away over time and we need to feel comfortable making kind of hard decisions that might upset a small set of users or a small set of us internally to do the bigger thing of make the product simple, make the product powerful, and adapt to where the future is going. So I think that's been very core. And then the second thing which ties back to the kind of pace of execution is just do the thing, which I find myself kind of repeating a lot even as we've kind of grown as a company is it's on you. You know, we're all in this boat together. We want to win. But if you see something that you think needs to happen, you know, this is not an ask for permission culture, you go out and you fix the thing and you pull in the resources that you need to make it happen. And I think that has made many people very successful here before and I think it's something we really share with the SpaceX AI as well. Agency as you may have heard. So interesting. One of the big questions that comes up and this might be my final question is around modes. And a lot of people look at cursor is a really interesting example of there in a market with technically maybe no modes, but they've continued to win and succeed. The two modes you think about with cursor is the data feedback loop of people a lot of completing, learning what they're doing and training models based on that. So that's unique. The other is just best in class experience. And being like a high daily active user product and finding over time what works and what people need. What have you just learned? And I guess any thoughts on modes in this space that might be helpful for folks that are trying to figure this out for themselves. Yeah, there's a lot of talk about modes and it is a pretty interesting moment in time to be starting a company. So I can understand why so many founders are kind of asking themselves that and trying to project out 12 months from now, 24 months from now just feels like an eternity. I will say that I think if cursor in many other successful companies of this kind of vintage, I think if they have thought about modes slash kind of tried to work backwards from some strategy diagram or like maybe more abstract notion of how a company should work. You don't think that would have created this outcome or this product. I think what really created the magic of cursor was an obsession with building a useful thing today. And I think it was constantly this exercise of you can kind of see where the world is going three months from now, six months from now models are going to get smarter. A thing that isn't solvable now is finally going to be solvable. And I think cursor was a little bit this recurring prompt of how could we pull that stuff to today, even if it requires a little bit of engineering on top to make it work or a lot of engineering on top to make it work, even it requires changing the product in a specific way so the user can interact with this new capability. How can we bring that forward? And then three months from now we should delete all that stuff because it'll just be good and like basic, you know, common bare minimum of the product and then we'll build the thing for three months from then. And then it was constantly just doing that over and over again that I think led to users really trusting us and placing, you know, their time inside of our product and trusting that we were kind of bringing things to this next frontier into the next future. And so I would really encourage many founders or people starting out today to be more grounded in that perspective of how can I make something that is not possible now possible? Users are going to come to me to use that thing. I'm going to pull them to the next impossible frontier and then through all of that, I'm going to gain a lot of distribution advantages, I'm going to gain data advantages. There will be value there, but I think that's really the place to play. I love that answer, essentially, like the way I'm thinking about is just build something people are obsessed with, don't overthink the most piece, and if you can continue to do that, you'll find something which in cursor case ended up being a few things. And that came up actually recently on another podcast I did. I don't know if it'll come out before or after this. This idea of the most are discovered, not planned ahead of time a lot of times. Okay, let me actually ask you for a rock bot tips, as an actual last question. Some people are going to be like, oh, shit, I got to try this thing. What's all this excitement all about? I'd be some advice for folks that are trying out. Let's say for people that are new to it, just like here's some keys to success and maybe some power tip for someone that's already with it and just say, oh, I didn't know that. Yeah, I want to stay away from the super hacky pro tip stuff because I think our philosophy as a team and as a company is that those things really shouldn't exist. There shouldn't be all these crazy knobs. You should be able to delegate something to rock bot and they should do it. So what I would encourage for someone new, who's just downloading the app, you're looking at this screen, I think the first thing is give rock bot the context it needs to be successful. So in a similar way as if you were onboarding someone to your team, it'd be really helpful for them to have access to Slack and your email and the company records that you use every single day. So I'd give it access to the tools and then I would actually ask rock bot what it can do for you and let it kind of go through those connections that you've initially set up. In my case, it might be, I give it my email, I give it Slack and I was really surprised when I was first onboarding. This is, we didn't have any onboarding speeds at this time, so this was kind of the first task I gave it was go through my Slack, go through my email and suggest like five things that you can take off of my place and what it would take for you to do that. And I suggested five and like two of them were actually really helpful and I just immediately spun off two bots to solve those two. And that was my big well moment of feeling like no other AI tool in the past could have done those two things. It was not like draft an email, it was like do a chunk of work and so I'd encourage people who are brand new to do it that way. And then for people who are not brand new, I kind of constantly finding new patterns for ways that my bots can interact with each other and can collaborate with each other. And so I've been creating a bit more of a scaffold of kind of where these artifacts that Gargbots create should live and how it can write to a place that's very legible to me. So I have like these frequent digest that I read every day and it kind of pushes to a database and I can just read it very easily. And so I would encourage power users to think about ways that Gargbot can actually write to like a single store where you can organize a lot of its outputs much easier. Damn. I have, we need another episode of going deep on Gargbot. Roman's Gargbot set up, which probably has way too much private sensitive information we can show it. But that's okay. That's an amazing tip. Roman, is there anything that you wanted to share or anything else you wanted to touch on before we get to our very exciting lightning round? Nothing. Nothing else on my side. We covered so much ground. I can't believe there was only an hour and a half each. I felt like we've been talking for ages and covered everything I was hoping to cover. With that, we've reached our very exciting lightning round. Have you got four questions for you? Are you ready? I am ready. What are two or three books that you find yourself recommending most to other people? Yes. Two books for you. One is, I love Kerbanigat. So Kat's Cradle has been a fun recommendation and a copy that I've bought many friends before. And then second is The War of Art by Stephen Pressfield that I find myself frequently coming back to even if it's just a page or two at a time and I'd recommend for anybody. War of Art incredible. It's like such a short book and it's like once you read it and it's not the art of war, which is what people might think they're here and guess. But it's the War of Art. It's a play on that and it's about the challenge of being creative and creating something new and how to overcome the resistance. I love that recommendation. Next question. Favorite recent movie or TV show if you had any time to watch any of these things? Yes. Recently, so every year I do a watch of Casablanca, which is one of my favorite movies. And it incidentally also has a character with my last name, Ugarte, which is like the only example of I think a Ugarte in the media. So Casablanca always a great rewatch. And then on the TV side, I sometimes sneak in an episode of Monk, the detective show, which was one that I watched kind of as a kid with my family and I've come back to now that I live in San Francisco. And it's just a great moment in time snapshot of San Francisco in the late 90s, early 2000s when it was shot, but I really enjoy. monk reference on the podcast. Okay. Favorite or most interesting AI product right now, you can say rock a lot if you want, but if there's anything else, you get bonus points. I've always been a big AI semantics search nerd. I love any sem search product, especially the kind of out of ordinary ones. So I was like a very early user of Metaphor at the time, which became exa, and I love kind of using exa to do all of these maybe more strange queries over the internet. But I see a lot of examples of people building like semantically search over, you know, an embedded image store of the momma or kind of things like that. And I always have so much fun playing with those. So anything semantics search engine over like a weird data set, I love. And exa in particular is one you'd recommend. I love exa. Yeah. Very cool. Okay. Favorite life motto that you often come back to in work or in life. Not as short as a single motto, but I love the Disadirata, which I don't know if you've read, but I have it on my, on my door, and I've had it since I was a teenager and everywhere I move, I kind of paste it there. And it's it's a very short poem, but each line I just, I find myself finding something new in it every time I read it, and I find it really grounding. Roman, this was amazing. What a, what a point in time, we're here right now. At this moment in time of Rockbot of AI in general, it's going to be a really fun to revisit this. I don't know any year and be like, wow, we were so right and so wrong about so much. Thank you so much for doing this. I know it's a very busy time on your team right now. So I really appreciate you carbon out a couple hours to chat. Is there any place you want to point people to anything you want to plug other than check out Rockbot? Is that check out Rockbot? Of course. And yeah, main thing would be please send feedback. I think we're in the very early innings of this still. I mean, we released a beta three weeks ago. And a lot of the feedback that we've been getting from this early set of users is directly translating to what we built and how we build it. And so really appreciate all the input that people are giving me. Nice job Roman and team. I know there's a whole team behind all this. Roman, thank you so much for being here. Awesome. Thanks Lenny. Bye everyone. Thank you so much for listening. If you found this valuable, you can subscribe to the show on Apple Podcasts Spotify or your favorite podcast app. Also, please consider giving us rating or leaving a review as that really helps other listeners find the podcast. You can find all past episodes or learn more about the show at Lenny's podcast.com. See you in the next episode.

Podcast Summary

Key Points:

  1. Brockbot was built from scratch with a small, isolated team to create a dedicated knowledge work product that feels like human teammates, not just AI tools.
  2. The core vision is to give users a team of autonomous AI bots that act as persistent, computer-equipped colleagues capable of handling real-world tasks.
  3. A critical decision was to run all bots in the cloud with their own persistent computers, avoiding local execution and device dependency.
  4. Early onboarding of hundreds of users—including non-technical profiles—revealed natural patterns like bot delegation and chief-of-staff roles, validating the product’s usability.
  5. The product prioritized simplicity and reliability over features, cutting unnecessary interfaces and focusing on seamless, human-like collaboration.
  6. Success stemmed from two foundational choices
  7. Brockbot’s rapid development (one month from first code to internal launch) emphasized speed, iteration, and user feedback over long-term planning.
  8. The product’s future lies in proactive, autonomous agents that monitor workflows, surface critical updates, and operate like human teammates across both work and personal life.

Summary:

Brockbot was conceived as a revolutionary AI product that shifts the paradigm from AI chat with limited tools to a team of persistent, autonomous bots acting as human colleagues. Built by a small, isolated team in just one month, it prioritizes simplicity, reliability, and real-world functionality over complex features. The core innovation lies in giving each bot its own cloud-based computer, enabling it to perform tasks that require pixel-level interaction—like clicking in dashboards or navigating software—without user intervention.

Early onboarding of hundreds of users, including non-technical profiles, revealed natural work patterns, such as bots being promoted to manage teams and handle daily tasks. This validated the product’s potential beyond coding, proving its value in knowledge work. Key decisions—like removing local device dependencies and hiding internal mechanics—reduced user friction and made interactions feel more natural.

The product’s success stems from its focus on user experience, not feature bloat, and its vision of AI as a true teammate. In the future, bots will become proactive, monitoring workflows, surfacing urgent alerts, and operating autonomously—offering a seamless blend of work and personal life management. This approach positions Brockbot not just as a tool, but as a new standard for how AI will integrate into everyday human work.

FAQs

Brockbot is designed as a true colleague with its own computer and persistent memory, allowing it to perform complex, autonomous tasks. Unlike tools that rely on chat interfaces or coding assistants, Brockbot enables users to delegate full workloads, including browser interactions and tool manipulations, creating a more natural, teammate-like experience.

The team built a small, isolated group focused solely on creating a knowledge work product. They developed a functional prototype in about a month and launched it internally, where early user feedback validated its potential. This rapid iteration allowed them to refine the product quickly before a public launch.

The team chose a fresh start to avoid the clutter and complexity of integrating AI agents into existing products. A standalone product allows full control over the user experience, removes technical barriers for non-developers, and ensures a consistent vision for how AI should assist in daily work.

A bot acts as a persistent, autonomous colleague with access to a computer and long-term memory. Unlike chat interfaces that are session-based, bots perform real work, manage tasks, and evolve over time—feeling more like a human teammate than a tool.

Each bot has its own computer and can perform real-time interactions with websites and applications, including clicking, typing, and navigating. This capability allows bots to handle tasks that traditional AI tools cannot, such as filling out forms or managing dashboards.

The team manually onboarded hundreds of users to observe real-world usage patterns. This revealed that users naturally formed bot teams and managed workloads, leading to design decisions like keeping the interface simple and avoiding internal tool visibility.

Chat with AI

Loading...

Pro features

Go deeper with this episode

Unlock creator-grade tools that turn any transcript into show notes and subtitle files.