The speaker discusses the transformation of Codex from a flawed, senior-engineer-only tool into a primary interface for knowledge work. They explain that six months ago, Codex was considered "trash," but recent updates, driven by competition from Anthropic’s Claude Code, have made it a powerful, general-purpose agent. The speaker now spends 80% of their workday in Codex, using it to access emails, Slack, Notion, and other data sources, and to automate tasks like creating run-of-show documents and triaging communications. They highlight the desktop app’s speed, organization, and ability to handle both engineering tasks (e.g., shipping PRs) and knowledge work (e.g., strategic planning, recruiting). The speaker compares Codex favorably to Claude Code, noting that while both are evolving, Codex’s app is currently faster and more intuitive. They describe a competitive landscape where model companies are racing to build the best agent management interface for knowledge work. Ultimately, the speaker advocates for trying Codex, as it offers a seamless, agent-first workflow that automates routine tasks and enhances productivity. They emphasize that once users experience the efficiency of Codex, they are unlikely to return to older tools.
Codex is one of those things where three months ago, six months ago, it was trash. If anyone from OpenAI is on the call and listening to that, I stand by that 100%. If you have a great general purpose coding agent on your computer, it's actually really great for any kind of knowledge work. If it can write software on its own, it can do any kind of knowledge work on its own. When I sign on during the day, Codex is the first thing I open. It is pulling in whenever I need from Gmail, Slack, Notion, Stripe, all of our data sources. It's where I spend like 80% of my time working overwhelmingly because the app itself is just so good. There's a new operating system for how and where you're going to get your work done and it's this kind of agent management interface. Hello everybody. Welcome to Codex Camp. Codex for knowledge work. Psych to have you. Psych to have you on this auspicious GPT 5.5 day after release day. Hope you're doing well. I'm here with our head of growth Austin. Austin say hello. Hello. We're psyched to have you. We are psyched to do this. It was really built for senior engineers doing pair programming. It would argue with you to make you feel stupid. It was just a little autistic. It didn't have any emotional intelligence. I think OpenAI had this interesting strategy or this interesting theory starting with GPT 5 that your vibe coding was going to happen in Chatchee BT and that was where all that stuff was going to live. Then senior engineers are going to use Codex to do all their programming work but we're going to hobble the model so it doesn't do anything bad. It's in a sandbox, all that kind of stuff. I think basically what happened is anthropic figured out that having a model that's pretty usable and fast and smart and also emotional intelligence on your computer that can access your computer is a really, really great experience for programmers and it means you could throw away a lot of the old stuff that used to have in a programming environment where you were built for typing code. You could just type commands into your terminal and then it would start working. Then I think what anthropic figured out is if you have a great general purpose, if you have a great coding agent on your computer, it's actually really great for any kind of knowledge work. If it can write software on its own, it can do any kind of knowledge work on its own and we started to move from this world where programmers had been delegating, had been delegating their tasks, starting to delegate their tasks inside of Cloud Code to now any kind of knowledge work is being delegated inside of Cloud Code and Cloud Code and all that kind of stuff. I think OpenAI, they had this original split. It's like, oh, you're going to do all your vibe coding in Chatchee BT and I think they saw what was starting to happen with Cloud Code and over the last maybe three months or so, they have done this hard pivot on Codex where it has gone from a senior engineer only tool that is really for pair programming to I think like it is my daily driver for this kind of work. I use Codex for everything from deep engineering stuff to writing to recruiting. I do a lot to actually do a fair amount of recruiting is really good for that and I'll give you some use cases later, but they sort of figured out that having this general purpose agent on your computer with the ability to write code, the ability to access your file system, the ability to have a browser and wrapping it in a desktop app is like the ideal next step for knowledge work and I think that they built the best current version of that. What it is starting to snap into focus now is that there's a new operating system for how and where surface, for how and where you're going to get your work done and it's this kind of agent management interface and that's whether or not you're using Cloud Code or Cloud Co-Work in the desktop app or Codex in the desktop app. It's becoming this race between the model companies where every each model company has their own surface like this for agent management, a desktop app for agent management that's at its core programming agent that's used for knowledge work and Thropic has Cloud Code and Cloud Co-Work, OpenAI has Codex, XAI recently essentially bought cursor and Google is the only one that I mean they have anti-gravity but I don't think no one is seriously using it for that yet but I imagine Google will do this too and that's the race, that is the race that's happening and so I think for us who gets who get all the benefits of being able to use these tools it's really important to be bouncing around between these so like use it for example using Codex so that you can feel what it's like to work in an agent first world because once you add an agent that is like your primary way of accessing and using software and the internet and all that kind of stuff it opens up all this interesting stuff that wasn't possible before because you can send your agent out to go talk to other pieces of software and come back and you know we can get into more of the details there but I want to get into like the more of the concrete use cases but that's the world that we're starting to live in you're doing work on your computer through Codex or Codex and your agent is your interface to a lot of the work that you're doing and a lot of the software that you use and a lot of the stuff that you do every day and that's actually really fun it's really cool there's a lot of good stuff here and so I wanted to bring Austin in to help do this because Austin was our head of growth and I think he had his real agent pill moment you tell me awesome but probably like three or four months ago and the agent pill moment was really Cloud Code and I sort of remember you just being like oh yeah on a Monday morning meeting like oh yeah I just was on my computer all weekend like I was like 12 hours a day I didn't go out or anything because I was using Cloud Code and you started to use it for all those all the kind of knowledge work tasks that a growth marketer would and over the last couple weeks as we've been using 55 and I've been telling you for a little bit you should try Codex it seems like you've you actually just shifted everything over to Codex in 55 and so I think you're a great person to talk about you know sort of what you're seeing and how and how that is how that is how this has changed how these agent management interfaces have changed your workflow and then why you like Codex and then I would love to get into some demos of your actual Codex workflow so that we can sort of see things from your perspective. Yeah that sounds great so I yes my like agent pill moment was spending a week going deep into Cloud Code in the CLI probably in like December, January, hooking it up to everything I do for work and for my personal life and finding that I use work as my like CLI interface and finding that the things that could automate the things that could handle for me and then the way it could work as a thought partner to make my work better was like this is the only way I want to do the kind of knowledge work that requires strategic thinking and data analysis and shipping marketing copy like a bunch of stuff that can get you spread out across a bunch of apps and tools during the day and and maybe in February you kept nudging me to be like you really should try Codex there were things you liked about it and I've if someone says that at every if anyone on the team says that like I'll go try it and I like to push myself and play around with more engineering-y tasks especially to see what these models are capable of and so I tried to build a personal vibe coded app in Codex because that was one of the things that you said that it was really good for and my immediate response was like I think it is better at building the app but I can't tell because it's nothing has ever made me feel more stupid than Codex like two months ago like I always I use compounded our campaign engineering plugin that Kieran class and made for basically everything including knowledge work but especially if I'm trying to build an app or ship a PR to the to the site so I made a plan in the plan it comes up with like three questions and for like which direction we should go and I had no idea what the hell it was talking about it was like do you do want any of these three and every question I was like please explain this to me in more detail and its response was basically like why like why don't you just do what I'm recommending and I found a way to I basically stayed in Codex for all engineering stuff because I did like the results you've been if I didn't love working in it but I would say 80% of what I was reaching for was was Cloud Code in the CLI and when we got our hands on the new GPT model a month ago the the the first thing I felt was at the very least there's parity between the latest opus model and the latest GPT model for the kind of knowledge work I do there's a few things that opus is better there's a few things that that Codex is better that feels a little more specific to me even like I actually
I'm kind of designed, which I still really trust, hope is for. It feels a little more like, oh, there's some stuff I like better than this, but the real differentiator to me is that, to me, there's no comparison for how fast and powerful the Codex desktop app is as just like an app compared to the Cloud desktop app. Like, I have never been able to get a cowork to work for me. And I think it's because I've been kind of ruined by the Codex app. It's so fast, the sub agents are so good, the way in which it suggests and then ships automations for me is just like, I can't imagine not using it. I wouldn't be surprised if any week the Cloud desktop app is just as good, right? They could ship versions where it's faster and better, but I'm now at the point where, when I sign on during the day, I Codex is the first thing I open. It is pulling in whatever I need from Gmail, Slack, Notion, Stripe, all of our data sources. This morning, I was like, oh yeah, we need to do a run of show for this camp. I message Codex, I'm like, make the run of show. It knows exactly where to look because we've already had conversations about what we're gonna talk about today. It pushed it to Notion, it sent it to Slack. It was perfect. It was like, oh yeah, exactly what we should do. And yeah, it's where I spend like 80% of my time working overwhelmingly because the app itself is just so good. And then the model has now gotten good enough to be the daily driver for me. - Yeah, and I feel the same way. I'd love to get into and interview someone asked, are you discussing the app or the CLI? We're discussing the app, the desktop app. And I think you're making a good point that both of these companies, I think, sort of see the end game here and they're pushing in the right direction. And for a while at least, it's gonna be a horse race where every couple weeks or every couple of months, like one is gonna pull ahead and have this like sort of amazing thing going on. And then the competitor is going to, like Anthropic, for example, I think we'll really selling in a couple of weeks or a couple of months. I don't have any inside information into this, but that will make it at least parody, if not better, and they're just gonna keep trading. And at some point, I think that'll slow down and you'll end up with sort of separate ecosystems. But for now, they're actually fairly easy to switch between. It's not trivial, but it's pretty easy. Like you can kind of ask, Codex, can you go grab all my cloud stuff and it'll go do it? - I think it feels the way when you do it, it's funny, I'm in New York right now, I usually live in LA. Most of my friends who are in the knowledge workspace have been asking me about like what they should be using. They're all cloud code or cloud desktop app code. And when I tell them that I have fully transitioned to Codex, this like look of horror shows up on their face and they're like, do I, do I, do I really have to? And I, of course, tell them they don't, but I'm like, you really should, right now, you really should, like I think you would get a big benefit from it and I've been showing them why. And it's interesting and to me, I'm surprising how resistant people have been to it because when you're a knowledge worker and you have all these new tools, the cloud desktop app is game-changing. It's amazing, right? So the idea that the Codex app is maybe 30 to 40% better is like, whoo, 'cause that's a lot of work. Which we can get into kind of how I migrated. I can show some of that. It was very easy and the ways that I'm starting to use it. So I'm happy to dive into that and start sharing my screen and show it work. - Yeah, why don't you share your screen? I think, yeah, I kind of agree with you. It's more of like an emotional thing of like, oh, I have to learn a whole new thing or whatever, but it's pretty similar. Yeah, I would love to see some of your workflows. - Cool. So this is the Codex app. I was gonna do like a very, very quick tour. I think a lot of the audiences has seen it or kind of like where I go and how I use it. One thing I love about the Codex app is like, I do think it's much better organized than the cloud desktop app. My ability to have these folders with persistent, consistent chats inside of it that I can go check out. And then especially like the big differentiators that because I do think this is much better for engineering for like occasionally, I will ship a PR for one of our products. It's great to not have to switch between the Cloud Code CLI, the Cloud Desktop App in Codex that I can be here. I can be working on our improving our KPIs sheet, which I'll like show when I was doing here. And then I can go down to plus one and ship a PR for plus one. And the other thing I found, 'cause I tried the new, I tried the updates of the Cloud Desktop App plus week when they shipped it. And the stress test I put on it was, make a go-to-market plan for our new product and ship a PR to sparkle in different chats. And it was so clunky and slow. And when you do stuff like that inside of Codex, it just works. It just works really quickly and well. And that's the thing that like once you start feeling that it's very hard for me to turn away from it. So I have these different folders for some like vibe coded apps that I play around with for fun for my personal open Cloud where I can go and manipulate it here. And then the one with all of the chats is this like every growth OS, all this is a folder with a bunch of secrets and keys so it's connected to everything we use for every and then some project instructional files that explain what the every business is, what we care about, how we like to work together. It has some like reviewer agents inside of it that are all informed by how Compound Engineering works inside of Compound Engineering, Kiran's plugin. There is a Compound Engineering review step once you do some work that reviews for like security and a few different things that are oftentimes not as helpful for like I'm doing a strategic plan for a go-to-market initiative. And so inside of here there's like a fork for it for strategic alignment with the company goals for data accuracy. And having that inside of this folder means that as I'm making plans I can get reviews from the model in like a targeted way. And so the first thing I wanted to show is like how that I was talking to our manager and chief Kate yesterday to show her like how I would recommend getting started in Codex. And this is my recommended prompt. I'm happy to put it in the chat for people. We can put it in the email as well. And so all I did was I'm putting in the prompts here. I only have posted panelists access. So I'll send it later for some time. - Okay, yeah, maybe read it out. We can all agree that housing is expensive. Rent or mortgage doesn't matter what you're paying. It stings every month. But builds can make it feel a little better. Build started out by rewarding members for their rent. Now as of 2026, build members can also earn points on mortgage payments wherever they live. Every housing payment earns points you can use toward flights with top travel partners like United in Hyatt, Lyft Rides, Amazon.com purchases and so much more. This is actually pretty cool. And I have some friends that use this in like a lot. Something that's underrated is that build members also get access to a neighborhood concierge. They can make restaurant reservations, book fitness classes and find new local spots, all while still rewarding you at 45,000 merchant partners. It's like having a personal assistant baked into where you live. It's simple. Being a renter and now owning a home is better with built. Make sure to use our URL so they know we sent you and now back to the episode. Yeah, yeah. OK, I can zoom in as well. I think there we go. So through the plugin tool with Codex, I went in and like did the manual clicks to connect all of the tools I use every day, like Gmail Slack notion. And then I went to a new chat inside of this folder that was built through Cloud Code. Cloud Code built this whole every growth system. There's a Cloud MD file in there. And it saved locally. It's also St. St. St. St. and pushed to GitHub. And so I just opened that project inside of Codex when I started working here. And I started a compound engineering brainstorm workflow. Because then again, just kind of like a thing I reach for of, let's think about this thing together, me and the model. And basically I said, it's like go take a look at the things I use the most, which are notions Slack and Gmail. And think of some automations that would help me with my work. I find that when I'm trying something new, whether it's a model or an app, having an agent, having a very smart frontier model, like tell me how to use it, tell me what it should do, is like exactly where I want to start, rather than thinking of it myself. And I usually start here. Sometimes I will describe a very specific problem. But this is very helpful for me. And I think a good generic problem for people to start with. And Codex comes back. It looks at what's going on for me and for the company right now. And I thought these were really good. That like, it has this kind of follow up radar. This is a big thing that happens with people who do knowledge work, who do partnerships, who do social media marketing. That there's all this stuff coming at you across a bunch of different sources. Like, what if it handled the triage for you? What if it had this kind of like command sensor when we run a camp or an events, which usually requires a bunch of moving pieces and moving parts? Like Dan mentioned, for recruiting and hiring, we don't use a tool like Ashby or something. We kind of have it all synced through notion, because apps like this and agents can kind of like handle a lot of the pipeline and tracking work for us. And you can just ask it to automate it for you. [BLANK_AUDIO]
So it does that and it asks me, which ones look good, what do you want to tweak? For the sake of this demo, I didn't give it any real feedback. I was like, looks good. And this is actually the thing I've always been most impressed with Codex for and for the models is that it's like, great, I made this automation for you. And I do find that they just work incredibly well. They require very little tweaking to be like, this is a thing I would and do use every day. There's this set of instructions that it comes up with based on what it knows about me. I can change when it runs. I can give it additional insights. I can connect it to other things. But mostly it just works. There's one that works for me that just at the end of each day now compiles all of the stuff that I haven't responded to yet, drafts their replies and we can kind of like, not get out together of what to say or like, actually all I need to do is just give like a thumbs up, slack reaction to something and it'll do that for me. It's kind of like a dumb agent. Like I think of agents like this is like the dumb ones that just do the right thing every time. And then the smart ones like an open claw or a plus one the products we have coming that's like, you'll work back and forth with it and like have a, have like a more like a creative strategic partner. And Codex is good at building both. And I can show kind of like the smart agent setup. But if someone is looking to be like, can I see what this thing can do to help me with knowledge work? I would start here in like a brainstorming automation state because it is, and I think you'll also be surprised by how fast it is. And you'll think, oh, I'm starting to get what this thing can do. - This is so sick. I, your codex used to just far surpassing mine in terms of interestingness. I'm getting a lot of ideas. I want to just actually pause here. Normally you take questions at the end, but I think it would be kind of interesting. If you have a question about what Austin has just showed, it would be nice to let people come up and just ask a question or two, just to see what the vibe of the room is like. So please raise your hand if you have a question and we will call on you. Margaret, welcome. Please introduce yourself and ask your question. - Hi, can you hear me? I'm Margaret. I'm in Plymouth. And my question is, what is your review step look like? So it's saying don't send post archive or modify without explicit approval. So what does that look like? Is that like do you call up say, hey, let's do the review flow now? Or is it doing push notifications to your phone? Or what? Thanks. - Yeah, so for this, what I prefer, and I was actually talking to a friend, a dinner last night, who said they did the same thing on their own. They came up with this too is like, everything, I work primarily in codex. I do all the drafting and set up in codex. And then it's helped perform my brain to have the final review step actually live in the external app. So it will draft all the Slack messages. And then I can go to Slack or Slack has that like, draft reply tab and I can go and knock them out. And I do find that it like, freshens up my brain a bit to be like, here's where I'll just make sure that this is what I want to send to a human being. Same thing for email. It like creates all of these drafts in Gmail. And I'll actually go open Gmail and look at them and knock them out. I know some other people who just have it actually come up inside of codex. And they're like, yes, you're saying that it looks good there. I do the same thing for strategic planning. It pushes to a either proof doc, the like agent friendly markdown file that Dan made or a notion doc. I use them for some different things. And I just like for like the last pass before humans engage with it to step away from this agent space and have a final check in another surface. That's really the only time that I'm like leaving the app to do something. That's brilliant. Thank you. Sweet. All right, we'll do one more and then we'll keep going. Alex, please introduce yourself and ask your question. Hi. My name is Alex. I'm a musician and I do a lot of gigs and get emails from clients all the time. So I have to sort my leads from my newsletters and all that informational stuff. So how do you make sure that you prompt codex to keep those emails safe for me? The ones that require a personalized response. And I just want to make sure that I don't send something that loses me money or something. Yeah. So for me personally, I rely a lot on Corea, the app that Kieran runs at every for the AI email assistant. That's a part of the every subscription. It's really helpful now that inside of Corea, there is a CLI and API connector that I can work in codex and tell Corea, which is managing my email, filtering, and my email rules, what I want and what I value. The way I do that is the same thing I would recommend whether you use Corea or not, which is to have the agent interview you to get an understanding of what the rules should be. I always find that I get a better result rather than saying what I think the rules should be. And so I'll do a brain dump using monologue, our Speech Detects app, saying, here's the problem I'm facing. My email's a mess. Let's figure out how to triage it. I think it would work perfectly well if you wanted to try starting it as an automation in codex or a rule in codex of saying, I think these are the things I want to make sure I get. I think these are the rules I want to set of like never send anything for me, only draft. I think I want to go through all emails at 3 p.m. on weekdays, but go take a look at all my email, go do a search, spawn sub agents to do a search. I'm always selling codex, spawn sub agents to do different types of work across different workflows. And then come back with a plan. Come back with a plan for like how you're going to set up my email and then you can read the plan and see, oh, it looks like it's actually going to brief or summarize or auto archive something that might lead to making money for you. And that's where you can tweak it. And then the other step I take is that I set reminders for myself. And I use to do lists for all of my like reminder task tracking. It's also connected to codex. So I can just message codex or message my open claw and say like, just add this reminder to my schedule to check how the new automation is working and like do an audit of it. So you can see like, it's been 72 hours. Let's see if I'm seeing if I've missed anything. You can prompt the model and see like what have you been archiving. And I find that really helpful. But I'm really excited by all of the work or product gm's that every have been doing to make it so that I can just prompt codex or a cloud code or inside of cursor any agent to manipulate those apps how I want. It works really well with our other tools also. Thanks Alex. I will add to that like one of the things that we found. Basically because Austin started doing it, I sort of was like, oh, that's really interesting. Is that Austin started? We have plus ones, which is our hosted open claw. And Austin started setting up his plus ones with codex and cloud code and realized that it's just a much better experience. So rather than, for example, the earlier version of plus ones, we had like a whole dashboard and a whole onboarding experience where you had to kind of manually click a bunch of buttons and give it a lot of context. It's much easier if we just expose plus ones via a CLI to codex or cloud code. And then you can just like talk to codex. And it will take everything it knows about you from your computer and your past conversations and throw it into a plus one setup and Austin showing this. And it's like a really powerful. And it's part of what I'm saying about how the world is changing when you assume every user has access to an agent like this. Because we don't have to have a settings dashboard. We don't have to have an onboarding experience. We don't have to gather as much context manually. It can just be given to us for free by codex. And that's really interesting. Yeah. One of my favorite use cases was I got-- I got really inspired by this interview Claire Vo did with Lenny where she said how much of a breakthrough she had when she stopped trying to just use an individual open clause as a master supercharged open clause and had this suite of six specified open clause. I think that applies to any kind of agent. There's the new chat, chapetit, like provisional agents. I got hooked on that. I think Claire's point was really good. And my path towards making this suite of agents to help with the growth function of every was just going to codex, going to this folder. I actually just sent it the transcript of Claire's interview with Lenny. And so I want to do this too, given everything you know about me and my work, make a plan to suggest six agents that we should provision into our Slack. Consider the fact that we might want to make some of them notion custom agents, which I find work really well. It's just like do the same thing every day, every time. Some of them might need to be smarter on emissions, but like do that, come up with a plan. The plan that made was really good. And like, I--
I tweeted a bit after seeing it, but there was that. And then now I have this suite of six agents in our Slack that worked really well for me. They still break. I find when you're making open calls and personal agents right now, you should accept they're gonna break a bit. But the really powerful thing is that rather than going back and forth with the agent or getting frustrated, I just go to Codex and I'm like, I either screenshot or I can add Slack in Codex and say, go find this conversation or the stupid thing happened and fix it. And it does a really good job of just like changing the architecture of the agent and making a fix from there. - I love that. Yeah, it's just such a step-changing how you work. And now I wanna paste that clear interview too. (laughing) - I want to show one thing that I like, this is like how to actually my favorite way to use the stuff for knowledge work. It's a thing that I wish I had for so much of my career 'cause this is one of the most time consuming to me frustrating things about knowledge work is that we are doing a real go-to-market public launch for a plus one, we're very excited about it. And we've been having a bunch of internal meetings and Slack conversations around like, how are we taking this to market? What is the strategy? What are we gonna do? And we've done all of the work that like, kind of like only humans can do, the like marketing case, the business case, the like the narratives and stuff. Not all of it is as refined as it used to be 'cause it still needs to be refined, but it's all sitting somewhere. And I had all these plans this week to make the go-to-market plan, which is like one thing I'm responsible for. And an inevitable thing that happens that happens at everyone's shop is like, all this stuff came up like, I've got to do interviews for hiring. We found out the release date for the new Jet2PT model. And so I had a day, I think it was Tuesday in between meetings where I'm just kind of like, I'm prompting Codex this way. Hey, I kind of done most of the work, right? Like in our notion, every meeting is recorded in a single place in all the transcripts are there. We talked about this a bunch in Slack. I have a template for a go-to-market plan that I really like. And I can go to Codex and say, like, can you just make the plan? And in my head, what I'm thinking is like, maybe it'll get like a six out of 10, or a seven out of 10. And we can keep nudging and I can keep like going along. And so it does that. What I'm asking you for is like, why don't you start by doing the compound engineering brainstorm step to just ship a proof doc and I can see how close you are. And I, one thing that it doesn't really do super well unless I tell it to, and I want to install this, like a workflow is that it doesn't go read or calendar of upcoming posts and launches. And so as it was going, I was like, oh, you always forget this. This is the message I'm sending of like, actually look at everything that's scheduled 'cause I have to account for that and go to MarketPlan. And then it makes a plan as a proof doc. I went and looked at it and I was like, again, I maybe have five minutes in between meetings and I'm like, this is really good. Like you kind of have every, you have the architecture enough that I want you to like factor in one other change and then just shift the plan to notion. And the plan at shift to notion, I was reading it and I was like, this is basically 80 to 90% of the way there. And that's not because I'm relying on the model to come up with our go-to-market strategy. It's that I'm relying on the model to look at all of the things that we've already said about the go-to-market strategy. He's it together and then review it, right? Come with what's not. There's a lot of important context loading that happens here. We're like, it knows what our target ICP is and knows what our goals are. It knows how we think about narrative positioning. And before this was possible, the only thing I could have done was either block off a whole day to sit and do this or get done with my work for the day at like six or seven and then stay up all night writing this. And this has been such a game changer for me. And the other part of it that I think I've found is really helpful is that I don't make this plan for humans. I make this plan for humans and agents and primarily for humans to understand through agents. And so when I send it to the team working on the go-to-market, they can read it and it's like digestible to humans. But the thing that it's really helpful for is like it's the full plan sectioned off all in one. And so Brandon or COO who's like deep in this product can ask his plus one, can ask code, as he has called code, like, let me know what Austin's plan is. Like summarize it for me. Let me know the business case. Brandon has to come up with the pricing modeling for the plan. So he can work with an agent against the plan. And as someone who spent so much time in my career thinking about like literally how the proposal or go-to-market document looks like, how is it going to look when I present to the CEO? Like this two-page plan for like a budget I'm asking for. Like is it going to make sense to their eyes and like really fine-tuning stuff? Giving up on that and just saying like, is the plan really good? And is it going to make sense to like, Dan's agent if he approves it? For to me, it makes me work faster. It makes the work better. It means that I don't have to think about all this like, kind of dumb stuff that doesn't matter. That like, it's to me a much more like powerful and fun way to work. - I totally agree with that. I just said so many things that are interesting there. The first one is just normalize sending agent documents around. And that's why we have proof. It's just such an easy way to send the markdown documents that we generate to each other and to review them together. And it's like, I think there's this whole strand of AI stuff that's like, make AI write in your voice. We even do this with spiral. But there's this other strand of just like, normalize AI writing because I would actually prefer to read your agent's writing than your writing in a lot of cases because I know that it's just easier for you to get all that thinking together in a format I can read. If you have your agent write it, the thing I care about is do you stand by it? Have you thought about it? And if I talk to you about it, it will be clear that if I talk about a particular bullet point in it, you've thought that through. And as long as we have the trust that that's gonna be the case, then I absolutely prefer the agent version. In the future, humans face a new problem. What do you do when your computer is doing your work for you? One answer? Think a Claude Walk, an idea by Avery. Avery, the only subscription you need to stay at the edge of AI. - Totally, like my friend Rachel Cardin who runs the great like, Sub-Sex Newsletter, like, you'd buy a video about social media, had a really good piece this week about frustrations for people working in social for like this pressure they feel that everything is run through AI and the quality going down. And one reason why is that there's that dichotomy of like, what do you actually stand behind? Are you running something through AI and you, like, maybe your manager did it and they don't even know what it said. And the thing I love about working at Avery is like, you show up to a meeting, you've, like, shared an AI written document ahead of time. And the expectation is that you're going to stand behind all of it that someone will ask a question of what's in that document. And you, if you say, like, oh, I didn't even know that was in there, it's like, you're exposed, right? But the other nice thing is that we continue to keep investing in skills and workflows and tools to kind of ensure that never happens. Like, I have rules inside of this project file to be like, if don't add anything that I haven't, like, said in another context, I want your suggestions. Send your suggestions to me and the chat, but don't put it in a document. And like, depending on how big the context gets, these models can follow or not follow those rules, which is another reason why I always leave codecs for that final review before it goes to the, like, humans I work with. Yeah. And I think that the, that last thing that, that I want to point out that you said is like, a lot of the time that you spend working is about taking thinking you've already done and putting it into a form that other people can read and consume. And the important part is doing the thinking. There is something obviously about, like, I love writing. Writing is a good way of thinking. And sometimes you actually want to do the writing yourself because you want to think about it for certain types of things and certain types of people. But there's a lot of stuff like company strategy where a lot of the thinking happens out loud in meetings. And there's also times, like, for example, I'm writing something that's sort of like a, it's like a retrospective on the last three and a half years of AI and like where I think we're going. And that's so hard to sit down and write, but it's much easier to just like dictate. So I just took a monologue note where I was just like saying stuff. And I'm using the AI to help me, like, figure out what I'm really trying to say. And in those cases, I think it's just so nice to record stuff, give codex access to everything, and then just have it spit out a strategy doc and go through it to make sure it's stuff you agree with. But it's such a time saver. And especially if you're someone who like awesome or like me, like you're in meetings a lot. And so you don't necessarily have huge chunks
time in your day to like go do a big strategy document because you're just trying to stay on top of whatever's happening. It helps you do that in the cracks of your day and do a lot of that thinking and I just I love it for that. Yeah, me too. I want to show one more thing where we get into more questions because it's like I wanted to show kind of like a more like mix of knowledge work and engineering-y stuff that like would never have been possible without these kinds of tools and that I really love Codex4 which is I've been rebuilding our KPI tracker every week. I'll just like show it here for a bit. So we have so many different parts of our business at at every and it's very difficult to get all of those data points in one source of truth in a traditional tool like even post-doc which I'd really like and a lot of our data runs through it. To get one dashboard that is again both human and human and agent facing that is up to date with all of the metrics we care about. I haven't found a great solution for just like you know going to post-doc and having it having to do it. So I've been rebuilding our KPI sheets inside of notion with the goal in mind of anyone can point their agent to look at it and see how our new paid subscription trials doing, how our page used doing, how is monologue IOS, MRR doing, all versus plan, all of this stuff because one it helps you work as a human but it also really helps you automate agentic work so that you can say like if you're agencies that we're tracking behind on SEO for a keyword we should be winning on they can go just like ship a bunch of landing pages for us to try to win more on it if the if the source of truth is good and so I have been doing this big kind of like to me complex workflow problem and codex of let's build the sheet together let's have it live in a notion database that all of our agents can point at and I've done a bunch of different versions of it the first version was like can codex one shot this right like it has all the API keys it has everything I'm happy to give it the context on like how we measure MRR and everything and each time it was like a little off it was like maybe five to 10% off of the formatting the numbers the framing and our MRR MRR number can't be 5% off like we can't run a business with a source of truth is even 3% off it has to be just exactly right and so the thing that I force myself to do and it's weird now I'm like it feels so stupid that I have to do this but it makes sense is like I'm going column by column and to and to ensure each column is exactly right and defensible because it's the only way that we can run and grow the business reliably and especially the only way we can we can confidently unleash agents to go take actions against what's happening in that kpi sheet and it's like it's so interesting to me that I'm frustrated that I have to do this that the that the model can't do it for me but it's just because of how like powerful these models are gotten that I expect it to be able to do it but and this is the thing where I'm like you know it's using notions workers tool which is this like dev tool to build always on tool calls of our stripe of our social it's like creating little scripts and stuff also if I don't really understand but I understand the outputs I understand that the output is a notion database that updates every six hours with all of our metrics and it's just nice that I can do that and I don't need to hire a consultant to do it or like I don't need to like yeah take away from our like our engineers times that that work on our data like I can do this now and I can do it just by like prompting the model and understanding how the metrics are supposed to work it's amazing I can't wait is the do you think it'll be ready on Monday? okay it'll be ready on Monday yeah feeling really good because we've been I mean just having it turns out that figuring out how much money you're making and how much you've grown is truly a philosophical question you know and you actually do need to like go in and like set that frame and so we've been dealing with an outdated sheet because it's like it's pulling numbers but is are the numbers correct you know even even outside of AI and and there's no one way to for example measure MRR you just want to do it the same way every time so you have to decide and that's kind of it's kind of wild that it's like almost impossible to tell how much money you've made in an objective way you have to just like pick but anyway that's just the way my brain works I want to say before we get into questions one other thing that I use this for that would like blew my mind from a knowledge work perspective is recruiting so we're hiring a lot and we're looking for an L&D head of L&D someone to help us run courses and there's this company in New York called General Assembly and when I think about people who've run like really great courses about technology to teach people how to get hands on with like programming or design or anything like that like they're the company that I think of from the like 2010s in New York and so I my theory was that if we're hiring someone to do like build our courses they would probably have a good person probably would have worked at GA and Jason yes and I think J.A. is quality has gone up and down but at the beginning they were amazing and so what I did was I just said to codex hey like can you find can you just get a list of GA alums I'm like hiring an L&D director and then I want you to filter and sort the list by people who have subsequently gotten into AI and it did it like it just gave me a list of people the first one I clicked on it was like I was like this guy is perfect and then I looked and he followed me on Twitter so I just DMed him and like I don't know if we're gonna end up working with him but like it was just one of those holy shit like bulb moments where normally what we're doing is sorting through a ton of applications and like trying to find the right person and I we're still gonna do that but especially for any kind of like outbound effort it can kind of find that needle in the haystack that you're looking for really really well so I highly highly recommend okay we've got about 10 minutes left and I want to take some more time for questions if you got a question please please raise your hand one thing that we have not gotten to actually is that if you are here you're getting codex credits Austin do you want to go through that really quick yes so open AI has given us a code I'm about to drop into the chat for 250 attendees of this camp to get a free month of chat GP to chat GPT Pro lights that's about a $100 value and you can redeem it at this link that we will drop in the chat right now sick Dan I'm actually gonna I'm gonna slack it to you so you can drop in the chat because some reason I don't have access okay I'll do that um so yes this is this is our gift to you as every subscribers we try to do stuff like this all the time we've been we've given out I think we've done cursor credits we've done we've done a lot a lot of other stuff we have more stuff like this coming so we just want you to be able to try these tools be at the edge with us and we just love having you as subscribers so here is the link it's a hundred dollars oh no shun yes we did give out no shun it's a hundred dollars and check it out we will send it out in an email maybe we actually may not send it out in an email because it's only 200 limited to 250 people and that's pretty much exactly the number of people who are here so if there's any left we will send it if there is one person there's 251 people here so if that person you let us know we'll figure something out for you interesting not available on my plan okay we will have to deal with this let us let us figure out what to do what to do here so correction this is only if you do not have a plan this is for new users and we'll try to we'll try to get something for existing users and send it out as soon as we can cool all right let's do some questions um rich please ask your question so I I saw at the beginning you were using kind of how engineering is kind of part of your workflow are you using kind of like the off the shelf plug-in or is there tweaks to it and kind of where does that work and maybe not work when you're outside of you know kind of code creation workflow so I find that there's no overwhelming need to fork your own version of compound engineering I used it for a long time um for all of my knowledge work and it was extremely powerful for me and then
And maybe about two months ago, the main thing I noticed was reading the agent's response to especially the review stage of watching the reviewers that Kieran and Trevin had built that are very specific to engineering. I was like, oh, this like, the thing you'll see the agent do is say like, I'm supposed to go through this review step. It looks like it's designed for engineering. It's thinking about security and front end design when this is a go-to-market plan. The agent will then like change the path. The agent will be like, I'm going to review this for something else rather than reviewing it for security. And so the thing that I did was I went and forked a version of it that is actually publicly available on our GitHub called Compound Knowledge, which is built exclusively from me taking the Compound Engineering plugin, which is also public and you can go forth. And going inside, I think I started in Cloud Code now I updated in Codex and saying like, I wanted to tweak this to general knowledge work. And this is the thing I was referencing around the reviewers being much more specific to knowledge work around strategic alignment and data accuracy. I think more than anything, this is like a really fun way to learn and a fun way to like push yourself on using models. You're welcome just to go use this one. We'll include it in the follow up email to the camp. But I think it's a cool, like I learned a ton just by doing this. I had never made like a plugin like this before. And to make your own version of say you do like social media marketing and you want to make sure all the reviews go through your style guide, your like past performance. I got a ton out of operating this way. If you just want the Compound Engineering to make your work better, it absolutely works really, really well for knowledge work just kind of out of box. Got it. Yeah, no, interesting, particularly using kind of all the the end of step pieces like Compound that that's still apparently a valuable step for you. Yeah, the Compound step is really valuable. We have inside of our notion a go to database of after you're done with a session, you can send the learnings from the session to actually a team wide shared Compound source of truth. Whenever I'm done with any session in codex or code code, the agents are instructed to ask me, should we compound this, save it somewhere for the learning and should we turn any workflow from the session into a skill so that we can just do it automatically each time. Got it. Cool. No, I'll check that out. Thanks. Cool. All right. Rory, please introduce yourself and ask your question. Hi, my name is Rory and I'm in your head. Are there anything about the way you work at every, like maybe taking some time after meetings, like ending them a few minutes early so that you can do those things that you'd recommend to teams that are adopting workflows like yours? Was that clear? Yeah, I think so. I'd like to say it back to you, like what I'm hearing, which is a very real challenge here, is that it's so exciting and tempting and alluring to spend a lot of your day playing with stuff. Also spend a lot of your day continuing to push on. If I just get this automation right or this tool right, my work is going to be like 100 times better and easier and I actually find myself on a lot of days spending most of my time not in meetings trying to build really good tools and automations that work well and not making the time to do the actual tasks that have to push the business forward, like shipping the social posts for the day or whatever. And I don't really have an awesome answer for it. That outside of the fact that the playing around one is like kind of core to how we operate at every. It's a thing that Dan pushes all of us to do. I just wonder why I love working here. It's also like to me the best way to learn and makes me better at everything I do. And then the only kind of guidance I've given myself is that like these automations in Codex keep me on track to get the work done so that when I'm too deep in like playing around and building this like there's like a social automation tool I'm working on that I've been like deep in for a while. The Codex automations make it so that I like make sure Brandon gets what he needs for this like some like business plan we're doing. I do find myself over indexing on learning and playing because of how exciting and powerful the models have been and that more I have to continue to pull myself into the required day-to-day tasks and the version stuff that's happening. Yeah. And I also sort of read your question, Rory, and you tell me if this is wrong, but as like how do we do more of the AI stuff, the more of the playing even to even get started on this stuff in our day-to-day if we're like busy all the time. And what are the organizational practices that we have for that? And yeah, I just think like Austin said, it's just like a culture, it's a cultural thing. We just love playing around and that's like that's part of our job. And I think there's this thing happening right now where the tools and the workflows are changing so fast that just focusing on how your job currently works, you can run as fast as possible and someone using a new tool with a new paradigm and a new workflow is just going to be you by default. And so if you just give yourself some time to play around, it may feel like a waste of time, but you're leveling yourself up to a different game at a different level. And I think that's really important and some of the organizational practices that we have to help people do that are really around. So one of the things we do twice a year is called Think Week. And we just literally don't do any of our day-to-day work and we just spend a week together just like playing around with new stuff and building stuff and learning and being together. And you don't necessarily have to do a whole week of that, but I think it's really good to maybe do that once a quarter of a day or something like that and just give people the time and space to if you can. Sweet. All right, y'all. So that is our program for today. Thank you for coming. We love seeing you. We love doing this with you. Remember, every is the only subscription you need to say at the edge of AI. We would love it. If today you would go tell one of your friends to go subscribe to every. We want to get more people in here. We just think we're right at this amazing point in history where we get to surf, ride this big wave together and figure it out together and please tell your friends. See ya. Thanks, y'all. Who is the epitome of awesomeness? It's like finding a treasure chest in your backyard, but instead of gold, it's filled with pure unadulterated knowledge bombs about chat GPT. Every episode is a roller coaster of emotions, insights and laughter that will leave you on the edge of your seat, craving for more. It's not just a show. It's a journey into the future with Dan Shipper as the captain of the spaceship. So do yourself a favor. Hit like, smash subscribe and strap in for the ride of your life. And now, without any further ado, let me just say Dan, I'm absolutely hopelessly in love with you.
Podcast Summary
Key Points:
Codex has evolved from a poor-quality tool six months ago into a powerful daily driver for knowledge work, not just programming.
The speaker uses Codex for 80% of their work, integrating with Gmail, Slack, Notion, and Stripe, and prefers it over alternatives like Claude Code.
The key advantage of Codex is its fast, well-organized desktop app with effective sub-agents and automation capabilities.
The tool is shifting from a senior engineer pair-programming tool to a general-purpose agent for all knowledge work.
There is a competitive race among model companies (OpenAI, Anthropic, XAI, Google) to create the best agent management interface.
The speaker successfully migrated from Claude Code to Codex, finding it easier to use and more efficient for tasks like creating run-of-show documents and automating workflows.
Codex’s ability to handle both engineering tasks (e.g., shipping PRs) and knowledge work (e.g., strategic planning, recruiting) makes it a versatile daily tool.
Summary:
The speaker discusses the transformation of Codex from a flawed, senior-engineer-only tool into a primary interface for knowledge work. They explain that six months ago, Codex was considered "trash," but recent updates, driven by competition from Anthropic’s Claude Code, have made it a powerful, general-purpose agent. The speaker now spends 80% of their workday in Codex, using it to access emails, Slack, Notion, and other data sources, and to automate tasks like creating run-of-show documents and triaging communications.
, strategic planning, recruiting). The speaker compares Codex favorably to Claude Code, noting that while both are evolving, Codex’s app is currently faster and more intuitive. They describe a competitive landscape where model companies are racing to build the best agent management interface for knowledge work.
Ultimately, the speaker advocates for trying Codex, as it offers a seamless, agent-first workflow that automates routine tasks and enhances productivity. They emphasize that once users experience the efficiency of Codex, they are unlikely to return to older tools.
FAQs
Codex is a desktop app from OpenAI that started as a tool for senior engineers doing pair programming, but has pivoted to a general-purpose agent for knowledge work. It can access your computer, file system, and browser, and is now used for tasks like writing, recruiting, and data analysis.
Codex is praised for its fast and powerful desktop app, with sub-agents and automation suggestions that work seamlessly. Claude Code is also strong, but users may find switching between them easy, as Codex can import data from Claude setups.
Codex can handle deep engineering, writing, recruiting, creating run-of-show documents, pulling data from Gmail, Slack, Notion, and Stripe, and shipping PRs. It's used for 80% of daily work by some users.
Connect your daily tools (Gmail, Slack, Notion) via the plugin tool, then start a new chat and ask Codex to analyze your usage and suggest automations. Let the model guide you on how to use it based on your work patterns.
Users find Codex faster, with better organization through folders and persistent chats, and more reliable for complex tasks like creating a go-to-market plan and shipping a PR simultaneously. Sub-agents and automation suggestions work more smoothly.
Yes, Codex is designed for general knowledge work, including strategic thinking, data analysis, marketing copy, recruiting, and project management. Its ability to write code makes it versatile for any knowledge task.
Chat with AI
Loading...
Pro features
Go deeper with this episode
Unlock creator-grade tools that turn any transcript into show notes and subtitle files.