The podcast discusses Google I/O's major announcements, emphasizing the transition from the AI era to the "agentic era." Key launches include Gemini 2.5 Flash, a powerful and cost-effective model optimized for agentic tasks and coding, and Gemini Omni, a unified multimodal model that can generate and edit text, images, video, and audio. Omni is seen as a tool to empower creators by simplifying video remixing and storytelling, potentially moving content away from "slop" toward more thoughtful, engaging formats. For developers, Google introduced "managed agents" in the Gemini API, which drastically lowers the barrier to building agentic applications by allowing a single API call to orchestrate multiple models. The conversation highlights that the agentic era is defined by asynchronous agents that work in the background, freeing users from needing to be actively in control. This shift creates a significant opportunity for startups to build agentic products that solve real problems within familiar interfaces like text or email, rather than requiring users to adopt entirely new workflows. The "anti-gravity" ecosystem was also overhauled, providing a comprehensive suite of tools—including an agent manager, IDE, CLI, SDK, and API—to support agentic development across different user preferences. Overall, the message is that the future of building and creating is now agentic, with lower barriers to entry and powerful new capabilities.
We are here. It's Google I/O, Logan Kilpatrick from the DeepMine team, friend of the pod, been on the pod, deep times, Logan by the end of this episode, what are people gonna learn? We're gonna hear about all of the new releases from Google that just happened to that, Google I/O, and also, like specifically, we should deep dive on like how to build new AI agent-native products, which was sort of the thread of a Google I/O this year, was agent-sagents agents, which we should talk in depth about everything that we launched and what it means for builders and developers. [Music] Okay, cool. So, I mean, let's start off by, okay, what was launch and what is it matter? Yeah, there's so much new stuff, and I want to get your reactions to this too, because I think you have a grounded perspective of sort of the technology. I think one of the highlights was Gemini 3.5 Flash. It's the best model we've ever shipped and we've made available. Really, sort of, if you look at the history of Flash, I think Flash started as sort of this like smaller workhorse model that was really great for chat, and was sort of the, you know, very cheap to use, very cheap to run. And I think we've sort of Flash continues to evolve to meet the era of what people are actually trying to use the model. I think the era that we're in right now is people are trying to use the models to do agent-tick, sort of long-running tasks. And I think we want Flash to sort of be the workhorse model for the agent era for agent-tick, long-running tasks, for coding, for all that stuff. So you see a Flash model that's actually really great at coding, that's sort of competing with a bunch of our sort of ecosystem competitors, with very large models, the Flash model is sort of pulling its weight. How should people think about Flash 3.5 versus the competition? Yeah, I think it's probably like a more like, sonnet level model. I think OpenAI with just mainline, GBT, and then GBT mini, sort of they don't, I feel like Sonnet is not like a small model, if you will. It's obviously it packs a bunch, and I think it's definitely smarter than the mini models. So I feel like it's, I think we're anchoring more on the Sonnet level intelligence. So if folks are using that model and want to try Flash, please let us know, and send us the feedback of how it stacks up. Yeah, and the reasoning, all the agentic tool used stuff is all incredible. So I think that was one of the launches that models available, actually for the first time, like, available to all the users and search available across 900 million users in the Gemini app, available to developers in the API, so many other places. So I think it's like the most widely distributed model launch on day one that we've ever done, which has a film set of challenges that we could talk in depth about. The other thread, which I think folks are very excited about, is Gemini Omni. So sort of this new model that we've created, actually sort of somewhat of a world model, I think is how Demis framed it when we announced it on stage yesterday, being able to take in any type of input and create any type of output. And I think to give context for folks like Google, you know, we had Vio and Vio was state of the art, and sort of pushed the frontier for video generation. We had Nado Banana, which could do sort of image generation and editing. We have all these audio models that do TTS. We have a new Leary, a music model that's actually really, really capable. And the idea is, how do you fuse all of those models into a single thing, so that a, developers' lives are easier, but we don't need to train nine different models. And actually get this really interesting cross-pollination of capabilities, so that the same model that can actually benefit from Gemini's world understanding and the ability to generate text can also make a video, can also edit a video. And you see these like really interesting things, and we saw this with Nado Banana, like what happens when you give world knowledge to an image generation and editing model is really interesting use cases, and we give feedback all the time about what that's unlocked for customers. And so I'm really excited to see Omni starting out in the Gemini app in YouTube, and in Flow, and then very soon already we're kicking off a bunch of early access tests, hopefully as soon as I get out of I/O and get back to the office, and get feedback from developers and sort of continue the iteration and bring it to developers in the API so that folks can actually build products on top of Omni. Yeah, I mean, it's products, building products, but also building content that can be seen. Yeah, for sure. One of the hardest parts about vibe coding or building any business is getting distribution to the thing. Yep. Right? So when I was watching the Omni demo, and just seeing how amazing it actually is, like what was going through my mind is like, okay, how can I make it commercial? How can I make an ad? How could I create an Instagram account using this model? And I think what you're going to see is a lot of people generate millions of followers, generate millions of views every single month, if they know how to storytell well and use the model. I actually also think, and what I would love to see, I don't know, we'll give some work to your editor team. Make the intro of this video have a bunch of crazy stuff happening, and you can actually change the intro of the video. We did this for some of the podcasts that I was doing, actually with the Omni team. And that hopefully it'll create a bunch of new creators. I also think it is a fundamental accelerator for existing folks who are producing content. Like editing video is hard. Yeah, I'm sure you have an amazing team that's doing really hard work. And there's not enough hours in the day. There's not enough editor time in hours and storage on SSDs in order to do all the stuff that I think could be done. I think Omni will sort of fill this really interesting gap for maybe the creator who wants to tell an interesting story that didn't have the means to go have a whole team support them in telling that story. And I'm really excited to see that happen. And I've already been seeing a bunch of examples on X and other places that folks like starting to be able to tell that story and bring content to life in new ways and actually like repurpose existing content. So it's going to be crazy. And this is only the first iteration of the Omni model. This is the flash variance. So it only gets better from here, which is really exciting. I think we're going to see some really cool stuff throughout the rest of this year. And from an API perspective, how do you see, you think people are going to integrate with Omni in terms of the products they build? Yeah, I think there's so many interesting creative suites. I think a lot of people will also just be doing what you describe, which is like doing content with it and using the Gemini App or Flow or YouTube in order to do so. But I really do think it's an example of like this video remixing, editing capability is like not something people have built into products thus far. So it is going to open up a new category. And I think if you're building businesses and like want to help, you know, actually if you want to build a business and help the next thousand, hundred thousand million creators go and tell their story. And there's a bunch of really unique opportunities because this is a fundamentally different way of interacting with video. Totally. Yeah, I mean, it reminds me like I remember when social media was coming up and social media agencies were just popping up, right? Yeah. There's sort of a similar opportunity right now where, you know, you can build an Omni agency. Yeah. And sort of deploy this first small businesses as an opportunity, right? Yeah. And there's so many ways to like make the my takeaway from seeing this like on my own content was it makes the content more engaging. Like you could like do all types of like goofy things that like don't actually dilute, you know, people are like the tongue and cheek thing is like the subway surfer. You know, engagement farming and videos because people don't have attention. And actually like I'm as a consumer of content, I'm actually really hopeful that there'll be more people doing like the much more thoughtful version of that in a way that like you could actually only do with Omni or like a really, really advanced like the FAC studio that was able to pull this stuff off. So I think we'll actually see that and I think it will like. Move us away from this like slop sort of era into something that's like a little bit more tasteful. And I think that that's maybe corollary to I think some of the narrative of like AI tool sort of just generating more bad stuff in the world. Omni actually might be a really cool tool to help people tell a story with a higher degree of taste than they've been able to. What wasn't launched today that you wish was launched today. I know it's a question that you probably don't like getting but one thing I like about you is you're not afraid of criticisms. And I see you when I see every see criticisms of you know any of the products you're the first person in the replies to be like we're working on it we're taking feedback. So any anything that you wish you wish to see over the next coming months. Yeah well first of all I do I do love the feedback I think actually I think folks look at sort of criticism in such an interesting way I see it as like it actually makes my job easy. Yeah like if you come say you know hey we wish the models were able to do this thing or we wish the API was able to do this thing very easy to like go and take that feedback the hard thing is actually trying to identify these things when people don't tell you what they actually want. So I love the feedback I appreciate it please please keep it coming. I think what I would love to see and what what the teams actively working on is the other set of Gemini models. So we obviously we're launching 3.5 flash 3.5 pro is in is in the works it's cooking there's there's many iterations and runs happening behind the scenes right now. I think I think we said we announced yesterday it'll be available early next month or hopefully sometime next month maybe not early next month. And so yeah I'm excited it would be it would have been great to sort of bring the whole 3.5 model family out at the same time but it is also fun because I think 3.5 flash sets the bar high and I think we we continued so I describe it as like pulling the rabbit out of the hat we can.
continue to take the pro level intelligence and then stick it in the flash model. And I was talking to Oriole and Jeff yesterday who are some of the leads for Gemini who actually invented distillation, which is the technique that allows us to take the pro level intelligence model and stick it into the flash model sort of every single time. And it's just crazy that it keeps working. And it brings the cost of intelligence down, which is really interesting. Even though I, I, I, lots of feedback about like the cost of flash so far. And the important thing actually is that the cost of intelligence goes down. The model is like one implementation of the like skew capturing the cost of intelligence. And I think folks, it's sort of a, it's a lossy way of expressing the cost of intelligence. But I think net net cost of intelligence has gone down with flash, which is really exciting. - What else was launched that, you know, a founder or someone who's trying to be more productive or make money on the internet to know about? - Yeah, I think 20, I mean, obviously the open-call revolution that sort of like took the world by storm is exciting and is an opportunity. And actually, you know, I love the folks who are working on open-clots and incredible open-source project, Peter's awesome. But actually for folks, if you've tried that product experience before, it's really tough. Like you really do need like, you need to be a confident person who's like willing to take risk and let things, but I think Gary Tans sort of described this as like, you know, it's a Ferrari, but you have to also then be your own personal Ferrari mechanic and sort of make the tools work. And so, I mean, I studied computer science at school and like, same, you know. And I didn't feel great, you know, yeah, I love what the team is done and stuff like that, but I was kind of nervous. Like my heart rate was going up when I was like installing and doing things. - Yeah, yeah, it's, and you have to like take the leap of faith. And so, I think for people building stuff like that, that is in itself an opportunity. And so, I think there's sort of two angles to this. One, inside the Gemini app, we sort of launched our sort of flavor of this like, always on 24/7 agent to sort of help you run your business or, you know, come up with your next idea and sort of, it's rolling out to trusted testers this week. And then next week, it'll start rolling out to Gemini Ultra customers, which is exciting. So I think, you know, for sending work, throwing work over the fence and having an agent do it is the best thing. And if you want to build experiences like that, we just launched managed agents in the Gemini API. So using the same harness that's actually power in the same model that's powering Gemini Spark in the Gemini app, we also have that ability for developers to go and build those experiences themselves. And I'm really excited about this, like the friction to build agents historically and choosing a framework and all of this sort of like, iteration loop that you have to do in order to like, get quality good enough, et cetera, et cetera. Trying to shortcut that for people who want to build and just like not have to deal with the infrastructure, not have to deal with any of the problems. Send a single API call, I demoed this on stage yesterday doing like an AI radio show. And I think we're calling like seven different models and doing all of this. I didn't write any orchestration code. I literally just like wrote skills and markdown and was able to have it sort of go and orchestrate this whole show. So I think it also like, even if you're not the most technical developer and you didn't study computer science and you're just using these tools off the shelf, like managed agents in the Gemini API, I actually think hopefully we'll lower the barrier to entry for people who want to build agents. - I think MCPs are coming soon to that product. - I think we will have MCP support for it to do tool calling. I think right now it is just like skills in order to start. So you can like describe and it can do some of these things. But yeah, there's a whole, Miss IO is sort of like step one of the managed agent story. And there's like 50 other things that are on the roadmap that we need to land in order to make the experience rich and feature full and all that's coming. - I mean, that was my one big takeaway from being IO is, you know, I always thought of this time as the AI era and you guys have been saying it's the agentic era. - Yes. - And a lot of the new products that I've been seeing that you guys have launched of last 24 hours has been agentic this, agentic that. Can you just, yeah, I'm curious your thoughts on this agentic era and, you know, the person who's listening to this is an idea of person, right? So they're listening to this and I'm like, okay, how can I use this in my business or create a new business? Do you have any requests for startups or any ideas? Given that we're now in the agentic era that we can be using Google products to go build? - Yeah, no, it's a great question. I think two things come to mind. I think historically there was this like one-to-one correlation of you spending your time and actually an actual work happening. And I think the exciting thing for somebody who has ideas and wants to build stuff is like asynchronous agents, agents running in the background, fundamentally changes that dichotomy of like, you don't, doesn't require you actively in the driver's seat every moment that there's actually useful work happening. And the important thing is like, you still work. I think we're all sort of, I was having a separate conversation about sort of every three to six months this like expectation reset that you have to do with somebody building in the AI era. I think a lot of people like tried, I'm trying to think of auto-GBT as one example. Like three years ago in that team, I think, at the time like did something really interesting, but like it didn't really work to do anything useful. It was like, it was a really, really cool and interesting demo. And so I think a lot of us like myself included, I was like, okay, agents are exciting. It's the future, the future was not there at that time. And it feels like we've really crossed the chasm and that future is right now. So if you've sort of written off that there's a bunch of opportunities, reset your priors, there are opportunities and folks should be actually building those products because the customer base, this is the interesting thing is like that customer base yet does not yet actually know that they need an agentic product. And so I think there's some alpha and like how you storytelling this. Like it's not clear that like the average person, some of that maybe are looking for an agent. I think there's a lot of people who are like, they have a problem and they just want that problem to be solved. And you can solve it in a way that like really brings time back into people's day instead of like requiring them to be actively in the driver's seat the whole time. So and I'll add one more comment, which is doing that in the form factors that people are already familiar with still feels like there's so much alpha in this. Like trying to convince the small business owner to go and adopt some completely new thing that they've never heard of before that's really good. They're gonna have to teach all their employees about and their family about. It's a high bar to pull that off. Teaching the small business owner how to text with an AI assistant or how to send an email. You don't have to do that. They already know how to do that. And so I think you can really use this like existing technology to meet people where they are at the same time that you like fundamentally introduce new technology to them. - Anti-gravity got a huge overhaul. - Yes. - We should talk about that. - Yeah, yeah. So I think the whole anti-gravity suite actually is coming together. I think we introduced anti-gravity I think like six months ago something like that sort of as an AI powered IDE. And I think if you look at where it is now, it's actually an entire ecosystem. So you have anti-gravity sort of the agent manager. If you don't even wanna like touch the code itself like personally and you want the agent to do all that for you, there's the anti-gravity agent manager on web and desktop that you can install. There's the IDE, the same anti-gravity product that you're using before continues to work as an IDE. You can do that. There's anti-gravity the CLI product. So if you're a developer and you're like, I love the CLI, it's the best thing in the world. It got that for you. If you want the SDK so that you can actually build agents on your own infrastructure, the SDK exists. If you want anti-gravity sort of powering experiences in the API because you don't wanna manage the infrastructure, we have that in the Gemini API. And so I think the story is like anti-gravity sort of as this agentic coding layer, meeting you wherever you are, however you wanna build products. Actually, anti-gravity sort of the agent harness powering the always on Gemini Spark experience for consumers in the Gemini app. So everywhere you go, actually even going to search. So everywhere you go, you sort of access anti-gravity. Now it's another layer sort of bringing the Google ecosystem together just like Gemini is as well, which is really exciting. So how should people think about Google AI studio and anti-gravity? Yeah, that's it. You're making my job easy. That's a question I have, honestly. It's like, I can tell you how I see it. Please. To me, it's like AI studio just feels way more comfortable by non-technical. But an anti-gravity just feels like more of a fully featured developer-focused product. But that's also changing a little bit. So yeah, I'm curious how you see it. Yeah, I think it's changing in both directions, which is interesting. I think AI studio does a lot of different things. We have sort of the playground experience if you want to test our latest models and agents. You can get your API key to build with Gemini API if you're a human or an agent, which is exciting. And we're now also, we started building this vibe coding experience probably, I think actually 12 months ago, which is crazy. And to see the progress of where we are today, you can now natively build Android apps and sort of share them with users and download them onto your phone. and you can natively integrate with Google works.
without ever leaving AI Studio, which is also really exciting. The way that I frame this, and actually, Andre Carpathy has the best framing of this, which is it's actually two different things. There's on one end of the spectrum is vibe coding. On the other end of the spectrum is a gentick engineering. AI Studio is going after vibe coding. We want to make it so that you can bring your idea to life, go from prompt to profitable company, without ever actually having to see any code. You could never look at a single line of code, and we should help you do that. You should be able to deploy, get a database, in the future add payments, get a mobile app, do all those things, never look at a single line of code. On the integrity side, I think the problem they're trying to solve is very much production quality code. You want to work in a million line code base. You want to build Google. We're using integrity to build Google. The bar is quite high. We are using AI Studio in certain ways to build Google, but not contributing to the Google search code base as an example. There's flexibility control trade-off versus batteries included. I think, I think, you can use other models. You can use, you could be using the open AI API if you want as part of the anti-gravity. AI Studio is really focused on the Google ecosystem. We're trying to make a bunch of opinionated decisions for you, make it so that you don't have to think about those things. It just all works. You can sign in with your Google account, it works out of the box. You can build an Android app natively for free, install it on your phone, which I just found out yesterday. It's crazy. It's crazy. I mean, we just launched it yesterday. It's a brand new feature that I'm excited about. And Page and I were talking offstage before yesterday with a bunch of folks. Neither of us had ever actually built an Android app before. It's opened up this whole new ecosystem that we probably would have. I mean, I love Android and Android. It's awesome. I was probably never going to build an Android app. And so to see, you know, literally in the first day, like tens of thousands of people go and build their first Android app in AI Studio, literally for free, I think is really cool. I mean, so much alpha and just that alone. Right? You can think about it. Android is the number one operating system. There's like billions of people using Android. Billions of people. And you don't need to be technical. And you can ship your Android app. And, you know, you need a good idea. You need a good niche. You have to understand distribution. But there's no reason that if you want to build a mobile experience, so you're not just trying it out. Yeah. And one of the intentional decisions on Android that we actually made was to make it so that these are native Android apps. So this isn't like a, you're not using one of the frameworks that sort of lets you do like all of the different ecosystems with a single app. It's intentionally a native Android app. And the reason for that is because we want to make sure you can actually build for all the different form factors. So you can actually, and this doesn't work perfectly well today. But we have this sort of groundwork in place and where this is where we're incrementally getting towards. But you can build for our Android XR wearables. You can build, so like glasses in the future. And the fall when the glasses land, you're going to be able to build in a studio, build a native app for XR and build XR apps without ever actually writing any code. You're going to be able to do the same thing for the other wearables, like the watches. Android Auto. You can build like apps for people in the car when they're on the go. Like all of those things, using the same sort of setup in the same agent in a studio is really exciting. And those form factors, you actually don't get using any of those other like compatibility frameworks. And so I think it'll be cool to sort of build these frontier type applications in a studio with native Android apps. For people who listen to this, well, end with this. For people who listen to this, what are your hopes with how, you know, there's hundreds of thousands of developers who and builders who are listening to this? What are your hopes for these people who are watching now to go and build? Yeah, that's a great question. I think the most exciting thing is like what YouTube did for creators. I think is what's happening now for software. I think there's going to be this entire generation of like software-esque creators where you can just like go by yourself or with a small team, like go and build a business using all of these AI coding tools. And I'm very excited to see similar to what YouTube has done where like you don't need this like historically to build software. You need to like either get really lucky or raise a bunch of money from people in order to raise a bunch of money from people. The market had to be massive. And so there was all these like really, really interesting problems that people have just decided they're not going to solve because the market's not big. And I think those markets are now in play. And I think it's just like it's alpha for the the observant person who's willing to do the diligence to like go try to find some of those opportunities. And I think why I love the content that you do is because I think you're actually helping people find those opportunities. And like that is the next generation of like value creation that's going to happen is people going after these things that you wouldn't have obviously gone after five years ago when the cost to build software was so high. And it you needed a 40 person team to build something. Now it's like one person or a small group of friends and some of these tools and anything is possible. You're off to the races. I love it. All right. Thanks a lot Logan. This was great. Thank you for. Appreciate you. Thank you for having me and first I owe first I owe. Hopefully hopefully we'll have you for the next year as well. And thanks for thanks for blazing the heat of Northern California. I appreciate you. Yeah. Take care.
Podcast Summary
Key Points:
Google I/O 2024 focused on the "agentic era," with major releases centered on building AI agent-native products.
Key model launches include Gemini 2.5 Flash (a powerful, cost-effective workhorse for agentic and coding tasks, comparable to Sonnet-level intelligence) and Gemini Omni (a multimodal model that fuses text, image, video, and audio generation/editing capabilities).
Omni is designed to simplify content creation by allowing video remixing and editing, potentially lowering the barrier for creators and enabling new forms of engaging, tasteful content.
New developer tools include "managed agents" in the Gemini API, which reduce the friction of building agents by allowing a single API call to orchestrate multiple models without writing complex orchestration code.
The "agentic era" is defined by asynchronous agents running in the background, fundamentally changing the relationship between user effort and work output. This presents a major opportunity for startups to build agentic products that solve real problems in familiar form factors (e.g., text, email).
The entire "anti-gravity" (likely a codename for a product suite) ecosystem has been overhauled, offering an agent manager, IDE, CLI, SDK, and API access, meeting developers wherever they prefer to work.
Summary:
5 Flash, a powerful and cost-effective model optimized for agentic tasks and coding, and Gemini Omni, a unified multimodal model that can generate and edit text, images, video, and audio. Omni is seen as a tool to empower creators by simplifying video remixing and storytelling, potentially moving content away from "slop" toward more thoughtful, engaging formats. For developers, Google introduced "managed agents" in the Gemini API, which drastically lowers the barrier to building agentic applications by allowing a single API call to orchestrate multiple models.
The conversation highlights that the agentic era is defined by asynchronous agents that work in the background, freeing users from needing to be actively in control. This shift creates a significant opportunity for startups to build agentic products that solve real problems within familiar interfaces like text or email, rather than requiring users to adopt entirely new workflows. The "anti-gravity" ecosystem was also overhauled, providing a comprehensive suite of tools—including an agent manager, IDE, CLI, SDK, and API—to support agentic development across different user preferences.
Overall, the message is that the future of building and creating is now agentic, with lower barriers to entry and powerful new capabilities.
FAQs
It is Google's best model, designed for agentic, long-running tasks and coding. It offers Sonnet-level intelligence at a low cost, making it suitable for developers and builders.
Gemini Omni is a multimodal model that can take in any input and produce any output, like text, video, or audio. It enables video remixing and content creation, with early access in the Gemini app, YouTube, and Flow.
The agentic era focuses on autonomous agents that run in the background, allowing work to happen without constant user input. This creates opportunities for startups to build products that solve problems asynchronously, meeting users where they are.
Managed agents allow developers to build agentic experiences with a single API call, using the same model powering Gemini Spark. It lowers the barrier to entry by handling infrastructure and orchestration.
Omni helps creators edit, remix, and repurpose video content easily, reducing the need for large teams. It enables storytelling with higher quality, potentially moving away from 'slop' content.
Anti-gravity is an agentic coding layer that includes an IDE, CLI, SDK, and agent manager. It powers Gemini Spark and is available in the Gemini API, meeting developers wherever they work.
Chat with AI
Loading...
Pro features
Go deeper with this episode
Unlock creator-grade tools that turn any transcript into show notes and subtitle files.