Go back

Ep 841: ChatGPT Computer History, New Gemini Model, Claude Flexes on the Browser and 7 more AI updates you should use Today

36m 52s

Ep 841: ChatGPT Computer History, New Gemini Model, Claude Flexes on the Browser and 7 more AI updates you should use Today

The podcast episode, hosted by Jordan Wilson, kicks off by announcing a 30-day replay of the "Start Here" series, which aims to educate listeners on generative AI fundamentals, drawing from over 10,000 hours of coverage. The main segment focuses on recent AI developments, emphasizing how faster and cheaper models are making AI more practical for business use. Key highlights include Google's Gemini 3.7 Flash, released just three weeks after its predecessor, offering improved coding and agent capabilities at half the API cost, making it a top recommendation for cost-effective performance. Zai's GLM 5.3, an open-weights model, demonstrates significant advances in cybersecurity, outperforming closed-source rivals on the CyberGin benchmark, though its weights are pending release. SpaceX AI's Rock 4.6 also emerges as a strong competitor, matching GPT-5.6 Seoul on independent benchmarks and available through Cursor, with no price hike. Additionally, Google Sheets introduces a Canvas feature that transforms plain spreadsheets into interactive dashboards, accessible to paid Gemini users, streamlining data visualization without external tools. The episode underscores a trend of rapid model iterations and feature expansions, urging listeners to leverage these tools for business growth, while noting that many updates, like Microsoft Copilot unification and ChatGPT voice improvements, didn't make the top seven list due to the sheer volume of changes.

Transcription

6410 Words, 34323 Characters

English
Welcome to the Everyday AI podcast. My name is Jordan Wilson and for the past three and a half years, we've put out more than 800 episodes. Yet, one of the most common questions I get, I didn't really have an answer for. Where do I start on the Everyday AI podcast? And that's why we started the Start Here series. And with the fall now back in full swing, the Everyday AI podcast is going back to school and playing back the entire Start Here series from front to back. We've hit pause on our normal Monday to Friday programming to run back our most popular series ever for the next 30 days. We made the Start Here series for beginners and AI champions alike. So whether you're just trying to get a grasp on large language models or grappling with the best coding harness for multi-agentic workflows, the Start Here series covers it all. Any language, no jargon, and easy to follow along each day. So make sure to subscribe to the podcast and check back each day for new insights day by day. The series is a culmination of spending more than 10,000 hours covering generative AI over the past three and a half years. So you don't want to miss a single episode of the Start Here series. Let's get into it. AI just got faster, more accessible and cheaper all in the same week, but there's a good chance you might have missed all of it. I mean, let's see what happened just recently. Google dropped its new Gemini 3.7 flash at half the price of the version that they just released three weeks ago. Then Elon Musk's team shipped a smarter rock at the exact same price and open AI answered with a new mode that runs its most powerful model 14 times faster. I mean, I still don't even know how that's even possible. But here's why you should care even if you've never touched an API because price and speed are the two things standing between AI demos and AI that actually is running inside of your business because every time one of those two things drops, a use case that didn't make sense last month suddenly makes a ton of sense. And that was just a model news. I mean, Claude started working inside your browser tabs, your spreadsheets can now turn themselves into live dashboards and chat to be tea quietly shipped a feature that watches everything you do on your Mac. And I'm here for it. So that's why on Fridays, we bring you Friday features are weekly updates of the latest AI features that you can actually use today, not announcements, teasers. Yeah, you can go use all these things right now. All right, let's get into it. Welcome to every day AI. If you're new here, my name's Jordan Wilson and we do this every day. So daily unedited unscripted live stream podcasts and free daily news that are helping business leaders like you and me not just keep up with everything that's happening, but I tell you what matters, what doesn't, how to use it and you take that information to grow your company and your career. All of a sudden it's like, whoa, you're the smartest person in AI. Everyone's looking at you like, what's your secret? Well, your secret is this, but also our website, your everyday AI dot com. So if you didn't know, you can go check that out. We have literally 830 plus episodes. You can go watch, listen to and read about all for free. It is a free generative AI university on our website. So let's get into it on today's show. Here's what you're going to learn. You're going to know the new Mac feature that watches everything you do at work, then tells you what you did all day and maybe finishes the work that you forgot to do. You're going to know why the smartest cheap model in AI just got 50% cheaper and the catch that's hiding in the fine prints. And you're going to know how quad can now do real work inside your browser tabs, login and all. All right. So let's take a look. Actually, livestream, it's having shot as you out in a bit. Big bogey face. Good to see you. Brent joining us from London, Brian from Minnesota, Jose joining us from Santiago, Miko joining us from Tokyo, where we're all over the place this morning. I'd love to see it. Michelle, good to see you. Join us. All right. A lot of new AI features, y'all. So on Fridays, it's one of my favorite shows and it's actually been one of our more listened to shows because it's hard to keep up. And some of these things literally I had the show planned last night woke up this morning. I'm like, all right, there's already new things, you know, some things had to get the cut. Right. It's kind of crazy because there's some big updates this week that didn't even make our top seven. I mean, let me just read these before we get into the ones we're actually going to cover. I mean, the Microsoft co-pilot started unifying its co-pilot offerings into the super app. So that's actually happening. Was named in the top seven thing. Chat and GPT upgraded its voice mode to include projects and file uploads, which is pretty big. I've been using that one a lot since it came out. Google Gemini released 13 new connectors and Google also released a slew of new AI features for Google ads and Google analytics. So let's get into the ones that we're actually going to be covering today. And let's start with, well, if you care specifically about the fastest, cheapest model that is still the most intelligent, technically, if you look at things like artificial analysis, they actually have a nice little model filter. And if you say those are the three things that are most important for you, the new Gemini 3.7 flash is the recommended model. So it is brand new. And if you're wondering, like, wait, Jordan, did she just cover this new Gemini flash? Like very recently? Yeah, three weeks ago. That's how long it took Google to go from Gemini 3.6 flash to Gemini 3.7 flash. So let's see what's new. Google released Gemini 3.7 flash. It's new AI model that it calls the most intelligent workhorse model for coding in agents. And it's a big step up from the version released just three weeks ago with major gains in coding, debugging, and building web apps. So who has access to this? If you're a paid subscriber, it's available already. So interestingly enough, it's not yet available inside of Google Gemini, right? So Gemini.google.com. But it is available in its Gemini Spark model, which is available, the kind of agentic offering for Google, which is available in Gemini. I know that's a little confusing. But if you go into Google's AI studio, it's in there as well as Android studio and the Gemini API and Gemini Enterprise platform. So yeah, if you're a normal paid Gemini user, which is, I think the majority of our listeners and you go in there, you're not going to find it there. But it is powering the Gemini Spark and if you're using Google anti-gravity. So many different ways that you could access Google's AI. It's like you need a roadmap sometimes. All right, but here's the key thing. API pricing is 75 cents per million in put tokens and $3.75 per million output tokens. All right, so let's read just quick kind of what Google said about their newest release. Let me zoom in on my browser here. Apparently, I'm old and can barely see. All right, so Google says today we're building on the progress of our widely used flash series by introducing Gemini 3.7 Flash. Our most intelligent workhorse model yet for coding and agents. The release comes just three weeks after Gemini 3.6 Flash and is a direct result of developer feedback and algorithm, algorithmic innovations that we look forward to bringing to future models. 3.7 Flash delivers substantial improvements across software engineering, knowledge work, and web development workflows with an introductory price of half of the original 3.6 Flash cost per million token. So if you recall when Gemini 3.6 Flash came out, I said, I don't understand this model because the whole point of the Flash series all along from Google has been to be extremely fast and extremely cheap. And Gemini 3.6 Flash, although it was fast-ish, it was not cheap, right? It was not its normal. You know, you would look at Gemini's Flash series and be like, all right, this is, you know, not like free. But this is a very low-cost model that you can put out at scale. So unfortunately, I think a lot of people were upset at Gemini 3.6 Flash because it wasn't really that Gemini 3.7 Flash gets back to exactly that. So who's going to find this useful? Well, it's useful because it got smarter, faster, and the price cut in half. Again, that's if you're paying on the API side. But the real thing is inside Google Gemini Spark. So if you haven't used Spark, that's kind of their new ish autonomous coding agent that is inside of Google Gemini. So it's kind of a mode. It's an agent mode inside of Google Gemini. So for that, it's going to make, I think, a big difference. So I tested out Gemini Spark. I wasn't super impressed, you know, when it first came out to paid plans, mainly just because I've been using, you know, codex and, you know, cloud code, cloud code work for so long. And I'm like, yeah, I'm not really saying anything that I can't do inside of those. But now with a better model, I think some people are going to find it much more valuable. So who's going to find this valuable? Well, if you're a business using AI on high via the API, this is might be your new workhorse model, right? People don't really understand from a serving tokens perspective. Google is and has always been a leader in the space, right? Even though even if you don't see a lot of people talking about the Gemini chatbot, but it comes to serving up tokens, right? Google has been a leader and will continue to. All right. So oh, Stephen Leitz says he's been waiting for this one. All right. you enjoy Gemini 3.7 Flash. All right, let's get into our next one. And this is the one that just came out like kind of like an hour, couple hours ago. So yeah, we may have a new king of the open source hill. Yeah, we don't even have full benchmarks for it. So we don't know if it's going to truly be the best open weights model in the world, but it might be. That's because we have a new version of Zai's GLM 5.3. So a new model, it's calling, it's best yet for coding. And the strongest open weights coding models that it has measured. So an interesting wrinkle, it's the same underlying model as GLM 5.2, but every single gain just comes from better post training. So yeah, not like a completely new model per se. This is more of a step change that's very, very normal. But the big surprise here in the big jump, at least according to Zai's benchmarks, like I said, it's too fresh to have any third party benchmarks. Yet big jump is cyber security. So Zai says the model got better at finding and chaining software vulnerabilities than the company intended. And it is already found over 1,000 critical security flaws in real software, like Linux and Apple's web kits. Yeah, ones that have already, you know, according to Zai, withstood, you know, the fable in the GPD-5.6 vulnerabilities that everyone has found. The thing that's interesting here, you know, so for our live stream audience, I have kind of their blog post announcement up on my screen. You can always watch the video version of our podcast on your everydayai.com, as well as you can go catch that link in the show notes. But the interesting thing was, yeah, some of these benches, right, I mean, they're definitely, they're definitely competing with the closed source proprietary model, specifically, you know, Fable 5 from Anthropic and GPD-5.6 Seoul. So the one that really stood out is CyberGin, all right. So this is kind of a metric on the cyber security side, which is very important. I talked about that, you know, this week, was that this week that I had my, yeah, my rogue AI agent show, it's all blending together now. But, you know, let's go back to when Anthropic teased its mythos and fable models and then they sat on it and released it for like three months. The main, the main benchmark that it led with was CyberGin, all right. And right now, you know, it is a pretty important benchmark, you know, CyberGin and exploit gym. But for CyberGin, this new open model is the best in the world, right, which is crazy, right. Yeah, an open, an open model that is the best. So it scored an 84.5 on CyberGin ahead of the other two models that were kind of tied there for first place previously, which is Fable 5 and GPD-5.6 Seoul. So a little bit more about GLM 5.3 and who has access, as well as available to everyone now. So if you are on Zai's GLM coding plan or if you have their Z code installed, which Z code is pretty much just a pixel per pixel copy of codex. Anyways, well, if you use codex and you're like, oh, I want to try this new GLM 5.2, you can try Z code and it's pretty much the same thing as codex. But it's been rolled out to all subscribers. So the open weights are not out yet, all right. So Zai says it will release them in about two weeks after further safety testing. So calling it open source today is a little premature. It will be open weights and we at least know that GLM, sorry, Zai has a positive track record of doing this. This is what happened with GLM 5.2. This is what happened with Kimi K3, with Quinn, 3.8, all the Chinese open models they've released and then they usually will release the weights in about two weeks. So technically the weights are not available today. If you're superdork and care about those things, but they will be available soon. So why is this useful? Well, it's been a budget favorite. So the GLM coding plan has been a budget favorite for AI coding and this makes the budget option meaningfully better at complex long running tasks. So the company said on its own tests, it's roughly 50% better at coding than the last version, which is crazy. And if you recall back a couple of months when GLM 5.2 was released, this was kind of the first open model that people started to look at and say, wait, has China kind of closed the gap on US frontier proprietary close models and it was like kind of, right? I think before it was always US models were tier one and Chinese open models were tier two. So the US models are still ahead, but it's now it's 1A and 1B with the Chinese open models. So GLM, Kimi K3, and Quinn, 3.8, are probably in that 1B tier with, you know, in Propix, Opus 5 and Fabel 5 along with OpenAI's GPT 5.6 sole are kind of still in that 1A tier by themselves. So pretty interesting now. Yeah, Josh here on YouTube, good comment said, actually OpenAI now has new cybersecurity daybreak, blue and red. That is correct. They rolled that out earlier this week. If you read our newsletter, you caught that as well. All right, more new models. Do you see a pattern here? A lot of new models this week. They're coming out faster and faster than ever. In this one, I was kind of surprised with Rock 4.6 is actually a pretty good model, right? It's probably joined that 1B tier, you know, all of a sudden it is a top contender. So let's talk about what's new in the model. So space X AI, yes, formerly X AI, they just jammed all the names together, but the real name is space X AI released Rock 4.6. It's newest top tier AI model built for coding in long running automated tasks. So on a widely watched independent scorecard, that is in a lot of the benchmarks that artificial analysis covers. And now ties OpenAI's GPT 5.6 sole in a lot of important benchmarks and jumped five points over the previous Rock. So yeah, on artificial analysis index, it is right there in that connect with GPT 5.6 sole. So who has access to this right now? Well, developers have it available through the X AI API. And it's also included in the coding tool cursor on all plans. So yeah, cursor obviously space X AI acquired cursor. So this is available if you are on a cursor plan as well. It's also now the default model in Rock build, which is X AI's attempt at building a codex or clawed code competitor. So why is it useful? Well, the price didn't change from the last version of Rock. So anyone using Rock got a smarter model for the same money if you are paying per tokens. But the biggest thing is this keeps the price war alive because the jump from Rock 4.5 to 4.6 was huge, right? Maybe some of that compute that the company has been investing in and the cursor acquisition is starting to pay off in a model that maybe now companies will start to look at. I'm kind of surprised maybe that space X AI isn't just going with the cursor brand name and going with composer 3 or something like that. I will say this and this isn't my own personal bias. This is just literally anyone on the internet. I think because of some of the areas that Rock has made a niche in, people from a business perspective when it comes to paying tokens, they're not always looking at Rock as a competitor. They're like, oh, this is the one that's been doing some wild things on the internet. So we'll see if now Rock 4.6 kind of teetering between that 1B and 1A to your changes that. But here is what the company says and said in its release. It said today we are releasing Rock 4.6 on Rock 4.5 with a particular release, with a particular focus on long running agents and more ambitious interactive and visual work. It stays with complex tasks across many steps, whether researching a topic, analyzing information, working across a code base or turning an idea into a polished application or work artifact. The other thing that's kind of related to this Rock 4.6 and we won't be covering this on today's Friday features, even though I would have liked to, but it's only available on the very expensive Rock and Curse or plans, but the new Rockbot, which does look exciting and I'm looking forward to it being released on the standard base plan. So yeah, on the Friday show, we're usually not covering things that are only exclusively on the $200 or $300 plans, which is Rockbot for now, but Rockbot is powered by Rock 4.6. So pretty big week of releases from the SpaceX/Rockbot, Rock/Cursor Teams. All right, let's move on to our next one. And this one, I don't know, I'm a Google Sheets dork. So maybe I'm more excited about this one than the average person, but yeah, Canvas mode, one of my favorite modes of all time from Google Gemini is coming and it's here inside of Google Sheets. So let's talk about this feature you can go use right now. Google expanded access to its sheet canvas. feature that turns a plane spreadsheet into an interactive dashboard. So yeah, any Google sheet that you have inside of Google sheets now, you can literally turn it into an app, which is really cool. Right? So this is, you know, at least for me and for a lot of people, all this technically does is cut down on, you know, two or three little steps in between before I would just take my Google sheet, you know, download it, export it or connect it inside, you know, codex, cloud code or even Google Gemini, the chat bot side and then go ahead and, you know, build a little app or dashboard, but now you can do it all inside Google sheets without leaving. So yeah, it can just build a working mini app on top of your data that stays in sync and it updates as that data changes. Any paid subscribers for Google Gemini, you have access to it. It is literally just inside of the right panel. So just go click that, that Gemini button inside Google sheets in plain language. So they turn this into an app that does A, B and C and it's done for you. Also, now if you have a Gemini work or school account, it's also available. All right. Well, building a decent dashboard in sheets used to mean a lot, right? So if you're just looking at building, you know, graphs and charts, yeah, you have to kind of know the formulas or some formatting hacks or use a separate business intelligence tool. Now you just type, build me a sales pipeline board and build one. All right. So the spreadsheet sheet though stays as your source of truth. And here's the kind of the unlock, at least for me personally, right? Thing. Most people, if you work a lot in Google sheets, you know this, but Google sheets is notoriously good at connecting with all of your different data sources. So, you know, I don't know if you have, you know, a male, chimp list or if you have a, you know, ads going in meta, whatever it is, right? So many of these platforms sync directly with Google sheets. So when you can say if you set something up once and then you can build the app inside of Google sheets and then if the data is syncing live, you can build a nice little app once and never have to update it again. And you can go in and it's updated every single day. So this is kind of similar to a, I would say a light version of like chat GPD sites or the live artifacts feature inside of Claude code. All right. Speaking of Claude, all right, they took a page out of the codex book and essentially the very good chat GBT Chrome extension that we covered here a couple of weeks ago. Now the exact same thing is available but for Claude, essentially bringing the power of Claude code work into the Chrome side panel. So here's, let's actually start with what Anthropic says. It says the Claude in Chrome side panel is now a Claude code work session. Conversations are saved to your history, your skills and connectors work in the browser and a task you start in a tab can be finished on the Claude desktop web and mobile apps. It's available on max in team plans today and is rolling out to pro users over the coming weeks. All right. So here's kind of the, the gist. So essentially it's asked that you start in a Chrome tab can now be picked up and finished in the Claude desktop if you're using Claude code work. So it is a co work to co work sync as far as my tests show yet right still one of the big downsides of Claude code, Claude desktop in general is Claude. Co work has no idea what Claude code is doing. Claude code has no idea what Claude check on the desktop is doing in vice versus so there are three very different silos, which is unfortunate. So the, the Claude Chrome extension here that works with Claude. Co work is just with Claude co work. So the exact same functionality has been out for about a month now with the chat GPT sidebar and then that one is unified and it works whether you're using the chat panel or the work panel. But regardless, a big step up for Claude users. Here's why it's useful. Well, Claude can now see the page that you're on and acts on it and then obviously sync that session co work to co work. So everything from clicking, typing, filling forms and using your existing logins, that's the big unlock there. So that means it can work on the tools that don't already connect to a I be a direct connectors or mcp's like your vendor portals or just, you know, internal dashboards that are, you know, a little more antiquated and don't have, you know, all that, you know, fun updated AI connections. So now, you know, Claude co work just go around, click around and use computer use inside your browser via the updated Chrome extension. I mean, anyone if you're a heavy Claude user and you still have to use the internet and a lot, right? That's the big thing. I always say first, always check directly integrated connectors or integrations. Second, see if whatever app that you're using has an mcp or model contacts protocol that's supported in Claude, Gemini, chat, CVT, Microsoft, co-pilot, et cetera. And if it doesn't have one of those two things, that's where something like this new Claude in Chrome, Claude co work extension really comes in handy because for all of the other, you know, probably 90% of the internet that can't talk directly, you know, to your Claude co work, this new extension will at least allow it to enter to understand and use the interface. So yeah, I think anyone is going to like this if you had those, you know, annoying tools, you know, that you're like, oh, man, you know, this, this app that I use, that's a big part of my job. There's no AI integration. Well, now at least Claude co work in Chrome can use it. All right. We have two more and both of these are from open AI. So this one has been talked about for a while. It was supposed to be released in July. So in extra two weeks of waiting. And if you are a heavy API user, if your company is using the new GPD 56 soul via the API, this one could be a legit game changer. This is the first of its kind across any model, literally in the history of AI that you can run a frontier AI model at legit blazing speeds. All right. So let's talk about what's new. So open AI announced ultra fast. It's new speed tier that runs its most capable model, GPD 56 soul at up to 14 times faster than normal. So how the heck is this possible? Well, it is a specific partnership with the now public company, the chipmaker, Cerebrus. It generates up to 750 tokens per second with open AI and Cerebrus saying it's the exact same intelligence, just legit lightning speed faster. So who has access right now? So if you are in this is this one is a little more select, but I think for open AI API customers, it's a lot of our listeners. That's why I still included this one. So it is in limited preview, but this is y'all, especially for our larger enterprises. This has been one of the most looking, the most highly anticipated features that I've seen in a while. Right. People have been talking about this for many months. So essentially open AI did have a Cerebrus model before, but it was an older model. It was GPT 5.3 codex. So it was a codex specific model and it was an older model. I literally just use this like yesterday still. So it's still a really good model, but it's older, but it's so blazing fast. Right. I when it first came out, I was just using it for computer use because it could use a computer, right. Like as an example, not to knock on, you know, clawed, but you know, cloth's computer use as powerful as it is. It is painfully slow. Right. Codex's default computer use is in my testing five to seven times faster. So when you use a Cerebrus model with it, it is like as fast as a human or faster than a human can use a computer, which for me is an exciting unlock. So why is it useful? Well, until now, if you wanted real time speed, you had to settle for a smaller, not as capable model. So this kind of removes that trade off. It's top tier intelligence at speed that are fast enough for live products. So, you know, as an example, we talked earlier about the new Gemini 3.7 flash. So this is at speed, I think technically a little faster. We'll have to wait. This has only been out for, you know, less than 24 hours now. So we'll have to see the third party speeds. But presumably, this is going to be faster. So this is going to be delivering a level of intelligence via the API that literally the world has never seen. So I haven't even quite wrapped my brain yet around what is going to be made possible with this. But you know, speed just matters for business. So like as an example, if you are using this to power your agents, you know, a 40 step task as an example, it's going to pay that speed tax on every single step. So faster steps turn a workflow that maybe took minutes into one that takes seconds. So you can see results faster. You can make improvements faster and you can find a good internal use case fit faster. Whereas before it just might take longer. Any large enterprise company that's building customer facing AI products where people are sitting there waiting for an answer or sitting there waiting for an agent to complete its tasks. So teams, you know, teams that are running long, multi-step AI agents are going to find this extremely valuable as well. All right. Let's go to our last. last one here. This one, another kind of niche one, but is super, super exciting for me. All right. So if you remember, like a year and a half ago, Microsoft had its recall product that was kind of recall, right, not really, but technically, right? They released it, privacy advocates are like, you can't release this. And then it got pulled the back and forth, right? Now, I think you probably have a much better version of this in the new update that open AI just dropped called computer history. So we had talked about chronicle before. So this is just an extension or chronicle made better. Let me break it down in simple terms. So open AI just released computer history. This is a new feature that watches your activity across apps and websites on your Mac. It is Mac only right now. Sorry, Windows and Linux. And yes, open AI did this week release chat GPT work slash codex for Linux as well. But this essentially watches all your activity across apps and websites on your Mac and turns it into memories in a timeline that chat GPT and codex can use. So you can literally scroll through every single thing that you've done on the new chat GPT slash codex desktop app. And you can ask things like, what was I working on before my break? Or where's that proposal docked from this morning or summarize yesterday for standup, right? Anything that you work on inside of codex or chat GPT work, which for me is quite literally everything. It can notice repeated workflows, right? So it can find things that you started and didn't stop. And then it can finish it itself. That is wild. So it can not only finish work that you were doing and just forgot or, you know, maybe you ran into a roadblock, got busy whatever. So this new computer history can not only finish those tasks that you started and just stopped or abandoned, but it can actually suggest skills as you go along, which for me, I talked about this on the show the other day in my the next 12 months of AI. So make sure you go check that out. That was yesterday's show. So Thursday show went over the 19 predictions every business leader needs to hear. And one of the things I said was, you know, lines of code is not important metric. Tokens burned is not important metric. One of the most important metrics I think for the rest of 2026 and early 2027 is how many times you are reusing skills, right? Because you are saving not only a ton of time and sharing those across your team, but also you're saving tokens technically because you are only running the most optimized version of a workflow inside of your AI operating system that you know is going to bring a certain level of output or results. So now with this new computer history enabled, it is going to complete things for you, which my gosh, I need because normally I have like, you know, eight to 20 different, you know, codex threads going at once. And I forget these things a lot. But then also to suggest skills. This is really cool. So it essentially allows chat to be to observe and learn your activity on computer, helping it understand your workflow, complete ongoing tasks. You might have forgotten and then recommend those skills and automations tailored to how you use your device. So who has this right now? Any paid business plan has it right now as well as chat GPT pro. So you have to go into the chat GPT desktop app. It is Mac only and it is off by default. So you have to turn it on. And if you are using this across your team, every person has to turn it on themselves. And obviously this is not yet available in the EU UK or Switzerland. So all of those, you know, strict privacy first companies, you're not going to have it. Well, this is the AI that finally knows what you did yesterday, not just what you typed into a chat box. Right. So the context switching and like the where was I moments, I think are some of the biggest silent time wasteers right now in knowledge work. So a real example, you know, the end of the day, asking for a summary of everything you worked on or start the morning by asking chat GPT to pick up where you stopped or, you know, looking at the skills that it's adjusted based on all the work that you did yesterday. So if you're just juggling a ton of different projects like me, if you're losing time, if you're a heavy AI user, you're really going to like that. It is a little, you know, so this computer history builds on the chronicle research preview with reduced token usage and more privacy control. So like I said, this is fresh. I haven't even used it yet because literally I had so many codex threads running. I didn't want to, you know, download the update just yet. So, you know, Chronicle did take a little bit of space on your hard drive. So, but this specifically, the computer history is the next iteration of that. And they did say it is a little faster and reduced token usage. All right. So that is a wrap for the seven AI updates that you should use today. Y'all, if you're not paying attention every single day, that's why this Friday show is extremely useful. All right. And a reminder, y'all, the live stream and the newsletter are taken a little break. Quick announcement. It is time to go back to school. So we're going to be doing some learning together on every day AI. So our start here series that we kicked off this year has been some of our most popular shows ever. And we've gotten a lot of comments of people like, Hey, you should just run these all. So that's what we're going to do. We're going to be running back the entire series 30 episodes 30 days commit to this. So we're going to be running that series back starting on Monday, starting with the very first episode. So yeah, a little break from our normal Monday to Friday schedule, but I guarantee if you stick with this series from a volume one through volume 30 and put the work in, you'll know more about AI that like 99% of the people at your company. So it's made for beginners and advanced users alike, no charge and no fluff, no computer science degree required class starts Monday. Don't be late. All right. That's a wrap. Make sure if you haven't already go to your everyday AI.com. Thanks for tuning in. We'll see you back next time for more everyday AI. Thanks y'all. And that's a wrap for today's edition of everyday AI. Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rating. It helps keep us going. For a little more AI magic, visit your everyday AI.com and sign up to our daily newsletter so you don't get left behind. Go break some barriers and we'll see you next time.

Podcast Summary

Key Points:

  1. The Everyday AI podcast is replaying its "Start Here" series for 30 days, designed for beginners and AI enthusiasts, covering generative AI topics without jargon.
  2. Recent AI model releases include Google's Gemini 3.7 Flash, which is cheaper (half the price) and faster, focusing on coding and agents; it's available via API, AI Studio, and Gemini Spark.
  3. Zai's GLM 5.3, an open-weights model, shows major gains in coding and cybersecurity, scoring 84.5 on CyberGin, ahead of proprietary models like Fable 5 and GPT-5.6 Seoul; weights release in two weeks.
  4. SpaceX AI's Rock 4.6 ties OpenAI's GPT-5.6 Seoul on key benchmarks, is available via API and Cursor, and keeps the price war alive with no price increase.
  5. Google Sheets now offers a "Canvas" feature that turns spreadsheets into interactive dashboards or mini-apps, available to paid Gemini subscribers, simplifying data visualization.
  6. Other updates include Microsoft Copilot unifying into a super app, ChatGPT voice mode upgrades, and new Google Gemini connectors, though not all made the top list.

Summary:

The podcast episode, hosted by Jordan Wilson, kicks off by announcing a 30-day replay of the "Start Here" series, which aims to educate listeners on generative AI fundamentals, drawing from over 10,000 hours of coverage. The main segment focuses on recent AI developments, emphasizing how faster and cheaper models are making AI more practical for business use. 7 Flash, released just three weeks after its predecessor, offering improved coding and agent capabilities at half the API cost, making it a top recommendation for cost-effective performance.

3, an open-weights model, demonstrates significant advances in cybersecurity, outperforming closed-source rivals on the CyberGin benchmark, though its weights are pending release. 6 Seoul on independent benchmarks and available through Cursor, with no price hike. Additionally, Google Sheets introduces a Canvas feature that transforms plain spreadsheets into interactive dashboards, accessible to paid Gemini users, streamlining data visualization without external tools.

The episode underscores a trend of rapid model iterations and feature expansions, urging listeners to leverage these tools for business growth, while noting that many updates, like Microsoft Copilot unification and ChatGPT voice improvements, didn't make the top seven list due to the sheer volume of changes.

FAQs

The Start Here series is a beginner-friendly series from the Everyday AI podcast, designed for beginners and AI champions alike. It covers topics from large language models to multi-agentic workflows with no jargon.

The podcast is replaying the entire Start Here series over 30 days as a back-to-school initiative, pausing normal programming to feature its most popular series ever.

Gemini 3.7 Flash is Google's newest AI model, offering major gains in coding, debugging, and web app building. It's 50% cheaper than the previous version and is available to paid subscribers via platforms like AI Studio and the Gemini API.

GLM 5.3 is a new open-weights coding model that shows major improvements in cybersecurity, scoring 84.5 on CyberGin, ahead of proprietary models like Fable 5. It's available now, with weights expected in two weeks.

Rock 4.6 is a top-tier AI model built for coding and long-running tasks, tying OpenAI's GPT 5.6 sole on key benchmarks. It's available via the X AI API and Cursor, with no price increase from the previous version.

Canvas mode is a feature that turns a plain spreadsheet into an interactive dashboard or mini-app, staying in sync with data changes. It's available to paid Google Gemini subscribers and work/school accounts.

Chat with AI

Loading...

Pro features

Go deeper with this episode

Unlock creator-grade tools that turn any transcript into show notes and subtitle files.