Ep 744: Leaks show insights on OpenAI and Anthropic’s next models, Claude Computer Use goes viral and more
42m 44s
This week's AI news highlights rapid advancements and strategic shifts among major tech companies. OpenAI has completed a new model codenamed "Spud," anticipated to launch imminently, with leadership claiming it will accelerate economic productivity and refocus efforts on enterprise solutions. Concurrently, Anthropic introduced a feature enabling its Claude chatbot to autonomously operate a Mac, performing tasks like file management and app control, marking progress toward agent-style AI. OpenAI also revived plugins for its Codex assistant, allowing developers to create custom integrations for improved coding efficiency. In a significant ecosystem move, Apple announced plans to integrate third-party AI chatbots, such as Google Gemini and Claude, with Siri in iOS 18, expanding user options beyond ChatGPT. Additionally, a U.S. federal judge temporarily halted government directives that sought to ban Anthropic's AI tools over supply chain risk allegations, permitting their continued use during ongoing legal review. These developments underscore the intense competition and fast-paced innovation in AI, with a clear trend toward more integrated, practical applications for both consumer and enterprise use.
This is the Everyday AI Show, the everyday podcast where we simplify AI and bring its power to your fingertips. Listen daily for practical advice to boost your career, business and everyday life. I think right now the best AI model in the world, if you have the budget and you have the patience, is GPT-54 Pro from OpenAI, which seems like it was released months ago. It was announced this month, yet we already have new AI news that OpenAI is working on a newer model, codenamed "Spud" and it could be right around the corner. OpenAI must think it's pretty good as they say it will reshape the economy and their doubling had count to prepare for it. And well, they're not alone, as their biggest AI competitor and Thropic, well, recent leaks show that they're also cooking up a new model that could arrive soon, but apparently it's so powerful the company is concerned about its capabilities. This week was a heavyweight week for both the big AI companies, a ton of AI news, leaks and updates and we've got it all. Oh, and if you stick around toward the end of the live stream version of our show, you'll actually be the first in the world to hear of a new AI update from what are the big four that will break to you live that well, we think it's actually really good. All right, I'm excited to get to you the most important AI news of the week. I hope you are to let's dive into it. If you're new here, welcome, what's going on? My name is Jordan Wilson. This is every day AI today, the live stream podcast and free tailing news that are helping everyday business leaders like you and me keep up with the non stop avalanche of AI updates because they are literally every single day. I tell you what's important, what's not you take that information to be the smartest person in AI to grow your company and your career. So it starts here with the unedited unscripted live stream podcast, but take it in the next level, make sure you go to our website at your everyday AI calm. Each day we recap the live stream in the newsletter. So you can quickly catch up and we give you all of the other important AI feature updates and everything else that you need to know to stay ahead. So without further ado, let's get straight into it and give you the AI news that matters for the week of March 30th. Well, the first one, it's a short name for a apparently really big model. So open AI, their new model nicknamed Spud. Well, could be dropping any week or any month now. So according to reports, open AI has completed development of a new AI model, codamed Spud capability. The company says could significantly boost productivity and help refocus its own internal strategy toward focusing more on enterprise tools and integrated products. So they come as opening at open AI has shelved a bunch of well, kind of popular products such as Sora and they've delayed some of their other efforts as well, the concentrate on their core business and well, at least it looks like for the future that maybe whatever this new codenamed Spud model is actually call right whether that's a GPT 555 a GPT 6 we're not sure. But according to the information open AI recently finished pre training work on Spud and plans to reveal it within weeks with a CEO, the CEO open AI CEO, Sam Altman reportedly saying the model can really accelerate the economy, it claimed that signals a push for broad commercial impact. Spud is expected to support open AI's plan to build a productivity super app by combining chat GPT codex and the company's browser call atless, potentially making multi to work flows faster and more tightly integrated. So while specific technical details about Spud's architecture or capabilities were not disclosed in the reports, the emphasis from leadership suggests improvements in productivity, multi modal understanding or more advanced reasoning and tool use. The timing of Spud's completion aligns with open AI's decision to redeploy resources away from consumer experiments like Sora and their adult node which they delayed and toward enterprise and robotics implying Spud will be a cornerstone of that strategic shift. So pretty big news here from open AI yeah I was like looking at the calendar and I'm like wait what has it been like two months since GPT 54 and I'm like no we're still in March and GPT 54 was released in March so the pace of development obviously if you've been following AI for more than a couple of months is straight up blistering here. And I think maybe one of the reasons we've seen this is continued pressure and great models and improved productivity gains specifically I think from anthropic where I think they've been winning 2026 so far at least when it comes to shipping products to consumer and then obviously from Google as well. All right speaking of things from anthropic that's our next piece of AI news and this one was big well it was at least extremely viral we'll see how big it ends up being but I think this is something I could kind of shape the next agentical layer so. Anthropic has rolled out a research preview feature that let it's that lets its chatbot clawed actively use a Mac to perform task for users making a step toward more autonomous agent style AI that can operate apps type click and use connectors for services like Google calendar and Slack. So the new feature is in a limited research preview and it's right now only available to paid cloud users so if you're on the pro or the max plan and it is restricted to computers running Mac OS making it immediately relevant only to yeah if you're a paid user on Apple. So here's what it is and well why it's pretty cool it can perform real world actions on your computer right so a lot of these capabilities were previously available kind of in cloud code work and cloud code but here's what's new well cloud can now perform real world actions on your computer it can open files use your browser and developer tools in the browser it can type and move the cursor and even choose connected apps that it has. Access to inside cloud or manually control to complete task such as transferring files to your phone or batch resizing photos. Anthropic says cloud will always ask for permission before acting and allows users to stop agent at any time aiming to balance convenience with user control also the company implemented automated safeguards to scan for prompt injections and other common attacks against agents and it's. Able's some apps by default but it warns the feature is new it may contain errors and it should be used with highly it should not be used with highly sensitive data. So I don't know live stream audience podcast people let me know in a comment. If you use this yet I've used it I think it's extremely impressive we may go give this the total hands on treatment on Wednesday so yeah if you are new to the show on Mondays we go over the a I news on Wednesdays we go pretty deep hands on demo with one tool and then on Fridays we kind of do a recap style of new features and Tuesday Thursday you know we rotate different shows in there but we may go super in depth with this one on Wednesday. My quick take on this is it's buggy obviously it's the first of its kind but it's extremely impressive and I do think it represents the not the final layer right but it represents the next layer of agent work you know I think until some of the protocols improve you know that the eight the eight agent agent protocol mcp's right all of these other things. I think there's still even with all of these protocols that essentially help AI agents talk to other agents and help large language models talk to other large language models even until these things improve I still think the next layer is well it's agents using your actual computer and this is a very you know open call style update here right but this is different. This isn't just being able to save files upload files access your local directory yes that has already been available in quad co work and in quad code this is actually controlling your entire computer so I like this so far like I says when like I said when it works great doesn't always work but being able to open different documents on your computer being having quad the desktop app control open use other apps on your computer. That's kind of the new capability here right you know some of my testing I'm having you know quad open perplexities comment browser open open a eyes atless browser right use different browsers use different desktop programs that normally might not be able to talk to or access other AI so pretty big one if you think we should give this the Wednesday deep. Deep.
Deep dive treatment just say computer use in a common there. So I I know if we want this so It was crazy viral though. So when open a or sorry when in drop-ig release this on Twitter I think the release video got like 75 million views almost instantly so pretty big All right Next piece of AI news this might seem like a smaller one But I actually decided to highlight this as one of our main AI news stories for the week Although there's probably even bigger AI news stories from open AI But I think this one is actually really important. That's because well plugins are back right if you've been a long time listener of the show You know I was very very bullish on plugins early on inside chat to PT super bum that open AI actually disabled them You know discontinue support eventually, you know led to GPT's and apps and other things But now open AI has quietly updated its codex program to let developers create custom plugins and elect general users use them So open AI has introduced plugin support for codex allowing developers to add custom skills and External integrations to their coding assistant. So the most newsworthy detail here is that plugins can include 3 package scripts which are skills plus configuration files so codex can run tested code snippets instead of just generating code from scratch each time Which that obviously reduces the hallucination risk and cuts inference costs and response time as well So plugins can connect codex to external services through MCP servers and developers can upload MCP configuration files to control sandbox middleware and environment behavior So open AI also with this launched a plugin directory with more than a dozen pre-built integrations including plugins that let codex edit Google Drive files in review get hub repo changes So this follows a similar capability launched by Anthropic for cloud code a couple months ago They're kind of plug-in ecosystem and Anthropics already supports sub agents and codex just rolled those out as well So open AI says plugins are designed to help software teams keep development tools and configurations synchronized across multiple engineers open AI accounts reducing code inconsistencies on collaborative projects Here's why I think this is super important number one Everyone including open AI is straight up sleeping on codex, right? I do think maybe this is one of the reasons that open AI is Shifting away from having codex as its own dedicated app and moving forward with the super app where it kind of rolls codex chat GPT and the Atlas browser all into one as because codex is a freaking beast. Yes, it is better. I don't care what anyone says I use these tools more than 99.9% of the population Right codex right now with GPT 5 for high extra high is better than cloud code with opus 4 6 Better than cloud co work with opus 4 6. It is just better You know if you want something that looks nice and done quickly You can use cloud code cloud code work if you want something done the right way And you have the time. I think codex is by far better. Yeah, it creates ugly front ends. We get it But it just gets the code right. But here's the thing Even if you are not using it to code it is amazing at just doing every day Now it's work tasks right so in the same way that I think that in Thropic kind of not maybe pivoted but kind of Remarketed cloud code essentially as cloud co work right because they realize that it's great for even non coders I think that's what we're going to see with open AI in this new super app because maybe they're slowly starting to realize that codex is actually really good for people that are not Running software not software development teams. Yes, it's good for those people obviously But it's good for everyday knowledge work and I think plugins combine with the kind of the news of the upcoming super app is Really significant of that and it's showing that I think the average knowledge worker is going to be using the codex platform a lot Whatever that looks like in the new super app. I mean that remains to be seen But what this means to you if you have not used codex yet start using it now, right? These plugins I think do make it easier And you can probably find some great use cases right even one of the plugins I just mentioned there Uh Google drive Right being able to edit Google drive files. That's pretty big right? That's pretty big It's a you know big short coming a lot of the you know front-end AI connectors not being able to actually edit files All right Our next piece of AI news well apple right we're gonna all the big companies here So apple may actually get AI that works. We'll see they've been saying that for years Like the the boy that cried wolf will see if we believe them this year, but Maybe it will happen because according to bloomberg apple will soon let third party AI chatbots such as google's john i and enthropics Claude integrate with Siri starting in ios 27 That means iPhone users can route unanswered series queries to their preferred chatbot app So also apple announced that at their upcoming Wdc keynote They're gonna have a lot more of an AI focus and third party Integrations that will work with ios. So essentially seems like apple saying like yeah We've spent billions of dollars and three or four years and we can't get it right So essentially we're gonna open up the generally extremely restricted apple ecosystem and allow users to integrate with other AI chatbots right so if you've been like me and you're using your iPhone and you're like this thing is You know a very expensive dump phone and there's no AI that works in it. It's kind of my thought right Siri doesn't work It doesn't do anything um, you know, yeah, they have these oh these right tools there. They're useless right so maybe now We can all finally get AI that actually works on our devices So according to reports users will choose which services Siri can access through the new extension settings in ios27 ipad os27 and mac os27 found in the apple intelligence and a Siri panel of settings So that the change ends the practical exclusivity of apple's prior open AI tie up right so they've had this feature Not very well integrated. I would say right where essentially Siri can just kick things over to chatbt But open AI's chatbt will still remain supported, but will no longer be the only external bot that Siri can call Apple also is still planning a major Siri overhaul which we've been hearing about for years and will ship its own Siri chatbot built on the google Gemini models which we've talked about pretty extensively on that show on the show While extensions give users the option to direct requests to other chatbots instead of just Siri So Bloomberg reports that apple will announce these new features and the new updated Siri that can talk to other chatbots at their Wdc keynotes of this summer in June with a feature arriving in ios27 All right We'll see if that actually happens right apple right it's it's it's it's kind of funny because Two years ago right apple was like oh Apple intelligence you know we are AI and it's gonna be the best AI ever and then they got sued because they Didn't release anything that actually works. There was nothing actually intelligent that apple released And then they kind of quote unquote took a year off You know from their big wwdc that's the world wide developer conference right that's their one big time a year They come out with all their announcements. So you know two years ago they're like oh yeah We are apple intelligence. We are the smartest AI in the world They got sued because they couldn't deliver so then they quote unquote took a year off from AI they didn't really You know announce anything of substance at last year's wwdc and now apparently this year they're back to being the apple intelligence So hopefully they've learned their lesson and they won't lean into trying to redefine the actual AI category Probably not a good idea Especially if all they're ultimately doing here is allowing users to use better and smarter AI. So It should actually be pretty telling How they approach this from a branding angle But I do think however the markets right if you care about that I think the markets will actually like this move from apple because they're like yeah apple We understand you can't build AI So you should probably start integrating with as many third-party AI chatbots as possible All right our next piece of AI news well this one is the ongoing AI moves too fast to follow but you're expected to keep up Otherwise your career or company might lag behind while AI native competitors leap ahead But you don't have 10 hours a day to understand it all that's what I do for you But after 700 plus episodes of every day AI the most common questions I get is where do I start That's why we created the start here series an ongoing podcast series of more than a dozen episodes You can listen to in order it covers the AI basics for beginners and sharpens the skills of AI champions pushing their
companies forward. In the ongoing series, we explain complex trends in simple language that you can turn into action. There's three ways to jump in. Number one, go scroll back to the first one in episode 691. Number two, tap the link in your show notes at any time for the Start Here series, or you can just go to StartHereSeries.com, which also gives you free access to our inner circle community where you can connect with other business leaders doing the same. The Start Here series will slow down the pace of AI so you can get ahead. Brahma, and it's not over, but we have a new chapter in the novella. So a federal judge has temporarily stopped the Pentagon and other federal agencies from enforcing directives that would have immediately halted the government's use of Anthropics AI tools. Using the company's widely used cloud system, available while a broader legal fight plays out. So a US district judge in California issued an order late last week, preventing the enforcement directives from President Trump and Defense Secretary Pete Hegseth that sought to bar Anthropics tools from government use for now. So essentially the federal government labeled Anthropic a supply chain risk for a couple of reasons, which we get to here in a minute. But essentially a judge said, no, doesn't make sense. So the order now lets Anthropics AI continue to be used inside the government for now and by outside contractors working with the military while the lawsuit proceeds. So the judge wrote that the government actions looked aimed at quote unquote, crippling Anthropic and a chilling public debate. And the judge characterized statements by officials as appearing to be classic first amendment retaliation. So the dispute began after public criticism from President Trump and Hegseth, who labeled Anthropic a supply chain risk. And the first public use of that designation against a US company ever in a label typically reserved for companies tied to adversary nation. So yeah, it was very strange that the federal government decided to label Anthropic a supply chain risk because that's never literally happened against a US company. Then in traffic in response, sued the Department of Defense and other agencies earlier this month saying the government's designation in public attacks harmed its business and violated its free speech rights. So the judge noted that officials public comments attacking Anthropic were more on political grounds. For example, they called the company woke and its employee in Anthropics employees, quote unquote, left wing nut jobs rather than pointing to specific security defects. So this drama is not yet over. Anthropic essentially, they had two little clauses that they wanted to have in their agreement with the government for reasons they said that would ultimately give them more protection over how the military would not use its AI. As an example, they said they didn't want the government to use its clawed systems for fully autonomous weapons in war without human oversight. Right. So we'll see how this continues to go on, but it is again, one of those stories that is going to continue to drag on. But the latest one here, pretty big update, a judge essentially saying no government, you were wrong. You can't do this and this was political and there's really not a lot of merit. So yeah, it will continue to be legislated. All right. Our next piece of AI news, a little technical one here, but it actually had pretty big ramifications both instantly and in the long run. So Google researchers have introduced a new methodology called TurboQuant, a two stage vector compression method that reduces transformer key value cash memory by about six times while preserving downstream accuracy on tested workloads. So the most newsworthy part here is that TurboQuant lets models quantize key value caches down to three bits without retraining, which could sharply improve lower memory requirements for large language models. So the system requires no model retraining, which that's huge, making it potentially compatible with existing open models and easier to adopt in current production stacks. So here's in a simple way, right? This is technical. I had to read this this one a couple of times because you know, as much as I talk about AI, I am not super technical on the pre training side. So it's kind of like a super shredder, right, for an AI's memory. So when you chat with an AI, it has to store a lot of data to remember what you just said. And usually that takes up a ton of well expensive memory in turboquant from Google. So essentially a new quantize technique, right? It just kind of squishes the data. So it shrinks the memory to about one sixth of that size without losing any of the information. And by doing that, it obviously speeds things up because the data that has to remember, right, about your conversations is smaller that AI can quote unquote read it up to eight times faster. And the impressive things here is well, this new technology or technique can apparently be applied to any new model. So it doesn't have to be only new models that haven't been developed yet. So there's no extra work. So you don't have to retrain the AI. So this can be applied to open models. So as an example, it can work instantly on models, maybe like Google's open source, Gemma. So also interestingly enough. So this happened late last week and it instantly triggered a pretty sharp, but probably temporary sell off on the stock market, you know, memory chip stocks. So pretty big deal here. We'll see what Google does with this technology, how it may be used across the spectrum. But this could ultimately, right, bring way more powerful models to way smaller devices. Right. As an example, as of recently, there's been this big, you know, open claw and open claw ask kind of surge, right. And a lot of people are buying very expensive, you know, $10,000 max studios or, you know, Nvidia DGXs to run the most powerful local models that they can so they don't have to pay for cloud inference, right. They don't have to rack up API bills with in traffic, Google or open AI. But the problem is, well, you have to have a huge, very expensive computer to run these models. Maybe you might be able to get a six times more powerful model on the same size or just have a much smaller computer or a, you know, not as powerful GPU to be able to bring all of these things. So pretty, pretty exciting news for the future of AI development where we just might have way better local models available on phones and computers in the future. If TurboQuant ends up being what it could be, which is a game changer for how large language models work. All right, here's, well, our pretty big, biggest story of the week, although we are going to be able to break, uh, break some news here in about five minutes from one of the big companies. So, uh, a major leak at inthropic has exposed nearly 3,000 internal files. Well, probably 99.9% of them weren't very important, except there was a new unpublished draft about a new powerful model from inthropic called Claude, uh, mythos. I think that's how it's pronounced, right? In details of an even larger tier named Copapera. So, uh, inthropic did confirm they were reporting to fortune that an accidental data leak exposed nearly 3,000 assets uploaded to its content management system on its website, but marked as private, making them publicly accessible in a data lake. So the leak collection included unused marketing assets, PDF, images, all that stuff, also employee, incorporate event information. But the big thing was a draft blog post about a new AI model labeled Claude mythos. So according to the lead drop draft, uh, mythos is by far the most powerful AI model that inthropic has ever developed and the company calls its performance a step change. The model is in trials right now with selected early access customers. So the documents revealed a planned new top tier as well, uh, right now codenamed Copapera. I think that's how it's pronounced, uh, which would sit above inthropic's current highest tier, which is opus, making Copapera the company's late, uh, largest and most capable offering. So inthropic's leaked materials warns that Claude mythos and Capapera could significantly increase cyber security risks, saying the model, uh, the models are currently far ahead of any other AI model in cyber capabilities. the league draft states.
and profit intends to study and share findings about near term cyber risks. So defenders can prepare and it is giving early access to organizations to give them time to harden their code basis against potential AI driven exploits. But I think the potentially bigger news here might ultimately be this new models access because reports point to a very limited early access rollout with select customers only. And those leaked materials described access expanding gradually through the Clawed API rather than broad availability in a standard paid plan. So yeah, this might be the first, well, major model we've seen from any company that might not be available to regularly paying business users. So if you have a Clawed paid plan where you're paying monthly and you expect, oh, well, I'm going to have always the most powerful models from Anthropic, well, maybe not or maybe not right away. Which I think is actually a pretty big pivot overall with how these companies work. We've seen reports that open AI and Anthropic are likely both going public this year. And this could be a new kind of tactics to bring in more revenue, maybe right before an IPO will see. But again, reports are saying that this might even if you're a $200, you know, like myself, you know, paying $200 a month for Clawed Max, you might not get access to the new Clawed Mythos whenever it's released. All right, speaking of release, well, we're safe now. The embargo has lifted and we can talk about new updates from Microsoft. So yes, if you are listening here on the live stream, you are the first in the world to hear about this. So Microsoft is now just introducing that co-pilot co-work is available through its frontier program. And there's some new, which I think are really, really good updates that they're rolling out to researcher. All right, but first co-pilot co-work, all right. So we know that Microsoft co-pilot co-work isn't new. They announced this a couple of weeks ago, but this is now pretty big news because they're making it available via the frontier program. So that's when a lot of companies are now going to get access to it. Also, they're introducing a Microsoft 365 co-pilot feature designed for long running multi-step work instead of just the one-off chat prompt. So that's with co-work. This is a kind of partnership. With Anthropic, it's not really a white label version of co-work, but it kind of is, right? They're leveraging that technology. And obviously Microsoft is a big investor in Anthropic. So initially, this was rolling out to a very small beta group. So now it's going to be rolling out to a pretty big group in the frontier program. So the company said co-pilot co-work, let users describe the outcomes they want, then create a plan, works across files and tools, and then shows a visible progress. So you kind of the screenshot here I have for our live stream audience, you can kind of see it work through its entire plan. Very similar to if you've used Anthropics, Clawed, Co-Work, but Microsoft style. So Microsoft said co-pilot co-work includes skills from both Clawed and Microsoft. So it's not just an exact duplicate of co-pilot co-work because it does include obviously a lot of specific and Microsoft exclusive capabilities, including calendar management and daily briefing, and the company said it can handle both one-time tasks and repeatable workflows such as monthly budget review. All right, but here's actually some of the announcements that I am maybe more personally excited about. So Microsoft has announced a new and improved researcher agent. Here's the big part though, built on multi-model intelligence. So it's aimed at a more complex knowledge work by synthesizing information across sources and producing cited reasoning analysis that users can act on. So one of the most important additions is the new critique feature where OpenAI's GPT model will draft a response and then Anthropics Clawed model, reviews it for accuracy, completeness, and citation integrity before delivery, showing that Microsoft is now productizing a multi-model review workflow rather than relying on a single model's first pass. Y'all, I can't believe it is essentially now quarter two of 2026. And we're now just getting this for the first time from one of the big players. Granted, it's really only Microsoft or Google that could do this, right? Essentially bringing a new model in from a different company, right? So that's completely different training. It works in a completely new way. And essentially having those two models work well with each other, but technically against each other, right? So having OpenAI's GPT create something and then Anthropics Clawed essentially tear it apart, right? For anyone that's a power of user of AI, if you're using it for high value work, this is what you've been doing manually, right? For me, this is probably what I spend 70% of my time doing, right? I think there are some great third party solutions, such as perplexities model console, but it's honestly very expensive, right? It eats up a bunch of credits. There's some other third party, less popular tools that kind of do this, but it's actually a pretty big deal. That Microsoft is the first of the big four, right? So that's Microsoft OpenAI and Thropic and Google. They're the first one to provide this. And like I said, they can't really-- anyone else can't really do this. I think technically Google could, right? Because Google is also a big investor in Anthropic. But I think they're probably competing a little bit too closely on the model side as well. Because right now, Microsoft is even though they are developing kind of their next generation of models in their Microsoft AI, Mustafa Salimon, now working on that side. They don't really have today, at least, frontier level models. So this is pretty big, both expanding co-pilot co-work to the frontier program, which is a lot of enterprise organizations here in the US. But then also with these new updates in the model console, which is huge, right? So that's a new feature. So you have the critique feature in researcher. But then you also have the model console feature, which lets users compare responses from different models side-by-side. So they can see where answers agree, where the diverge and what each model contributes. So like I said, that's something that's been kind of available inside perplexity. But pretty cool to see this offer now from one of the big four. All right, that's it for our big stories of the week. But let's roll into our what's new and what's next. So this is the kind of bullet point roundup of all the other big stories. In other weeks, some of these might have been some of the biggest stories of the week. But there's a lot going on this week with all the new leaks, big bottles released from Microsoft, a ton going on. So let's go over this, what's new and what's next. So in traffic, release a new economic index, which shows a widening AI fluency gap. Google released their updated Luria 3 Pro. You can create audio tracks up to three minutes. Open AI reportedly close their funding round. That's reaching $120 billion. All right, meta introduced, SAM 3.1. So that's their segment, anything model. So that allows you to segment any object and image or videos from simple prompts like click or boxes. Soft bank reportedly secured a $40 billion loan tied to their open AI investment. A report so that Chattu B.T.'s new ad pilot passed $100 million in annualized revenue. So we saw some reports that even though maybe it wasn't as successful as they had hoped, well, they've already brought in $100 million in annualized revenue. My hot take is they're going to like 10X that by the end of the year. It's going to be a cash machine. The White House announced members of the president's council of advisors on science and technology, including Mark Zuckerberg, Jensen Wong, and others. Google released Gemini 3.1 Flash Live and Search Live. We went over that on our Friday show, going over our new Friday features, our weekly show. Open AI upgraded their Chattu B.T. shopping and product discovery, really emphasizing shopping, less in product discovery. More Google release and updated version of Google translate. It's actually really good. I've been using it. We went over that Friday as well. Arm introduced that it will start making its own chips. Open AI revamped Chattu B.T. shopping with a Genetic Commerce Protocol and the Walmart app. Luma AI launched a pretty impressive and kind of out of nowhere, a new multi-modal image model called Uni1, fairly impressive so far. Drugmaker Eli Lilly agreed on a $2.7 billion deal within Silico for AI-developed drugs. Sam Aldman reportedly has now shifted his own priorities and will no longer directly oversee security or safety at Open AI. Chattu B.T. rolled out its library feature for central file management. And Throbbing is reportedly.
working on a computer use for mobile after their very viral computer use for Mac OS. In Gemini Business, Google is testing skills. Mata conducted another round of layoffs this time around 700 employees amid their AI shift. Open AI shelves, reportedly their adult mode, in definitely alongside cutting and killing off Sora as a dedicated app for now. Mata rolled out new AI shopping experiences on Facebook and Instagram. Open AI officially relaunched their nonprofit arm and named leaders and they planned to spend at least a billion dollars in doing so. Manus now offers a computer control from their iPhone app. I haven't tried that one yet. I do have a manus subscription, so I'll have to give that one a try. There's a new benchmark in town, Arc AGI 3 since Arc AGI 1 and 2 have essentially been saturated, but we're still saying we haven't achieved AGI. Well, now there's Arc AGI 3. A new benchmark that has launched with $2 million plus in prizes. Apple will reportedly, like I said, we are to cover that one. They are opening a series to rival AI assistants. Google Gemini is launching a chat history and memory import tools to make it easier to start with Gemini to transfer over. A chat GBT is getting unified Google Drive integration for business and enterprise. Soono release version 5.5. A lot of cool features in there allowing you to also kind of use your own voice, which is cool. And then last but not least agile robotics in Google DeepMind announced a strategic research partnership. Y'all, that was a ton happening this week in AI. If you missed anything, well, we just covered it. So you don't have to spend hours each and every week saying, Oh, what's happening in AI? What's worth paying attention to? What's not? What should I be using in my company? That's what we do on our Monday show. So I hope this was helpful. If so, please subscribe to the show. Please leave us a rating. I'd really appreciate that. It takes like 30 seconds, whether you're listening on Apple podcasts or Spotify. And then if you haven't already, make sure you go to your everyday AI.com. Sign up for the free daily newsletter where we recap, not just each day's podcast, but everything else you need to know to be the smartest person in AI at your company. So thank you for tuning in. We hope to see you back tomorrow and every day for more everyday AI. Thanks y'all.
Podcast Summary
Key Points:
OpenAI has developed a new AI model codenamed "Spud," expected to launch soon, with claims it will significantly boost productivity and reshape the economy, signaling a strategic shift toward enterprise tools.
Anthropic released a research preview feature allowing its Claude chatbot to autonomously perform tasks on a user's Mac, such as opening files and using apps, representing a step toward more advanced AI agents.
OpenAI reintroduced plugins for its Codex coding assistant, enabling custom integrations and external service connections to improve development workflows and reduce errors.
Apple plans to allow third-party AI chatbots like Google Gemini and Anthropic Claude to integrate with Siri in iOS 18, moving away from exclusivity with OpenAI to enhance its AI capabilities.
A federal judge temporarily blocked U.S. government directives that aimed to halt the use of Anthropic's AI tools over supply chain risk concerns, allowing continued use pending legal proceedings.
Summary:
This week's AI news highlights rapid advancements and strategic shifts among major tech companies. OpenAI has completed a new model codenamed "Spud," anticipated to launch imminently, with leadership claiming it will accelerate economic productivity and refocus efforts on enterprise solutions. Concurrently, Anthropic introduced a feature enabling its Claude chatbot to autonomously operate a Mac, performing tasks like file management and app control, marking progress toward agent-style AI.
OpenAI also revived plugins for its Codex assistant, allowing developers to create custom integrations for improved coding efficiency. In a significant ecosystem move, Apple announced plans to integrate third-party AI chatbots, such as Google Gemini and Claude, with Siri in iOS 18, expanding user options beyond ChatGPT. S.
federal judge temporarily halted government directives that sought to ban Anthropic's AI tools over supply chain risk allegations, permitting their continued use during ongoing legal review. These developments underscore the intense competition and fast-paced innovation in AI, with a clear trend toward more integrated, practical applications for both consumer and enterprise use.
FAQs
According to the host, the best AI model currently is OpenAI's GPT-54 Pro, assuming you have the budget and patience for it.
OpenAI's 'Spud' model is expected to significantly boost productivity, reshape the economy, and support a productivity super app by integrating ChatGPT, Codex, and the Atlas browser.
Anthropic released a research preview feature that allows Claude to actively use a Mac computer to perform tasks, such as opening files, using browsers, and controlling apps, moving toward autonomous agent-style AI.
OpenAI quietly reintroduced plugin support for Codex, allowing developers to create custom plugins that enable Codex to run tested code snippets and connect to external services, reducing hallucinations and improving efficiency.
Apple plans to let third-party AI chatbots like Google's Gemini and Anthropic's Claude integrate with Siri in iOS 27, allowing users to route unanswered queries to their preferred chatbot app.
A federal judge temporarily stopped the Pentagon and other agencies from enforcing directives that would halt the government's use of Anthropic's AI tools, citing concerns over free speech and supply chain risk designations.
Chat with AI
Loading...
Pro features
Go deeper with this episode
Unlock creator-grade tools that turn any transcript into show notes and subtitle files.