Everyday AI Podcast – An AI and ChatGPT Podcast - Ep 826: ChatGPT goes Jarvis Mode, Claude can learn from you, Google unleashes spark agent and 7 more AI updates you can use today

Episode Date: July 24, 2026

Over 3 hours, OpenAI, Anthropic, Google AND Microsoft all dropped new AI upgrades that are live. How you use AI in your work literally changes every day, as frontier labs are racing to roll out big q...uality of life updates between big model drops. How can you keep up? With our Friday Features show, where we break down the latest AI updates that are live and available to all, and we tell you how to use them and why they matter. This week did not disappoint. You don't want to miss what's now at your fingertips. JARVIS mode, anyone? ChatGPT goes Jarvis Mode, Claude can learn from you, Google unleashes spark agent and 7 more AI updates you can use today -- An Everyday AI Chat with Jordan WilsonNewsletter: Sign up for our free daily newsletterMore on this Episode: Episode PageToday's Episode on LinkedIn: Thoughts on this? Join the convo on LinkedIn and connect with other AI leaders.Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineupWebsite: YourEverydayAI.comEmail The Show: info@youreverydayai.comConnect with Jordan on LinkedInTopics Covered in This Episode:ChatGPT Health Syncs Apple and Medical DataClaude Voice Mode Adds Opus and SonnetClaude Voice Mode Supports ConnectorsMicrosoft MAI Image 2.5 Pro Launch DetailsMicrosoft MAI Image Model Benchmark PreviewGoogle Gemini 3.6 Flash and Flashlight ReleaseGemini 3.6 Flash: Token Efficiency UpgradesGoogle Gemini Spark Agent for Task AutomationClaude Cowork "Record a Skill" With Voice NarrationChatGPT Voice on Desktop: Full Jarvis ModeChatGPT Voice Controls Apps via App ShotsCross-Platform AI Skills Sharing (Claude, Codex, GPT)Timestamps:00:00 Recent AI feature updates05:22 Unified health data management09:52 New voice feature explanation11:28 Launch of Microsoft's new image model16:17 Explaining the Gemini 3.5 models17:11 Developers benefiting from 3.6 Flash22:45 Introducing Gemini personal intelligence25:10 Claude Cowork's new skill feature28:32 New default feature in Claude Cowork34:22 Using AI like Iron Man35:09 Excitement for future AI advancements38:20 Wrapping up and subscribingKeywords: ChatGPT Jarvis mode, ChatGPT Health, OpenAI, Anthropic, Claude voice mode, Claude Cowork, Claude record a skill, Microsoft, MAI image 2.5 Pro, AI image generator, Google Gemini, Gemini 3.6 Flash, Gemini 3.5 Flashlight, Gemini Spark, Google AI agent, AI-powered personal assistant, AI agents, Agentic workflows, Multimodal AI, Token efficiency, Image generation, Voice-activated AI, AI-powered task automation, App shots, GPT Live, Remote browser, Computer code execution, Slack integration, GitHub integration, Notion, PowerPoint AI features, Workspace plans, Apple Health integration, Medical records AI, Health data privacy, Consumer AI, Chronic condition management, AI-powered document processing, AI for business, AI model benchmarking, AI for developers, AI economics, Personal intelligence, Automated triggers, Google Docs AI, Team collaboration AISend Everyday AI and Jordan a text message. (We can't reply back unless you leave contact info) Ready for ROI on GenAI? Go to youreverydayai.com/partner 

Transcript
Discussion (0)
Starting point is 00:00:00 This is the Everyday AI show, the Everyday Podcast where we simplify AI and bring its power to your fingertips. Listen daily for practical advice to boost your career, business, and everyday life. If you listened to our Monday AI News That Matters segment, I told you we were kind of in for a wild week. That's because after a somewhat slow-ish week last week, I said on Monday that this week was going to be a banger in terms of new AI features you can actually use. Sure, I knew ahead of time a few things that were going to come out, but even I was kind of blown away just yesterday at the sheer amount of new updates that all dropped within hours of each other. I mean, no joke. On Thursday alone, we got huge new AI features from Google, Microsoft,
Starting point is 00:00:55 Anthropic, and Open AI in a matter of hours. That's why on Fridays, we break down what's new, in AI that you can actually use. And we do it in a very specific way. We give you the AI features that are available for you today, not rumors, leaks, or weightless. These are instant AI upgrades that you should be putting into play today. And we've got a lot to cover on today's show. So stick with me for the next 25-ish minutes. And I'm going to tell you how chat GPT went full Jarvis mode and the new app shot secret.
Starting point is 00:01:32 that unlocks the future of work. I'm going to let you know the new Claude feature that took a page out of the Codex Playbook, but actually made it a little bit better. And I'm going to fill you in on the kind of new Google agent that you can start using today, but with one caveat. All right. Let's get into it. Welcome to Everyday AI.
Starting point is 00:01:55 My name is Jordan Wilson, and we do this every day. It's your unedited, unscripted daily live stream podcast and free daily newsletter, helping business leaders like you and me keep up with the avalanche of AI news and updates. I tell you what matters. How to use it. You take that information. You're the smartest person in AI and you can grow your company and your career. So it starts here, but make sure you go to the website at your everyday AI.com.
Starting point is 00:02:20 We're to be recapping the highlights from today's show. So if you miss any of the details, maybe you're out on a walk and you want to go back and say, wait, how did this work? It's all going to be in the newsletter as well as all of the other AI updates. you need to know. All right, let's get straight into it. I'm a big fan of the Friday Features show. I wish I would have started it like three and a half years ago.
Starting point is 00:02:42 We've been doing it now for about six months. But just for our podcast audience, I'm always showing on my screen some basic, you know, information from the company. You're not missing anything. So just FYI. But let's dive in live. I'm going to be sharing my screen.
Starting point is 00:02:58 Just kind of showing the windows for the different. announcements that we're going to be going over. All right. So let's start first with your health. Yeah. This is actually one I'm personally excited about. I know some people aren't going to be, you know, diving in headfirst necessarily to give their help data to anyone, let alone an AI company. For me, I'm absolutely wildly going to be using this. So here's what's new. Chat, GPT Health is finally roll out to everyone. It is launched for all U.S. users today. And it's expanding the feature that was first introduced in January, but it was introduced as a wait list feature. And it connects your Apple health and supported medical records directly into chat GPT. So what the heck can it do? Why is it
Starting point is 00:03:54 useful? Well, it can compare your lab results, summarize changes, since a prior appointment, track your medications, factor in your sleep activity and workouts, essentially anything that, you know, your Apple health might track and all the different applications that connect in there, well, it can bring in that data and anything else that you can update or upload into chat. So Open AI says that health conversations obviously live in a dedicated space siloed from all your other chats, right? So yeah, you're not going to have all of that information popping up via like a memory, into your other chats.
Starting point is 00:04:31 So like I said, this is already rolling out to US users. You do have to be 18 or older. But the cool thing is this is available on free plans, on paid plans, and also on iOS. So one small caveat, though, free users, you're just going to be using kind of the model that you have the most access to, which is GPD 5.5 instant. And then for other paid users, you'll be able to use the more powerful GPD 556 sold. All right. So why is this useful? Right. Because right now, obviously, there's no one place for all of your health information to live. Like, I can't tell you the amount of conversations I've had with people about this over the last few years.
Starting point is 00:05:16 And, you know, there was a company out there that was kind of, you know, AI native. They were called Forward Health. They are no more. And I was kind of excited about that concept because I'm like, there needs to be, you know, something that's not tied to a single. provider that allows you to just essentially bring in all your health records and to have a smart system that just knows all that. So obviously, there is nothing previously that would have stopped you from doing this inside of a large language model like Claude or ChatGBTGPT, Gemini, copilot, etc. But it wasn't really set up for this, right? So think of this as a specialized version of ChatGPT that's siloed from everything else. So it keeps all of your health data separate. But it is literally built for this.
Starting point is 00:06:00 this. So, you know, right now your health info is scattered across different patient portals, EMRs, apps, and it just pulls it all into the conversational layer. And there's also some kind of practical prompts baked up for you, right? So how your cholesterol might be trending or how you can summarize your blood work before an appointment, what you should ask your doctor tomorrow, etc. And that context kind of bleeds usefully into everyday tasks, like factoring a dietary restriction into your restaurant picks. So who's going to find this valuable? Well, this is obviously a consumer play.
Starting point is 00:06:36 But I think this just goes to show where opening I has started to focus more recently, right? Not necessarily just on the models, but how to make their models more useful for more people in more ways. So, you know, whether you are just someone that wants to get a little bit more serious about your health, you just have questions, or if you're managing. chronic conditions. If you're a caregiver trying to track something in your family health or just preparing for your upcoming appointment. So obviously, OpenAI says this is not intended for diagnosis or treatments and health conversations aren't used to train foundation models.
Starting point is 00:07:18 So pretty big one. I'm excited to get this. I signed up for the wait list. previously wasn't, uh, wasn't kind of selected to be part of that first group. So this is one I'm going to be jumping into head first. All right. Our next AI update, a new update, well, new ish feature from Anthropic, just kind of borrowing from the codex playbook, but I think they made it a little bit better. So Claude has some new voice mode upgrades. Oh, no.
Starting point is 00:07:47 This is not the one. Sorry, that's a little bit, uh, better than, uh, Chad GBT's. So this one. is voice mode not quite as good but also this one is a little different yeah just so many new updates even i'm getting confused and this is all i do every single day all right so this new update from anthropic is essentially expanding how and where their voice mode can be used so Anthropic has upgraded the Claude Voice Mode to run on Opus and Sonnet for the first time. That's because previously it ran only on their least powerful model, which is Haiku,
Starting point is 00:08:30 which is why for me I absolutely never used voice mode inside Claude on my phone for that very reason. But now I probably will. The cool thing as well is you can switch models mid-conversation in voice mode now works with connectors. So that is the big difference right now with the one thing that's maybe a little bit better is that their upgraded voice mode works with connectors, but it's not bi-directional. So it's still kind of the quote unquote older, dumber version of a voice mode. So you've heard me kind of rave about the new GPT Live. That's because GPT Live can listen to you and speak at the same time, right?
Starting point is 00:09:15 You can interrupt it. It might interrupt you. and it's kind of listening, thinking, and doing all these things at the same time. So this is still the old school walkie-talking mode where you talk, you wait, Claude responds. But the big benefit here is it can use connectors. So pulling in your live data. So that's not something that GPT live can do,
Starting point is 00:09:37 although you can individually upload files into GPT Live. And Open AI did say that they are working on bringing connectors. So here's who has access. You have to have a paid. Claude account to get that upgraded sonnet in Opus voice mode. Otherwise, you will be defaulted to Haiku. So free accounts do have a very limited usage and only a single connection. And all prompts then obviously go through Haiku.
Starting point is 00:10:04 And right now, voice mode does support 11 languages. So why is this useful? Well, if you want to be able to chat with your connectors and you have a paid quad plan, this is really good, especially if you don't mind kind of working in walkie talking mode. For me, I've really enjoyed kind of the, the bidirectional or duplex new way of talking. It seems like that's the future of the voice interface. So as long as you don't mind kind of this more kind of waiting and holding and, you know, only being able to talk or listen at once, not that bad.
Starting point is 00:10:42 So who's going to find this useful? I mean, anyone that's a power user of Claude, if you're doing any hands-free work, commuters, you know, executives between meetings or just operators who want to think out loud, pretty big update and release here from Anthropic. All right. Next, we have a, yes, another new image mode from Microsoft. All right. So here's what's new in MAI images or sorry, MAI Image 2.5 Pro. All these image models are always a mouthful, right? We should just, I don't know, Microsoft get a fun like a nanobanana name and then we can call it.
Starting point is 00:11:24 You know, why not like Clippy Pro? You know, 2.5. All right. So this is launched and it's launched in a lot of different places. So it is in Foundry Preview. It's in Cope. It's in Microsoft copilot PowerPoint. And it's in also a variety of other places.
Starting point is 00:11:47 But the new image mode, it's pretty good. I don't think we have a lot of benchmarks on it just yet. But I would assume that we get those probably within a couple of days. So this is Microsoft's newest, highest fidelity professional grade image model built, they say, for superior high quality imagery, detailed editing, and precise in image text rendering. So this does, Microsoft says, handles text to image generation plus controllable edits, object removal, replacement in painting, text updates all while preserving composition. Are you still running in circles trying to figure out how to actually grow your business with AI? Maybe your company has been tinkering with large language models for a year or more, but can't really get traction to find ROI on Gen.
Starting point is 00:12:41 Hey, this is Jordan Wilson, host of this very podcast. Companies like Adobe, Microsoft, and Nvidia have partnered with us because they trust our expertise in educating the masses around generative AI to get ahead. And some of the most innovative companies in the country hire us to help with their AI strategy and to train hundreds of their employees on how to use Gen AI. So whether you're looking for chat GPT training for thousands or just need help building your front-end AI strategy, you can partner with us too, just like some of the biggest companies in the world do.
Starting point is 00:13:13 Go to your everyday AI.com slash partner to get in contact with our team. Or you can just click on the partner section of our website. We'll help you stop running in those AI circles and help get your team ahead and build a straight path to ROI on GenAI. So who has access? So yeah, it is now available in Microsoft Foundry. If you do have a Microsoft 365 copilot plan, you can use this right now inside of PowerPoint. Or you can use it via the API. The pricing is $5.
Starting point is 00:13:46 per one million text input tokens or $8 per one million image input tokens. All right. And then $106 per one million image output tokens. So why is this useful? Well, I'll say this. Right now, if you're one of those organizations that can only use Microsoft copilot and you can't use anything else, then this is great. This is a nice upgrade over the previous MAI image to point of,
Starting point is 00:14:16 but it's still presumably, right, we'll see what the benchmarks say, at least from my trained eye. This is still fairly far behind GPD images too, which is the best AI image, you know, model in the world and also the nanobanato two. So those are kind of the top two. And then you have some more, you know, Chinese open versions of these. But the best two still are GPD images and Google's nanobanana. So not really in the same.
Starting point is 00:14:46 tier. We'll see what the benchmarks say, at least by my kind of quick trained eye. I still think that there's probably a quality drop off. But regardless, if you are someone, especially I think in, if you are building decks in PowerPoint via co-pilot, being able to use this image generator in there is going to be big. Right. So yeah, I know that so many people out there are kind of living inside of PowerPoint and maybe your visuals and the old clip art just won't do, I think this is going to be pretty big. So, you know, if you're trying to create product imagery,
Starting point is 00:15:22 marketing visuals, brand assets, big, it is big for those purposes when you can't use anything else. All right. Our next AI update you can use today. Yeah, it was this big of a week that I didn't even really honestly get a chance to play with a brand new model from Google Gemini,
Starting point is 00:15:43 aside from just some daily, driving tasks. I haven't really given the new Gemini 3.6 Flash a run. So here is what's new with Gemini 3.6 Flash. There's actually three different models, but I say most people are focused on Gemini 3.6 Flash, but they also unleashed Gemini 3.5 Flash light. I know confusing. So they released both the 3.5 upgrades and updates and new variations in the 3.6 flash,
Starting point is 00:16:14 which is the first of the 3.6 Flash series. Even more confusing, the Pro series is still stuck on 3.1. So, yeah, now your latest and greatest Google Gemini models technically have three different step changes. You have 3.1 Pro because we are still waiting for 3.5 or maybe they'll skip straight to 3.6 Pro. So you have 3.1 Pro. Then you have the 3.5 Flash light in 3.5 Flash Cyber. And then the only one on the 3.6 tier is Flash. So extremely confusing.
Starting point is 00:16:50 They're all for different purposes, but let's break down at least what's new in these models. So Gemini 3.6 Flash, essentially not that much smarter per se, but it is cheaper and more token efficient. So Google says that Gemini 3.6 Flash consumes 17% fewer tokens than 3.5.3.4.5. 5 flash while taking fewer reasoning steps and tool calls in multi-step workflows. Then you have 3.5 Flash light, which is the cheaper and faster version of Flash. All right. So Flash is like your, you know, cheaper version. And then Flashlight is the cheaper and faster and more lightweight version.
Starting point is 00:17:35 So the 3.5 Flashlight is the fastest model in the 3.5 series at 350 output tokens per seconds built for low latency, high throughput work like a genetic search and document processing. And then you also have that third model, Gemini 3.5 cyber, which is tuned specifically for cybersecurity tasks. So who has access? Well, everyone. If you're a paid user, you can go into Gemini, you know, your Gemini.com, the Gemini API inside AI studio, Android studio, anti-gravity, Gemini Enterprise, right? Literally, wherever you're using Google Gemini for the most part, you can find the new 3.6 Flash there. Flash cyber, though, is limited access only. That's the government and trusted partners similar to, you know, a mythos or Open AI has their
Starting point is 00:18:26 daybreak, I believe is that what their version is called. So why is this useful and who should be using it? Well, if you are a developer, especially if you live inside of the Google ecosystem, If you're working on, you know, just document processing in bulk, if you're working on any agenic tasks, that's where it's going to pay off, right? It's cheaper and more token efficient. You pay less per token and burn fewer tokens in the process. And that's kind of the agentic economics 101 story there. So 3.6 does also, 3.6 flash does deliver higher precision coding with fewer unwanted edits
Starting point is 00:19:05 in reduction, reduction in loops, essentially, right? So if you're a dev running production agents inside Gemini, or if you found the previous version of 3.5 Flash to work in terms of price per performance, I think that 3.6 Flash, although it's not a big bump in terms of capabilities, it is just a little faster and cheaper and more efficient. So any developers that have already been using the Flash series is going to be good for you. And then also Flashlight, important here, will power the future Google search AI overviews where speed really matters the most. So the big news here is, well, we still don't have a new pro version.
Starting point is 00:19:47 Yeah, we were kind of reports this week. If you follow our newsletter, we covered this, said that the pro version may get delayed a few more months. And then we similarly got news from frequent open a, frequent, every day, I guess Logan Kilpatrick found on the show a handful of times. I did say that Google is pre-training Gemini 4. So who knows? Maybe we technically won't get a Gemini 3.5 pro or a 3.6 pro. Maybe we'll just be in for a longer wait and we'll get a Gemini 4 pro.
Starting point is 00:20:24 Not sure, but Google did say that they're pre-training Gemini 4. And we saw reports that the pro series is delayed. So if you put two and two together, maybe that just means there will be no 3.5 or 3.6 pro and we'll just be going straight to 4 pro. We'll see. All right. Let's go to our next one. And yeah, this one, I'm both excited, but a little bummed. But this is kind of how Google rolls out some of their products.
Starting point is 00:20:53 Maybe I'm just being a little snobbish in, you know, wanting the best tools for where I actually need them. But yes, we finally have a new 24-7 personal agent from Google. That just works. I'm talking about Gemini Spark. So not technically new, but new to most everyone. Because this did launch in May, but only those two on the ultra plan, which is that $200 a month plan from Google Gemini. So we didn't cover it on the show in May, right?
Starting point is 00:21:31 For the most part, I only. cover on the Friday features, those that come out to the base paid plans, right? Your $20 a month plan, not your plans that are $100, $200 a month. So that was launched for ultra subscribers in May. But now Google has released Gemini Spark to all pro subscribers as well. So that's Google's personal AI agent for automating complex tasks. It's rolling out. I have access to it right now.
Starting point is 00:22:01 So it performs multi-step tasks on your behalf, research, planning, complex workflows, while you focus on, well, whatever else. So it's kind of built on these three different concepts. So you have your tasks, which is your high-level goal. You have your skills, right, that opens skills standards. Those are your reusable instructions with context and then schedules. So those are automated triggers like times or conditions. All right. So like I said, if you are on that,
Starting point is 00:22:31 that base paid plan, you should get access to it now. Especially if you're in the U.S., Google did say that other countries are going to be getting access to it soon. So here's the downside or the caveat. I have no clue when it's going to be rolling out to workspace plans. This is usually one of those things that they say soon. So if your business runs on Google, on a paid Google Gemini account, maybe you have access to it, maybe you don't.
Starting point is 00:22:59 But if you are on a standard paid, you know, Google Gemini plan like me, right? So my personal Gmail, I have a paid Gemini account for my personal Gmail. I have a paid Gemini account for my work emails. So on the work side, we don't have this. So for me, it's like, well, I don't know how much I'm going to be using Gemini Spark in my personal Gmail, right? I just use it more for testing and for certain purposes in Gemini that I don't have access to in my workspace account. So why is it useful? Well, it works inside of the Google apps, right?
Starting point is 00:23:35 So any app that you give Spark access to inside of the Gemini workspace, it can essentially go and, you know, monitor for updates in those apps, you know, push updates from, you know, docs to slides or from your Gmail to your calendar, right? So this also works with your personal intelligence, that new feature from Google. there's a remote browser and computer with code execution, chats, canvas, all those different things. So an example that Google gave is, you know, saying Google Doc, the task is to plan and manage your business trip coming up. It will look through your, you know, your Gmail, your calendar, and your schedule, and it will tell you, oh, your flight is delayed or, you know, a travel booking skill, you know, plus a Gmail writing skill. and rebook the room and send a confirmation or something like that.
Starting point is 00:24:32 So Spark can now open, read and edit Google Docs directly with that expanding to spreadsheets and presentations, including reading comments and editing shared team docs. So I think Gemini Spark was previewed a super long time ago. And at the time, I was personally very excited. But it took obviously many months for it to actually roll out. And it's still, at least according to my research, not available for every single paid workspace plan. So obviously, this is great if you run Gemini in your business. But not all business paid accounts have access to this.
Starting point is 00:25:15 All pro accounts on the personal Gmail side do. So if I'm being honest, six months ago, if this rolled out to my workspace account, I would be talking about Gemini Spark every single hour because it's great. But right now, you can do most of these things anyways with Codex or ClaudeCode desktop. So that's why I'm not like going to be, you know, writing about Gemini Spark. You know, it's kind of interesting that in a lot of cases, Open AI, Anthropic. And the crazy thing is even sometimes Microsoft offer some of these features and functionalities before Google does with their own products. right so even Microsoft rolled out a essentially an open claw ask integration and you can connect with
Starting point is 00:26:03 Google so although it's exciting technically Google a little bit behind even in their own backyard although I do know that millions of people will like this update and that's why I decided to cover it on the show all right two more and we are getting to our big one but before we talk about how chat GPT can now be your Jarvis. We have to talk about the, this is the one where Anthropic took a page out of the Codex playbook and made it a little bit better. That is the new Record A Skill feature
Starting point is 00:26:38 that just shipped this week in Claude Co-Work. So this is essentially you can screen record yourself doing a task once and then Claude analyzes the recording and turns it into a reusable rerunable, skill. So Claude captures screen activity, cursor movements, keystrokes, and the new addition that we don't have just yet in codex, at least not by default, is your voice narration. So then it processes all of that together in the recording, and it turns it into a structured skill that saves it to your library. So who has access? It is rolled out to all paid plans now. So free users can still use existing
Starting point is 00:27:23 skills, right, but you can't obviously record new ones. So it was a little confusing to find this. So it's only available on co-work. So you have to go into the chat interface. So now obviously Claude Desktop changed their interface a little bit. It used to be the three different tabs. It was chat, co-work, and code. Now it's just chat and code.
Starting point is 00:27:48 But in chat, there's a co-work tab, right? So you have to be in chat. You have to go find your co-work. tab, you hit the plus button, and then you can find the new record of skill. So it actually took me a little bit of clicks to find it. I'm like, wait, where is this? So you do need to update this. This is desktop only.
Starting point is 00:28:04 So update your Claude desktop app. Go into chat, click the cowork tab, click the plus button. And then actually super useful. So I did a little bit of testing. It's really good. So why is it useful? Well, until now creating a skill meant either chatting with Claude and having it create one for you. You know, you can obviously create those skill MD files by hand.
Starting point is 00:28:28 But so many times, it's just like at the end, and you should be doing this, FYI, if you're not already doing this, the way I look at new chats or tasks in any, you know, large language model, especially any agentic one, when I'm done and I get an output where I'm like, this is good, you should be turning that into a skill, right? So this is instead of having to work through it or to hand hold, right, clot or code or codex or anti-gravity or anything, right? At the end, you would normally say turn this into a skill. So the difference here with the teach clod a skill is, well, you don't have to work through it. You just click start recording.
Starting point is 00:29:10 You do your work. You dictate through it. You say, here's what I'm doing, A, B, and C. I'm opening this site. Here's what I'm looking for on this page. I download it. I open it in this program, right? And it'll just record all of that.
Starting point is 00:29:23 and turn it into a skill. So like I said, technically taking a page out of the codex replay or no, record and replay skill that did this exact same thing. But the downside with codexes is by default, there is no voice narration, right? There is kind of some workarounds around that. But it is nice that in this new one for Claude Co work, it is just enabled by default. You don't have to have any workarounds to actually narrate your way through that skill. which is very helpful, obviously.
Starting point is 00:29:56 So who's going to find this valuable? Well, literally anyone. So if you're a Claude desktop user and you use Claude Co-Work, it's going to be great. I would love to see Anthropic roll this out for Claude as well. I think that would be really helpful just because seemingly, I'm guessing Claude Co-work is not kind of riding that same wave of popularity that it was in earlier. 26, just because the interface kind of got changed over a little bit. And I think more and more people are starting to use Claude code and just more of the agendic features and the normal chat section.
Starting point is 00:30:35 So you can only set this up in the cowork, which is the downside. But literally anyone that uses Claude and you're doing the same processes over and over, I mean, this is big, right? I love. I absolutely loved and use this codex skill all the time for my workflows. So I'm personally happy to see this in Claude Code. The other thing, here's the big unlock, y'all. You can do this.
Starting point is 00:31:00 And I talked about this in Codex as well. You can do this in Claude Code or Codex, go through, record the skill. And skills are shareable. And they're fairly openly supported across all different platforms. So you can just create a skill, very detailed skill as an example using this new feature from Anthropic and then use that skill inside of Google. or use that skill inside of codex or chat gbt. All right.
Starting point is 00:31:27 And our last big feature in y'all, this, this one, this is one of those instances where I'm like, okay, this is crazy after I tried this. So yeah, hasn't even been out a full 24 hours. But OpenAI went to full Jarvis mode. And they launched chat GPT voice on desktop. And it's a lot more than it sounds like, but let me go through the bullet point details first. So OpenAI launched chat GPT voice in the desktop app, powered by GBT Live, which I already referenced, that lets you talk through work and coordinate tasks across chat, work, and
Starting point is 00:32:05 codex. So per OpenAI announcement, you can control your computer and direct multiple agents running in chat GPT work or codex using just your voice with GBT Live, speaking, listening, and coordinating work simultaneously. So here is how open AI kind of describes the main benefits. So it says you can work across projects. You can start a new task, coordinate work in progress, and pick up where you left off without micromanaging.
Starting point is 00:32:35 You can work across your desktop. You can use your files, apps, and connected tools from Slack and GitHub to Notion and beyond. And you can explore, plan and learn. You can talk through a poll request, ask questions about a code base, or learn a new topic through natural back and forth. So who has access? Well, if you are on any paid plan
Starting point is 00:32:57 and you're using the desktop version of chat GPT work slash codex, it's the same thing. It is out now. And it works also with chat GBT's remote feature, which is pretty cool because you don't even need to be in front of your computer. You can just have your phone, right? I have one right. I'm traveling right now.
Starting point is 00:33:19 I'm not in Chicago, but I have essentially a home Mac studio that never gets shut off and I can be controlling it right now and just with my voice and it can be doing literally anything and I can be watching and seeing it go in real time. So this just kind of brings that GBT live mode that was announced about two weeks ago to the desktop. So before, that was only available via mobile.
Starting point is 00:33:46 So it doesn't just bring it to the desktop. This is the first time where I'm like, Like, wait, you technically don't even need to type or use your mouse anymore. Then, yes, I haven't been able to run it through many hours of demos. But the possibilities for this one are absolutely bonkers, y'all. And it is so good. So let me talk about a couple of things. Number one, it works with Chad Shoebtee's remote feature, which is great.
Starting point is 00:34:15 That's just on your mobile app. You can click remote. and you can technically control any computer that you have connected to that chat gbt account. So that's cool. But here's the other thing, app shots. All right. So what app shots are if you're a power, you know, a codex or chat GPT work user, the default,
Starting point is 00:34:35 you know, is the two command keys. And it takes essentially a screenshot, but not just a screenshot, but every single thing, it brings into context that you have in that program. So as an example, right now on my. screen, I have up some notes, some bullet points that I wanted to go over, right? But it's a very long list. So obviously what's on my screen is probably only like 5% of what's in that document. So if I use the app shots, it's going to obviously take a screenshot of what is actually on my screen, but it brings in the context of everything else in that program that I have up that is not
Starting point is 00:35:12 even on the screen. So the cool thing is with this new, you know, I kind of wish we could have got a fun name for it. I know they can't use Jarvis. I'm just using Jarvis, right? Like, you can have this app shot feature and functionality while you're using chat GPT, live voice on desktop. So in my kind of playing around, I had, you know, six or seven different programs open on my computer, hands free, not typing and just telling chat GPT work or codex what to do. Open this file. Okay, great. Can you find this change? that I made, it's probably somewhere in the middle. And it can go, it can find that using the app shots, right?
Starting point is 00:35:53 Bringing in all that other context without having me to always tell it, oh, somewhere in this document, you know, I, I think that I outlined something about that, that certain KPI, right? It's all there. And it's absolutely bonkers. So, you know, why we keep saying Jarvis is this is literally the scene out of Iron Man, right? When Tony, Tony Stark is at his computer, just saying, hey, Jarvis, you know, pull up this, go do this, right? Now we actually have this for the first time and it works. So not just having the remote is cool. The app shots is kind of that that hack. But what I'm really excited about
Starting point is 00:36:29 is for the future of this. So obviously, Open AI's computer use with the new GPD 5.6 model took a huge leap forward. But the thing I'm actually excited about is when we get that super blazing fast Cerevra's integration. So, you know, ObedAI did say that it was coming in July. So when we can bring this with this ridiculously fast version of GPD 5.6, what you are able to accomplish in front of your computer without even sitting down is amazing. So I don't know, maybe this is going to bring back the popularity of having like a treadmill desk or something like that.
Starting point is 00:37:12 or just being able to go touch grass. I'm like, okay, is this going to also, you know, the XR glasses and all these other things? I'm actually geeking out and excited about this one because I don't know. Like, I don't like sitting in front of a screen all day. I like getting up and walking them around a lot, right? And obviously you can have voice dictation apps, but that's different. So this is just a completely different way to work, a different way to interface with smart AI. It's one I'm super excited about.
Starting point is 00:37:40 So this is useful to anyone. If you're using the new chat GPT work desktop app, which is essentially the developer, the non-technical version of codex, this is in theory a way to completely change how you interface with the computer, not just how you use chat GPUT and how you work, but literally how you interface with the device. So I'm geeked out about that one. I am too, or I hope you are too. Maybe I'll do a Wednesday demo of that one.
Starting point is 00:38:09 So that's a wrap on all the things that are new this week. Let me give you a very quick update of our seven AI features you can use today if you have paid plans. So chat GPT health is out and launched. Claude has some nice voice mode upgrades. You no longer have to talk to the dumber haiku model. You can talk to the smart opus and sonnet. We have the new Microsoft MAI image 2.5 Pro. Google released Gemini 3.6 flash in Gemini 3.5 Flash light.
Starting point is 00:38:44 We got the new Gemini Spark. Google's always on 24-7 AI agent that is available now to pro users. Anthropics, very useful. Claude, record a skill, including that voice, being able to narrate your skill, which is great. And then last but not least, open AI went to full Jarvis mode with chat GPT voice on desktop. I hope this was helpful. If so, do me a favor. If you're not already,
Starting point is 00:39:09 please subscribe to the podcast on Spotify or Apple. And then go to Your EverydayAI.com. We're going to be recapping the highlights from today's show, anything you missed, and everything else in today's newsletter. So thank you for tuning in. We'll see you back Monday and Every Day for more Everyday AI. Thanks, y'all. And that's a wrap for today's edition of Everyday AI.
Starting point is 00:39:31 Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rating. It helps keep us going. For a little more AI magic, visit your everyday AI.com and sign up to our daily newsletter so you don't get left behind. Go break some barriers and we'll see you next time.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.