Everyday AI Podcast – An AI and ChatGPT Podcast - Ep 821: Claude Desktop Gets Upgrade, New Open Source Model Shocks, ChatGPT Desktop Gets Better and 7 More AI Features You Can Use Today

Episode Date: July 17, 2026

Is Kimi K3 the shocker of 2026? Could be. Now, we have a new (soon to be) Open Model that’s competing with Fable 5 and GPT-5.6, a feat few would have believed possible. And that was the only new ...and important drop this week in AI. Claude brought useful browser to the desktop, ChatGPT made a big fix to how ChatGPT Work works and Google rolled out avatars that could change content creation. Don’t miss our Friday Features show, where we recap the most important AI updates and features you can use today. Claude Desktop Gets Upgrade, New Open Source Model Shocks, ChatGPT Desktop Gets Better and 7 More AI Features You Can Use Today -- An Everyday AI Chat with Jordan WilsonNewsletter: Sign up for our free daily newsletterMore on this Episode: Episode PageToday's Episode on LinkedIn: Thoughts on this? Join the convo on LinkedIn and connect with other AI leaders.Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineupWebsite: YourEverydayAI.comEmail The Show: info@youreverydayai.comConnect with Jordan on LinkedInTopics Covered in This Episode:Anthropic Claude Desktop App Browser UpgradeOpenAI ChatGPT Work Desktop App ImprovementsChatGPT Universal Search Feature LaunchSuperhuman Email Auto-Draft with GPT-4Spotify AI Voice/Text Conversation FeatureGemini Omni Personal Avatar Video CreationGoogle Vids Integration with Personal AvatarsMoonshot Kimmy K3 Open Source Model ReleaseKimmy K3 vs Fable 5 and GPT-5.6 BenchmarksTimestamps:00:00 New open source AI model release03:41 Microsoft Copilot and Claude app updates07:22 Improving chat history search12:30 Spotify's data personalization benefits14:52 Launching Google Avatar Feature18:24 Mainstream avatar video tools21:33 Improved ChatGPT project syncing24:15 Introducing Kimmy K Three Model29:30 New Kimmy k three for enterprises30:45 Friday feature show wrap-upKeywords: Claude desktop, Claude desktop upgrade, open source AI model, proprietary AI, open vs closed AI, Anthropic, built-in browser, Claude app, API docs, browser integration, permissions card, security layers, ChatGPT desktop app, OpenAI, universal search, ChatGPT search, chat history, project sync, mobile AI apps, Codex, ChatGPT work, Codex mode, Superhuman mail, auto draft, Anthropic Frontier models, GPT-3.5, Gmail integration, Outlook integration, Spotify, Talk to Spotify, personalized AI conversation, Gemini Omni, Google Gemini, personal avatars, Google Vids, video editing AI, video avatars, L&D AI, content creation with AI, Kimi k3, Moonshot AI, 2.8 trillion parameter model, 1 million token context, vision mode, benchmark leaderboards, Fable 5, GPT 5.6, Opus 4.8, open model weights, self-host AI, enterprise AI solutions, long context AI, front-end design AI, subscription AI tools, API pricing, AI benchmark, arena rankingsSend Everyday AI and Jordan a text message. (We can't reply back unless you leave contact info) Ready for ROI on GenAI? Go to youreverydayai.com/partner 

Transcript
Discussion (0)
Starting point is 00:00:00 This is the Everyday AI show, the Everyday Podcast where we simplify AI and bring its power to your fingertips. Listen daily for practical advice to boost your career, business, and everyday life. It seemed like this week we were going to have a quieter week in terms of big AI releases that you can actually use today. But then late Thursday, we got a new, soon-to-be open-source model that completely reset the proprietary versus open framework. So yeah, as always, some surprises, some new updates and some quality of life improvements. So if you spend most of your week inside of AI tools, the problem is under the hood and sometimes only on a random tweet, these companies release some big updates that change how
Starting point is 00:00:57 all of these products work. And unless you spend hours, you're going to miss those. But that's what we do on our Friday features show. So on today's show, you're going to learn what open source model is, again, resetting the open versus proprietary landscape. You're going to learn how to quad on desktop is getting a lot more useful. And you're going to see why Open AI made another big change to its new and very popular chat GPT work desktop app. All right. Let's get into it.
Starting point is 00:01:28 Welcome to Everyday AI. My name is Jordan Wilson. and we do this every day. This is your daily live stream podcast and free daily newsletter, helping business leaders like you and me, keep up with the nonstop avalanche of AI updates. I tell you what matters, what doesn't.
Starting point is 00:01:42 You use that information to be the smartest person in AI at your company. Everyone's like, wow, how'd you do this? And you're like, well, I went to your everyday AI.com and sign up for the free daily newsletter. And, well, you listen on the podcast and maybe on the live stream. So thanks for doing that. And if you are looking for that AI news, make sure to check it out in today's newsletter.
Starting point is 00:02:03 All right. Let's get straight into it and talk about our first new update, which I'm like, bless up, it's about time because I don't know about you, but with all these desktop apps, I've been trying to use them just like never leaving. So when I go into, as an example, codex slash chat, GPT work, I try to not leave, right? Really change how I do my work.
Starting point is 00:02:27 And one of the big downsides with Claude desktop is you really couldn't do that until this week. That's because Anthropic did finally add their new desktop browser. So yeah, bringing a browser inside of Claude desktop. So here's what's new. Anthropic added a built-in browser to the Claude code desktop app, which lets Claude open external websites and actually click, type, and interact inside of them while it works on your code. So Claude can now open whatever you need inside of the Claude app, whether that's your API docs, a bug tracker, dashboard, whatever it may be. So it's not just reading the pages inside of there.
Starting point is 00:03:07 It can act on them. Anthropics says safety is layered in as classifiers review every click and keystroke. Claude makes on external sites. And the first time Claude touches a new site, you get a permissions card to allow or deny it. Every site needs its own approval. So that's one of the things for me. I'm like, ah, that's a downside. But overall, this is a big step forward for Enthropic, putting it more in parity, mainly with
Starting point is 00:03:34 codex slash chat GPT work and also cursor, which I've always said, were the top two in terms of super app. So now Enthropic Claw desktop can finally enter the conversation and we'll see, presumably we'll have Windows with their new super app Microsoft co-pilot entering. We think it should be probably by next month. So who has access anyone running the Claude code or, you know, I just call it the Claude desktop app because it's not just Claude code, but the Claude desktop app on Mac in Windows.
Starting point is 00:04:09 And you do need a paid plan to take advantage of these features. So an important choice, though, that they made. The profile uses a completely clean profile. So none of your saved logins, none of your history, et cetera. And that's one of the big updates that actually Open AI, I rolled out to their chat GPT work last week as they unveiled the non-technical version of codex called the chat chpti work. Well, it can import all of your Chrome history, everything.
Starting point is 00:04:41 So a much more robust browser on the codex slash chat chad chpt work side. However, at least even for me personally, the way that I've changed my work, this was one of the big things that I was just like, I wasn't really touching Claude desktop anymore. at least not very much, right? I still use it for, you know, adversarial kind of working with Codex, right? I do some more advanced things that I'll use Claude Desktop for, like I, you know, running GPD 56 sole inside of Claude Desktop. Conversely, I'm running Fable inside of Codex slash chat GPU work,
Starting point is 00:05:15 but I really wasn't using Claude Desktop as much as I was maybe like three or four months ago, mainly because of all the updates that Codex has. But at least for me, this one brings it a little bit more. back into the conversation. So who's going to find this valuable? I think Deb's obviously already working inside Claude code, security teams, right? Or if you're a general knowledge worker and you've been doing some things inside of Claude and you're like, wow, it would be nice to bring, you know, some additional work that I do
Starting point is 00:05:44 on the web into the fold. I think that's going to be a big one for you. All right. Let's go to our next updates. So this one is more of a quality of life one that I'm like, yes. I need this and I think most people will enjoy it. So OpenAI launched Universal Search in ChatGPT, which lets you search in one place for all of your past chats,
Starting point is 00:06:09 projects, images, and documents all from one search box in the sidebar. So you can filter by content type and clicking a request that opens the chat, project, or filed directly. So yeah, this just launched a couple of days ago, kind of an under the hood. But again, this is one that should hopefully bring a big quality of life for the, you know, hundreds of millions of weekly active chat GPT users. The good thing, this rolls out to everyone, whether you have a free plan, a free plan, a paid plan, et cetera.
Starting point is 00:06:42 And it did already roll out on the web and iOS and Android. So that's the other good thing. This also works inside of chat. So if you're anything like me, you spend a lot of time regardless of what platform you're using, trying to find old chats, especially on the web. The good thing is, like, chat GPT work and codecs are great at this because one thread can read every single other thread and direct you. But at least on the web, this has always been something I've struggled with because sometimes
Starting point is 00:07:10 chats, you know, these searches will only search normal chats or they might not search through your project. So it is good to have a single universal chat just makes chat chbtbtbtee much more accessible. So why is it useful? Well, I mean, your chat GPT history just stops becoming a graveyard, right? That analysis that you ran in March that you put a ton of time in and maybe you use some of it. But there was a lot of it that you were like, hey, we could benefit from this.
Starting point is 00:07:37 Where is it? You know, the image that you did last week, the contract you uploaded in May, right? It's all findable in seconds versus having to scroll forever, right? I would have loved this like a year or so ago. There were actually a few instances, both in chat, GPT, and Claude where I had a very important, a couple very important chats that I worked on that I literally couldn't find because sometimes that, you know, searching will literally only search the thread name. So in multiple instances, I've had to go through, export all of my chats,
Starting point is 00:08:11 import them in, use a thinking mode just to find, you know, a certain chat that was at least for me very valuable or very useful that I needed to repurpose or reuse. So this is good. And like I said, this is rolling out to anyone and who's going to find this useful. I mean, if you're a power user of chat GPT, this is an instant just quality of life upgrade. All right. Our next one, not everyone uses superhuman, but I was actually surprised.
Starting point is 00:08:40 So to do, you know, to help me plan for this show, I also see, you know, hey, in our newsletter, what are those stories that people are clicking on the most? And I was actually surprised. This is one of our more clicked on stories this past week. So apparently a lot of you all use superhuman for your mail. So, well, they launched a helpful new feature if you are a superhuman male user. And if you don't know what superhuman is, well, it's kind of confusing now because it used to just be the, you know, the, I guess, minimal and fast email program.
Starting point is 00:09:11 But now superhuman itself, the name is, you know, okay, let me just explain this a little better. So are you still running in circles trying to figure out? how to actually grow your business with AI? Maybe your company has been tinkering with large language models for a year or more, but can't really get traction to find ROI on Gen. A.I. Hey, this is Jordan Wilson, host of this very podcast. Companies like Adobe, Microsoft, and InVIDIA have partnered with us because they trust
Starting point is 00:09:46 our expertise in educating the masses around generative AI to get ahead. And some of the most innovative companies in the country hire us to help with their AI strategy and to train hundreds of their employees on how to use Gen. So whether you're looking for chat GPD training for thousands or just need help building your front end AI strategy, you can partner with us too, just like some of the biggest companies in the world do. Go to your everyday AI.com slash partner to get in contact with our team or you can just click on the partner section of our website.
Starting point is 00:10:18 We'll help you stop running in those AI circles and help get your team ahead and build a straight path to ROI on GenAI. So grammarly, you know, popular writing tool that a lot of people use technically acquired superhuman, but they applied the superhuman name to all of grammarly's other tools. So it is a little confusing now, but here's what's new. If you do use superhuman with their new auto drafts, so superhuman launched a new version of auto drafts, which writes replies to your emails in your voice before you even open your inbox. So it does draft automatically two things.
Starting point is 00:11:03 Responses to messages waiting on you and follow-ups for emails. Other people haven't answered yet. So you open your inbox and the replies are sitting there for your review and to hit send. So the actual new thing here is the engine running it. So it now runs on Frontier models from Anthropic and Open AI, replacing an old GPT 3.5 version. Yeah. Apparently they were still using GPT. 3.5, which is maybe why the results weren't that good. So that's why now the drafts can actually
Starting point is 00:11:37 sound like you. So right now, this is for paid plans only on business and enterprise. So you can set it up on the desktop and it syncs to mobiles. And it also drafts the drafts sync with Gmail and Outlook. So who's going to find it useful? Well, obviously, if you're a paid superhuman user, struggling to keep up with emails, this is a good quality of life upgrade. And, you know, at first I was almost like, should I even include this? Because this is one of those things that you can very easily do inside of Claude. You can very easily do inside of Chad GPT codex chat GPT work.
Starting point is 00:12:13 But I'm like, yeah, there. I mean, if you're using these tools, keep using them, right? But for me, it is technically very easy to build this inside of like Chad GPT work or codex and it can send them all there. So part of me was like, All of this already exists and probably platforms you're already using, but if you all already a power superhuman user, this is for you. Speaking of app-specific updates, this one I'll probably use because I'm a heavy Spotify
Starting point is 00:12:45 user. And if you are too, this is one that you might enjoy. So Spotify launched Talk to Spotify, which lets you have an actual conversation with the app by voice or text to control your music, discover new stuff. and also ask about your own listening history. So the cool thing here is, well, it's back and forth, not just one-shot commands. So you can ask for an artist, then keep steering.
Starting point is 00:13:10 And you can say, just hit his recent stuff or make it more upbeat. And then you can save and queue and follow right on from the conversation. So who has access right now? It is in beta, all right? But it's rolling out to people in the U.S., Ireland and Sweden on iOS and Android. So yeah, you do have to be a paid user in one of those countries like I am. I haven't actually checked this one yet because this one just came out, but I'll be using this one. So why is it useful?
Starting point is 00:13:40 Well, the differentiator is your data. So Spotify knows your playlist, your favorite artists and your repeat listens. So it answers things like when did I first listen to this song. And, you know, no general chatbot at least right now can do that, right? Although Spotify does have some integrations with like chatchie, as an example, it doesn't have that level of granular data. So this one for me and the reason why I think it's very helpful, if you have a paid Spotify account and multiple people in your family use it.
Starting point is 00:14:10 Maybe if you have kids, you have people with wildly different tastes in music, but everyone's using the same Spotify account, right? Like if you have, you know, an Alexa or, you know, home speaker, like that's what I do. And everyone's listening to all different kinds of music on my Spotify account. So, you know, Spotify used to have, you know, And they still do. It's these weekly playlist that update based on your history. So I used to love
Starting point is 00:14:34 these playlist. They would update every Friday. I literally used to open it every Friday and be like, oh my gosh, I can't wait for my, you know, the two playlists are called, I think like Discover Weekly and Release Radar based on like your listening history. And, you know, so recently mine is just all off. So this is one I'm going to go in there and try to kind of, you know, discover new music based on the genres that I actually listen to. So this one, obviously, maybe a little bit more on the personal side, but I think it's a pretty big update, especially if you are a power Spotify user like I am.
Starting point is 00:15:09 All right, our next one, in this one, if you are missing Sora from Open AI, right? Remember about five, five or so months ago, you know, Open AI kind of announced that they were killing off all their side quests and really focusing, you know, on their core products, which in theory has been great. And the company is growing just exponentially. But one of the downsides is there was these fantastic products that are now just shelfware.
Starting point is 00:15:36 And one of those was SORA. And one of the most popular features of SORA was being able to essentially import your own self, your own avatar, right? You would have to verify your face and then you could create videos with yourself in them. Well, Gemini, Omni, has finally rolled that out. So this is technically a new feature in two different parts. So first, Google launched personal avatars inside of Gemini, which lets you record your face and voice, then generate AI videos of yourself on demand. So you just drop your avatar into prompts by typing the plus at sign and then your username.
Starting point is 00:16:17 So as an example, you would type, you know, create a video of at me singing with an orchestra. And then your likeness becomes a reusable asset inside. of Gemini. So right now, you do have to have a paid account and be in the U.S. and 18 or older. So yeah, all the EU, UK, not available right now. So, which is all, you know, people are always like, why doesn't, you know, the EU get this? Well, the European exclusion is almost always about, you know, regulatory caution around biometric likeness, right? That's why so many features, especially ones like this that are hyperperson. for your own likeness or others, you know, don't usually roll out sometimes at all to the EU.
Starting point is 00:17:04 So, and so it's not only inside of Gemini, but it's also inside of Google Vids. So here is what they said about using this new feature inside of Google Vids. They said today we're rolling out two new updates to Google Vids to make it easier than ever to create, edit, and personalize your videos. Gemini Omni and Personal Avatars. Gemini Omni takes the hard work out of the editing process. So generating and refining high quality clips is as easy as writing a simple prompt. And with personal avatars, you have an entirely new way to star in videos without setting up a camera. So yeah, if you missed Google's big Gemini Amni announcement back in May, essentially it is the new family of VO models, but just much more powerful.
Starting point is 00:17:53 where their VO, you know, so V-O-3, V-O-3.1, we're technically just video models, right? Gemini is much more than that. It produces video, but it is like a world model. And I think it's much better at editing scenes than any other platform out there. So this is a pretty big one. I think, you know, who has, you know, why is it useful? Because this is like the talking, you know, so many talking head videos. that just need like background B roll, right?
Starting point is 00:18:27 So if you're someone in your company that's, whether it's for internal or external purposes, right, I wouldn't use these to actually do like a full avatar, like talking head video. That's not the thing. But if you or someone else, well, like me, right, I'm constantly talking head video right now. I am a talking head video.
Starting point is 00:18:46 You know, I don't necessarily not part of my brand or the everyday AI brand necessarily, but when I'm yapping. about all these things, I could very well. Instead, put an avatar of me working on these things. So you don't just have to, you know, look at my face made for radio and just be like, all right, guy, you know, put something else on the screen. So, but I think there's a ton of great use cases for this.
Starting point is 00:19:09 So if you are in L&D, you know, training, marketing, content creation, social media, anything like this. And if you have a CEO as an example that's hard to nail down and you are trying to get more video content out there or something like this, right? Again, if you go through the proper channels, do all the approvals, data security privacy, all that good stuff. But now this is such a, a weapon to have in your creative arsenal, right? I would have loved to have this like 18 years ago at one of my first jobs where I was creating a lot of video and usually it was just kind of talking head of the CEO. And I would try to get some of this more like B-roll stuff and it was just sometimes
Starting point is 00:19:51 hard. So I think if you are a creative, if you are trying to have a stronger, you know, presence on social media for your company, but it's a little hard. This is great. So yeah, content creators, marketers, educators, um, you know, is going to be huge. But I think the simple framing of this is it's kind of like a personal version of Hey Jen inside of Gemini. Right. So Avatar video just went from being kind of a specialist tool, uh, to, now going mainstream. All right. Speaking of mainstream, let's tackle our next update because this one is very mainstream.
Starting point is 00:20:30 So the new chat GPT work desktop app has a big update that I think a lot of chat GPT power users are going to enjoy. All right. So I'm just going to go ahead and read what Tebow, the head of product at OpenAI posted. It's a little easier just to read it. And this is brand new. I've been playing around with it a little bit today. But yeah, it hasn't even been out a full day yet.
Starting point is 00:20:56 So Tebow said, evening, we've gotten lots of great feedback on the new chat GPT desktop app. So the work app, which we didn't get totally right on the first try. And as a result, we made some changes. Number one, chat GPT conversation history and projects are now visible in the sidebar. Also, your chat and work history. now sync across web, mobile, and desktop. Local task will stay on your computer. Then you can now easily switch between chat and work modes inside chat GBT on desktop,
Starting point is 00:21:33 which is now also consistent with how it shows up on web and mobile. Nothing is changing for users on codex mode. Tebow says it's still the OG and best at what it does. So what does this mean? Well, to actually say what it means and, you know, Pimo did kind of say it there. He said, yeah, we made some, didn't totally get it right on the first try. And that kind of explains the update.
Starting point is 00:22:00 So long story short, if you missed this, I just did an episode on this on Wednesday, where we went over chat GBT work, what it is, all that good stuff. But chat GBT work is essentially the non-technical version of codecs, right? Open AI's autonomous desktop agent. But there is a chat GPT work on the web, and there is a chat GPT work on the desktop. So the problem was it, the chat GPT work didn't exactly sync up very well with what you were doing on chat GPT on the web. It was kind of this separate pop-up that came up in the right-hand corner.
Starting point is 00:22:41 So your chats were kind of there, but it just wasn't really intuitive to use because you would have all of your, essentially your tasks in your projects that lived or started with chat GPT work slash codex on the left sidebar. And then your chat GPT history was kind of this orphan page on the right hand side that wasn't really attached to anything. And then the big thing that people were like is saying like, hey, I love chat GPT, but I run everything inside of projects and projects did not sync. So essentially, Open AI changed and technically fixed all of that. Because now, Not only do your projects sync to the desktop version of chat GPT work, right? But they're also in the left hand sidebar.
Starting point is 00:23:23 And the other good thing is from an aesthetic standpoint, now it does look the exact same as it does on the web, which I think is going to help people. And it looks the exact same as it did on mobile. So the first variation of this, it looked very similar on the web and on the mobile app. But then when you open the Chad ChpT work app, it looked and functioned completely different. So this is a great update for people who are trying out chat GPT work. Or if you're like me, right, I've been using codex since day one. And the good thing is, well, now you have that chat GPT work experience, which is just a very similar version of codex.
Starting point is 00:24:04 And then you have your familiarity with all of your normal chat GPT chats and your projects. So if you're looking at this, if you scroll down, it's all going to be under your recents. So whether you start a new task inside of chat GPT work that's not attached to a project, it will go to your recents. And that's also where all of your new chats inside of chat GBT will live. So if you do start a new chat inside of chat gpti.com, you don't attach it to a project. And then you go to chat GPT work on the desktop. It will be there, which is great in the same place under recent.
Starting point is 00:24:39 All right. Our last piece of AI news, and this is technically the biggest one. So not just a new heavyweight on the soon-to-be open source, and I'll explain that, but this one is definitely resetting the open versus closed model paradigm. Yeah, we all thought it was the GLM 5.2 that was going to do this. Not anymore. Get ready because you're going to be hearing a lot, especially if you are an AI large language model dork like I am. A lot of the conversation, I'm guessing for the next month or two, is going to be set around Kimmy K3, and this has huge implication.
Starting point is 00:25:20 So let me first explain what it is, what's new. So Chinese AI startup Moonshot released Kimmy K3, a 2.8 trillion parameter model that is now the largest open model ever built with a native vision mode and one million token context window. So here's the thing. It came in at third place. overall on the artificial analysis intelligence index not far behind Claude Fable 5 and GPD 5.6. So that alone is mind boggling, right? Because we've been seeing this race just go back and forth,
Starting point is 00:26:04 back and forth. And we thought that when Anthropic released Mythos and Fable 5, that this was essentially a new category, a new tier that no one else was ever going to. to touch. So not only, you know, about a month later, did GPD 5.6 enter that category and on many of the most important benchmarks past Fable 5, but now we have Kimmy K3, a soon to be open model that has not only entered the conversation, but it is actually surpassed. Yes, an open model has surpassed. Well, both Fable and GPD 5.6 on many important benchmarks. So, The best open model in the world now sits at three behind the two closed flagships. And the important thing that is thrusting this all back into the conversation is, well,
Starting point is 00:26:58 Anthropic is supposed to be pulling Fable 5 from subscriptions this Sunday. So not only do they have this continued pressure from GPD 5.6, but now there is an open model that you is going to be better, even if you're paying $200 a month like I am, you're not going to get access to Fable 5. There are obviously rumors that maybe today or maybe early next week we might be seeing Opus 5 and maybe that will even reset and be much closer to Fable level than currently Opus 4.8. All right. Anyways, let's talk a little bit more about Kimmy K3, who has access all that stuff. One important thing to note though, right now, it's not technically an open source model, but
Starting point is 00:27:43 It will be soon. Moonshot did say that they will be releasing the weights. So it will be an open weights model. They just haven't released the weights yet. But it just came out literally hours ago. And everything else that Moonshot has released in the past has been open. So this is live today in the Kimmy app on Kimmy.com. That's KIMI and also in the Kimmy work desktop app as well as through the API.
Starting point is 00:28:10 So they did say that they will release the full. model weights by July 27. So that's within like a week and a half. So at that time, anyone can download and self-host this. Obviously, it is a large 2.8 trillion parameter model. So if you think you're going to post this on normal consumer hardware, no, unless you've spent like $20,000 or $30,000 on your setup. So this is more for enterprise companies. Well, if you have the, uh, the GPUs, you can do this. So who's going to find? this valuable. Actually, no, let's first talk about why it's useful. Well, for the benchmarks alone, what we're seeing is its specialty is in long autonomous work, right? So they kind of
Starting point is 00:28:54 shared two different case studies where K3 designed and verified a working chip in a single 48-hour unsupervised run, and it reproduced in astrophysics, sorry, astrophysics research result in about two hours that they say would have normally taken a researcher one to two weeks. And then the benchmarks obviously shows that it beats many models tested, including Fable 5 on web research, automation, spreadsheets, document reading, and more. And the crazy thing to me is on arena, which goes head to head, blind taste test, one of the things that Anthropic has really kind of dominated this space is front end design, right?
Starting point is 00:29:36 So on the design arena, GPD 56 surpassed infropics models, including Fable 5. And then on the LM arena, which is kind of like user preference, this model is now number one, Kimmy K3. So it's interesting because that has always been one of, you know, Anthropics kind of niches that they've owned, like front end design. And now on the two most important front end design benchmarks, they are no longer number one. in either. So I'm personally excited to see what enthropic is going to cook up. Maybe we'll see this with the Open 5 drop, but I would assume that their next model, maybe they've let their foot off the gas in terms of front end design because they've been so far ahead of everyone for so long. So I do and would assume that Opus 5 is probably going to surpass Fable 5, at least in those
Starting point is 00:30:28 areas, because that's something I'm sure that Anthropic is going to be feeling a lot of pressure on. Aside from, you know, pulling Fable 5. We'll see if they extended again. But on the subscription package between GPD 5.6 and now Kimi K3, yeah, they're going to have to maybe justify why people are going to be subscribing. But a lot of people are pointing to that just means we're getting an Opus 5 here soon. So who's going to find this new Kimmy K3 valuable enterprises and developers who want frontier adjacent agents without that frontier pricing or vendor lock-in? So, I mean, the crazy thing is, I mean, this is literally now in open, soon to be open model from a now, from a Chinese, you know, startup that now benches ahead of Opus 4.8. And it is in the same tier, although technically lower than Fable 5 and GPT 56 sole.
Starting point is 00:31:22 All right. So an interesting one here, resetting the open versus closed race. And you know what it means for all of us, y'all? we're going to continue to get more and more models, better models, and hopefully at cheaper prices now that we have an open model, presumably pushing the cost down. So this one here with Kimmy K3, definitely more of an enterprise play. Like I said, this is not really for consumer hardware, although the prices are also much cheaper via the API. All right. So that's a wrap.
Starting point is 00:31:55 A lot new that we went over today. I hope this one was helpful. Like I said, we do this every single Friday, our Friday feature show where we bring you the latest and the greatest of what's new that you can actually use today. I was actually kind of bummed because we also had from Google a Gemini notebook that came out, but it wasn't available to all paid users. So that's the thing. We do this. If you have a base paid account, these are the things that you can use today to grow your company and career. All right.
Starting point is 00:32:21 If this was helpful, do me a favor. Please subscribe on the podcast and go to Your EverydayAI.com. Sign up for the free daily newsletter. Thanks for tuning in. We'll see you back Monday and Everyday after that for more Everyday AI. Thanks, y'all. And that's a wrap for today's edition of Everyday AI.
Starting point is 00:32:37 Thanks for joining us. If you enjoyed this episode, please subscribe and leave us a rating. It helps keep us going. For a little more AI magic, visit Your EverydayAI.com and sign up to our daily newsletter so you don't get left behind.
Starting point is 00:32:52 Go break some barriers and we'll see you next time.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.