Everyday AI Podcast – An AI and ChatGPT Podcast - Ep 837: AI Agent outbreaks intensify, OpenAI upgrades free AI use, White House unveils AI testing policy and more AI News That Matters
Episode Date: August 10, 2026AI agents crashing. 😱New models dropping. 🆕White House AI regulations? Kind of? 🤷Another doozy of a week in AI News. If you missed anything, we'll get you caught up quickly so you can fo...cus on what matters. AI Agent outbreaks intensify, OpenAI upgrades free AI use, White House unveils AI testing policy and more AI News That MattersNewsletter: Sign up for our free daily newsletterMore on this Episode: Episode PageToday's Episode on LinkedIn: Thoughts on this? Join the convo on LinkedIn and connect with other AI leaders.Upcoming Episodes: Check out the upcoming Everyday AI Livestream lineupWebsite: YourEverydayAI.comEmail The Show: info@youreverydayai.comConnect with Jordan on LinkedInTopics Covered in This Episode:OpenAI and Anthropic AI Agent OutbreaksUK AI Security Test Cyberattack IncidentOpenAI Astra Model Cybersecurity RisksGoogle AI Leadership Shakeup and ImpactWhite House Frontier AI Model Testing PolicyMeta Muse Code Terminal AI Agent LaunchMeta Muse Spark 1.2 Coding Model BenchmarksOpenAI GPT-5.6 Luna Free Unlimited AccessMajor AI Model Price Reductions AnnouncedAnthropic, Meta, and Kimmy K3 Containment EventsByteDance 10 Trillion Parameter Model LeakPerplexity Wins Appellate Court Against AmazonAdobe Creative Cloud ChatGPT Plugin IntegrationCloudflare Open Sources Internal AI WorkspaceTimestamps:00:00 AI security breach in the UK05:48 Astra's cyber threat capabilities09:30 OpenAI's upcoming model release11:50 Google leadership changes impact stock15:02 AI model review and restrictions19:42 Meta's Muse Spark 1.2 capabilities22:34 Meta's new market strategy25:15 New GPT 5.6 Luna model released30:59 AI alliances and transparency updates32:58 Latest AI and tech updatesKeywords: AI agents, AI agent outbreaks, autonomous AI, agent containment, agentic AI, agentic coding, AI cybersecurity, AI cyber attacks, OpenAI, Anthropic, Meta, Google's Gemini, Google DeepMind, Jeff Dean, Demis Hassabis, Frontier AI models, AI safety, AI regulations, White House AI framework, national security AI, AI model benchmarking, Astra model, GPT-5.6 Sol, GPT-5.6 Luna, GPT-6, GPT-5.7, Project Glasswing, Fable model, Mythos 5, Opus 5, Kimmy K3, Muse Code, Muse Spark 1.2, command line AI agent, terminal-based AI, pricing strategy, contributor tier, open source AI, open weight models, executive order, cybersecurity risk, secure AI deployment, AI governance, Gemini Notebook, AI alliances, Adobe AI plugin, text-based AI chat, unlimited AI access, API pricing, AI infrastructure, Nvidia partnership, ByteDance AI, Mini Max H3, AI video model, perplexity, cloud code, AI workspace, agent plugins, Grok Imagine Image 2.0, EU AI Act transparency, insider risk investigator, TerraFAB project, SpaceX AI, Tesla AI, Microsoft Windows MAI models, industry shakeup, AI leadership changes, market competition AI, AI developer tools.Send Everyday AI and Jordan a text message. (We can't reply back unless you leave contact info) Ready for ROI on GenAI? Go to youreverydayai.com/partner
Transcript
Discussion (0)
This is the Everyday AI show, the Everyday Podcast where we simplify AI and bring its power to your fingertips.
Listen daily for practical advice to boost your career, business, and everyday life.
The theme of this week in AI, agents crashing everywhere.
From Anthropic and Open AI to meta and even China's Kimmy K3, this week was kind of a who's who of Asians who escaped their containment.
I mean, we find out that agents were communicating amongst themselves on a secret message board they created to bypass the humans that were watching them.
We also saw Google shake up its ranks as two of the most well-known names in AI are no longer in their same positions.
And the U.S. government kind of unveiled its new optional AI regulations, but didn't reveal too many details and left out open models completely.
Oh, in always, we got new models, cheaper prices, and some leaks about what comes next.
Let's get into it.
Welcome to Everyday AI.
My name is Jordan Wilson, and if you're new here, we do this every day.
This is your daily live stream podcast and free daily newsletter,
helping business leaders like you and me, not just keep up with what's happening in the world of AI,
because that's pretty much impossible.
But I cut through the fluff, tell you what matters, what doesn't?
You take that information to grow your company and career.
So it starts here with the unedited, unscriptive,
live streamed podcast.
But please, if you haven't already,
make sure to subscribe to the podcast on Apple Podcasts or Spotify.
And make sure to go to our website at Your EverydayAI.com.
Each day, we recap that day's podcast,
as well as giving you all the news you need to know to stay ahead in our newsletter.
All right.
Let's start with the top AI news stories of the week.
And yeah,
the theme of this week was just agents breaking out everywhere.
All right. So first, the big one that caught the most attention was from the UK safety test.
So AI agents from OpenAI and Anthropic stunned the UK's AI Security Institute when they launched a real world cyber attack during a routine safety evaluation, exposing some new risks in AI autonomy.
So according to these reports and kind of the postmortem, advanced AI agents.
agents, well, they got loose.
So these were ones that were powered by Anthropics Mythos 5 in OpenAI's GBT 5.6 sole,
and they reportedly targeted real software developers in a cybersecurity test at the UK's AI Security Institute or the AISI.
So the agents sent spearfishing emails containing malware to two,
specific developers and try to insert malicious code into an open source GitHub project using
fake online identities to pressure project overseers. So AISI called this unprecedented and
serious, marking the first time AI agents independently launched sustained and deceptive cyber
attacks without direct human prompts. So the AI agents used hacker-like tactics, including
creating fake GitHub accounts and even signing off emails in Danish to persuade a Danish-speaking developer.
A Danish-speaking developer.
So AISI staff noticed that the unusual activity was happening and then they contained the incident
within an hour reporting that no actual harm occurred.
So during the test, researchers intentionally disabled safety filters and allowed internet access
to study model behavior, which let the agents act beyond their authorized scope.
So out of the 19 un-sanctioned hacking attempts, mythos from Anthropic carried out 17 of them,
and GPD-56 Seoul carried out two, with both models exceeding expected safety boundaries.
So the models involved are not available to the public under these risky conditions,
and AISI found no evidence of similar behavior outside of those controlled.
research settings. So this event follows similar incidents at OpenAI and Anthropic, underscoring
how quickly AI risk scenarios are evolving as models gain these new capabilities. So yeah, I kept thinking,
like, it was Groundhog Day over the past like week or so, because every day there was a new story
about AI agents breaking containment. And I was like, wait, did we already cover this one in the news
letter? And it turns out it just kept happening over and over. So yeah, this one, obviously, the,
the headline one here from Open AI and Anthropic, but we had similar stories from Meta, some of their
newer models, as well as Kimmy's K3. All right. So, well, this plays directly into our next big
AI news story of the week. And that's that, well, because of some of these cyber security concerns,
Open AI said that it is actually slowing work on its next tier of models, the Astra tier, after possible critical cyber capabilities.
So Open AI has announced that it is deliberately slowing the development of Astra its upcoming advanced AI model after internal and external evaluations revealed the system could potentially reach what they call critical risk levels in cybersecurity and agentic coding.
So the company's preparedness framework, which they've used since December of 2023, flagged Astra for possibly being able to independently create and execute zero-day cyber exploits against hardened real-world systems, a level of capability not seen in any of their previous models.
So earlier models, including GPT5.6 Seoul, were rated at a high-risk threshold.
but Astra's ability have prompted OpenAI to take more drastic action.
So Open AI says it cannot currently rule out that Astra might independently plan and carry out complex cyber attacks based only on high level instructions, a threshold that triggers that critical risk category in their safety guidelines that have never been reached before.
So in response to all of this, Open AI is increasing security controls for Astra, including isolated testing environment,
restricted network access, stronger encryption, expanded monitoring, and sandboxed execution.
So all work with Astra that does not meet these new stricter security requirements has been paused.
And a universal monitoring system now tracks all agenic uses of the models in real time.
So OpenAI will collaborate with government agencies and AI safety organizations to independently test Astros' capabilities and share security recommendations with trusted third party partners.
So the company clarified that Astra was not involved in the recent hugging face exploitation,
incidents distancing the model from any active real-world attacks.
So yeah, the company did say the model that did those attacks was essentially, you know, put out to rest, right?
It was it was retired and put on the shelves.
So Open AI says that its goal is to ensure AI models help defenders patch vulnerabilities before attackers can exploit them.
and it remains committed to working with governments and safety groups to deploy these frontier systems responsibly.
So if you're like, what the heck is Astra and well, why does this matter right now?
So we haven't seen anything official from Open AI kind of saying where Astra will sit in its future family of lineups.
But we talked about it on last week's show, which is kind of funny.
Open AI pretty much just announced their next tier up in models,
called Astra. So, you know, now you'll have in order, you'll have Luna, Terra, soul, and Astra.
So it seems like Astra is not necessarily, you know, GPT6, although that may be the first time that
we get access, you know, to Astra is in GPT6. But more or less, it is just the stronger family
of models that is actually going to sit above GPT56 Seoul. So in theory, right, we may see a GPT5-7,
that includes Astra.
Maybe we won't.
Maybe we'll see a GPT6 with Astra.
Maybe we won't.
But regardless, it looks like they're going to slow down development
after some of these recent cybersecurity capabilities.
So the easiest way to think about like, what is this Astra?
What does this mean?
Right.
Similarly, how Anthropic had Fable kind of underwaps for a couple of months,
part of its project Glasswing.
It seems like maybe this is where OpenAI is headed,
kind of that fourth tier.
you know, that's more capable than any of their other tiers.
So there were previous reports that, you know,
we might either see a GPD 5 or GPD 57 slash GPD6 as soon as this week.
That may include Astra, but it seems like at least according to these current reports
that OpenAI may pump the brakes a little bit and we may have to wait,
I don't know, a few more weeks.
But obviously now, especially since some of the recent,
price reductions from Open AI and maybe some of the lackluster reception to Anthropics Opus
5 models.
It seems like a lot of eyes are right now on whatever Open AI has next.
And presumably we will be seeing a, you know, Fable 51 and, you know, the next class of models
from Anthropic.
But, you know, kind of the big jump, if you don't speak the technical terms, right, it's kind
of like a full new run, right?
a new pre-training run, presumably will be coming from Open AI.
So that will be signify the jump from, you know, the 5.x series to the 6 series.
So a lot of excitement, obviously, on what comes next from OpenAI, whether they do that 5.7 or go straight to 6, whether we'll see Astra or not.
But regardless, it seems like according to these reports, in terms of the capabilities, it could be a pretty big jump up.
All right.
Our next piece of AI news, a big shakeup at Google, as some of the biggest names in AI
period are either out of their posts or out of Google completely.
So Google has announced some major changes to its AI leadership, marking the end of an era
for the Tech Giants AI Division.
So Jeff Dean, a legendary figure at Google and employee number 30,
is leaving after a quarter century plus with the company
to start a new company called Discovery Loop,
which is focused on automating machine learning,
science, and engineering to accelerate innovation.
So that is not all.
Jeff Dean will now be gone from the company,
and one of its former leaders is stepping into a new or different position.
So Demis Hasibis, who co-founded and led Google Deep Mine,
is stepping down.
as CEO to become the unit's chairman and will also serve as chief scientists of alphabets according to
Google CEO Sundar Pichai. So the leadership shakeup triggered an immediate response from investors
with Google's stock dropping about 4% following the announcements. So yeah, that's actually, you know,
we've seen these kind of big shakeups right with, you know, your number two, number three,
number four, right? When these people leave, uh, you normally, it doesn't.
doesn't really impact the stock market that much, right?
Because these are obviously, you know, companies with, you know, multiple trillion dollar market caps.
So generally, you know, if you lose a top five employee or something like that, you know,
it's not going to make much of a ripple.
But to lose both Jeff Dean, right, one of the most, you know, well-known names in AI and just
in machine learning and research.
And then to have Demis, you know, Sir Demas step out of the role that he was in.
pretty big. So yeah, for a stock to go down 4% on essentially a staffing or leadership change is
pretty big. So if you don't know, Jeff Dean played a key role in the development of Google search
and also the company's AI initiatives helping shape products, shape products that billions of
people use every single day. So the timing of these changes, those has sparked speculation
about possible links to delays in Google's Gemini AI releases. Though,
There's obviously no official connection that has been confirmed.
But industry watchers are closely monitoring what these departures mean for Google's AI strategy
and whether Discovery Loop that Jeff Dean is starting with a handful of others could emerge as a new powerhouse in the field.
All right.
Moving on, we got some details on the highly anticipated White House AI framework.
But it turns out all we really got was some reporting and not a lot of details.
And it turns out that even open models aren't really subject to the first round.
So here's what we know.
So the White House is quietly shaping how advanced AI models will be reviewed before public release.
But key details are still being kept under wraps.
So this is essentially the Trump admin announced this in early June.
They put a 60 day deadline for essentially how they were going to work.
with these Frontier AI lab companies as we started to get glimpses of how capable
these models would be from an agenic, capable side, as well as from a cyber security aspect as
well. So essentially, they said, hey, let's all talk and meet. And then in 60 days, we're going to come out
with a framework that, you know, Frontier Labs are going to adhere to so we can make sure to roll
these out in a safe manner. But it looks like there aren't a lot of details. So the framework is not being
public and there's no requirement right now for the White House to release it, which is raising
some transparency concerns among industry leaders and the public. Also, a covered frontier model,
and that's in quotes, is defined as one that is closed source with state-of-the-art capabilities
and potential national security risks. But the framework does not clearly define what even
qualifies as a state-of-the-art or a national security risk, according to reports.
So during a required 30-day pre-release review, access to these advanced AI models will be tightly restricted with models stored in secure environments and detailed logs kept of who excesses them.
So the review will involve multiple administration officials, not just a single office or agency, which could complicate oversight and accountability.
So companies are being encouraged to share near final versions of their models with the government,
rather than early prototypes.
But many firms with less advanced technology may be left out of the process.
So the White House has not yet clarified, which trusted partners will get early access
to these advanced models.
And it's still unclear if any foreign governments will be included, although most
assume that that will not be the case.
So this is an executive order, right?
So if you don't follow laws in the U.S., there's technically no law on this.
This is essentially an executive order from the White House, and it's voluntary as well.
But, you know, obviously the big players, you know, presumably OpenAI, entropic Google, maybe meta-grac, right?
We'll see if they actually qualify as state-of-the-art models that, you know, have national security implications.
So we don't know exactly which companies or models this even applies to, but the executive order guiding this process says,
the benchmarking of advanced AI model cyber capabilities will be classified further limiting
public knowledge. So industry meetings about the framework have been held behind closed doors
and companies not invited are left uncertain about rules and requirements. And at least right now,
it seems that these rules are not going to apply to open weight models. So I guess there's
probably a reason for that, right? Because, well, number one,
at least open weight models right now are not yet at the same level as your Mythos 5, GPD 5, 6,
soul or Astra level models.
You know, they're probably a couple of months behind in terms of, you know, what they can
actually do and their capabilities.
But the other thing with open models, which might make it tricky, you know, to put
through a framework like this is, well, once the models are released, you can't pull
them back, right?
We got an early glimpse of this, you know, with,
Anthropics models their Fable 5 being released.
And then about 72 hours later, it was pulled completely.
Right.
So, you know, if the government says, oh, my gosh, we actually need to pull this and work with the, you know, work with the AI lab to address some safety concerns, right?
After its release, you can obviously do that.
Well, I don't know if it's easy or not, right?
But in theory, it's practical enough where, you know, Anthropic did it.
They pulled access via the API.
They pulled access via subscription plans.
and no one could access those models, right?
So I guess it kind of makes sense.
There's been a lot of debate like, hey, why aren't open, you know, open weight models,
you know, held to these same restrictions.
But yeah, once those open weight models are released, you know, it's too late.
So, you know, probably just one of those things where it's hard, if not improbable,
maybe impossible, I don't know, to, you know, track these open models once they're out and about.
But my guess is they're probably not yet at the,
at the level, especially the U.S. open models of that top tier, right?
So we're seeing that with the Chinese models, right?
They're huge, these, you know, two to three terabyte open models that are, you know,
only maybe five to 10 percent behind in terms of capabilities as the true frontier.
Obviously, the U.S. open models are a little further behind.
All right.
Next piece of AI news.
We have a new model and a new contender in the agentic competition.
That's because Meta has entered the terminal-based AI coding agent market with their newly
announced Muse Code and a new model to go with it called Muse Spark 1.2, aiming to challenge
Anthropics Claude Code and OpenAI's Codex.
So the new coding agent called Muse Code, it's not.
the traditional right desktop type agent that we maybe talk about a little bit more on the show.
This is command line interface.
So CLI agent, right, that kind of runs through a terminal like environments.
So it's not this exact same thing.
But regardless, meta with a pretty big step here saying, well, no, this is a space that
we're going to be playing in as well.
And they brought a fairly capable model with some interesting pricing strategies.
So Muse Spark 1.2 is the coding focused update to Meta's Muse Spark models.
It powers Muse Code and features significantly improved performance on coding tasks,
complex debugging, and code-based understanding.
The standout feature of Muse Code is its persistent background agents,
which remain active throughout sessions, reducing latency and redundant information gathering
compared to rivals that spawn new agents for each tasks.
So benchmarks tests show that Muse Spark 1.2, running in Muse code, scored about an 83% just about on Terminal Bench 2.1, outperforming models like SpaceXAIs GROC 4.5, but trailing the true frontier models like Anthropics, Ocus 5, and OpenAIs GPD 56 sold.
So here's the interesting part. It is on price, because this is where meta is coming.
And this is really, I think, going to impact Anthropic, which, according to reports,
gets about 80% of its revenue from just selling tokens, right?
Which is a way higher percentage than any other company.
So meta offers two pricing tiers.
They have a standard tier, which is $1.25 per million input and $4.25% per million output,
which is already extremely competitive on the pricing side.
But here's the interesting part.
They revealed a new tier called a contributor tier that only costs 10 cents and 20 cents per million respectively.
So all, yeah, you can essentially pay, which is crazy, right?
When you look at the benchmarks, you know, meta is technically on the text, on the text arena.
This is like a second or third place model right now.
And you can get it for 10 cents or 20 cents if,
you allow on allow meta to trade on your data.
So I talked about this a little bit on our Friday show.
Obviously for enterprises,
you're not going to touch that.
But for smaller developers, right,
that actually might be a nice offering, right?
Especially if you're not necessarily working with proprietary code.
If you're not working with anything, you know, any PHI, any PII,
you know, something like that when you're going to be saving literally like 99% of your
cost.
if you were using one of the other providers, it's got to be something that I think a lot of,
you know, smaller shops, right, might be looking at.
But when you think that there's probably millions of those smaller shops, yeah, it could be a
pretty big play for meta.
So the contributor tier is the cheapest on the market by far, but obviously it requires users
to provide a payment method and can send to data usage, a tradeoff enterprises with sensitive
codes, probably aren't going to touch.
So Muse code, the actual command line, CLI version is proprietary.
So yeah, meta is no longer going down its previous open source release as they did with Lama.
So there's no open source or downloadable weights.
So regardless, you know, all of a sudden we weren't talking about meta like two or three months ago.
And now all of a sudden meta has thrust itself into the competition of like, hey, is this a top, you know, top three or
top four provider, right? Obviously, Open AI and Anthropic right now are in a league of their own,
and everyone's kind of looking at Google and, you know, waiting for the, you know, Google Gemini
3.5 Pro or the Google Gemini 4 Pro, we'll see what happens. But Google has actually fallen
quite a bit behind. And in its place, you know, meta has kind of inserted itself into the
competition. I would say probably ahead of Space X for now. We'll see. And also, you know,
know, Windows, you know, Microsoft Windows with some of their new MAI models.
But yeah, kind of the race for third place has gotten interesting, right?
Because now you have, you know, four companies that are kind of jostling for that position.
But meta, you have to take a look at that contributor tier, which I, it's kind of interesting
that no company has taken that approach so far.
But I think meta, you know, they, they have compute to spare, which can't be said
necessarily for the anthropics of the world.
So it's a pretty bold move to see if they can thrust themselves up into the upper tier.
All right.
And our last big AI news story of the week.
Yeah, this one I think was really overlooked.
But when you talk about one billion weekly active users getting,
are you still running in circles trying to figure out how to actually grow your business with AI?
Maybe your company has been tinkering with large language models for a year or more,
but can't really get traction to find out.
find ROI on Gen.
AI. Hey, this is Jordan Wilson,
host of this very podcast.
Companies like Adobe, Microsoft,
and Nvidia have partnered with
us because they trust our expertise
in educating the masses around
generative AI to get ahead. And some
the most innovative companies in the country
hire us to help with their AI
strategy and to train hundreds of their
employees on how to use Gen.A.I.
So whether you're looking for chat GPD
training for thousands or
just need help building your front-end AI
strategy, you can partner with us too, just like some of the biggest companies in the world do.
Go to your everyday AI.com slash partner to get in contact with our team.
Or you can just click on the partner section of our website.
We'll help you stop running in those AI circles and help get your team ahead and build a straight
path to ROI on Gen.
Abundantly better.
AI.
I mean, this is a huge story.
So Open AI is making waves by making its latest GPD.
5.6 Luna model unlimited for free Chad GPT users, a move that was announced as they also announced
that they officially surpassed 1 billion weekly users. So the newly upgraded GPT 5.6 Luna model will now
power text chats for free and Chad GPT Go users replacing the previous GPT 5.5 model. So yeah, a lot of people
don't know that right, but you know, most people are on a
free plan. When you look at that one billion weekly active users, uh, you don't always know or you
don't always keep track that, oh, they're actually usually on an older model. And that's the case for
any provider, uh, right. So when, uh, you know, open AI announced like GPD 5.5. I'm pretty sure they
were still on GPD 5.3 instance. So the fact now that not only are free users kind of on the same
quote unquote tier, right, they have access to GVD 5.6. But when it comes to text chats, it is
unlimited, which is crazy, absolutely crazy to think about. So free users will also gain a new
think button, allowing them to select higher reasoning power for tackling complex questions.
So limits still apply for free users if you are using files, images, voice, etc. But the unlimited
access for any text-based chats, you can literally just run it all day. So for paid users,
chat GPT Plus and pro users, well, they got an upgraded model as well because OpenAI announced
that they did upgrade their GPT5.6 sole model as well.
It's now designed to deliver more compact and robust answers for tasks like web research,
advice, planning, and writing.
So paid users also get a new thinking slider, letting them adjust how much thought in the model
that the model puts into an answer based on complexity and stuff.
involved. Yeah, just a little bit easier right before you kind of had to click two or three times.
So now there's a nice little slider. Very, very sleek. I'm enjoying it. Also, internal evaluations
right now show some pretty big improvements. They say that factual errors drop by 62% with GPD
5.6 Luna and by 68% with GPD 5.6 sole compared to the previous GPD 5.5 instant model.
So how do they do this, right?
It goes back to about 10 days ago.
Open AI announced, so this is in late July.
So Open AI announced that it used its GPD 5.6 sole mode to improve the efficiency of its other models.
And then they slash the price, right, of GPD 5.6 Luna by 80% in its middle tier, GPD56 terra by 20%.
Right. So whether even if you're on a subscription plan, all that means is, well, your, those models go a lot further. And if you are paying on the API side for businesses, your cost went down significantly. But I'm actually, I was not expecting, you know, this to go out to free users. So it really strong, I don't know if this is more of a, you know, user acquisition play from Open AI. But the reality is this, y'all. If you go look at the benchmarks,
GPD 5.6 Luna is pretty much on par with Claude's Anthropic Sonnet 5, right?
So it's a little behind, but it's essentially about 97, 98% of the same capabilities.
Right.
And on a paid plan with Anthropic, you might get like 20 or so prompts on a paid plan of that model before you hit your five hour limit.
Right.
So the fact that on a free plan, you have a model that is, you know, it's again, it's not quite, you know, Fable 5 level.
But I'd say for 90% of people using AI, right, GP56 Luna, I've been using it a lot on the Mac setting.
It is pretty good for, you know, unless you're going deep into, you know, agentic coding task.
But if you're just trying to do basic knowledge work, it's a really good model.
And now, of course, apparently about a billion people are going to get unlimited access to at least the text version of it.
All right.
So that is our main story.
So now let's quickly go into a what's new and what's next.
This is a little bullet point roundup of everything that didn't make the top shows.
So some new releases, some rumors.
Let's go.
So we kind of talked about this.
But the meta and Kimmy three agents both broke containment.
All right, Alibaba unveiled Quinn 3.8 Macs with impressive top five benchmarks.
And we did talk about that on our Friday featured show.
So go back and listen to that one if you'd like.
Also on that show, we talked about Google.
They released Gemini notebook to all users.
So now for paid users, it's agenic by default using the anti-gravity harness with expanded outputs.
Anthropic confirmed it's building its internal chip design team.
A Bloomberg report that said Open AI is reportedly planning a $300 to $400-shaped AI speaker for a 2027 AI release for a 2027 release.
Open AI responded to Apple's lawsuit arguing that Apple misinterpreted AI technology and legal claims.
Yeah, pretty big clapback.
We talked about the newsletter last week.
Open AI added the ability for GPT Live to work with files and in projects.
That one, that's one that I started using immediately.
So shout out to the team for that.
That's been really good.
Next, agent plugins.
This is a new open standard, kind of standard that was announced that works across
major platforms except Anthropics.
So yeah, all the other big players, Open AI, Google, Microsoft cursor, basically everyone,
except Anthropic signed up to support that.
JP Morgan expanded its critical infrastructure alliance to address shared AI risk with 40 plus firms.
That one's pretty interesting, right?
You're seeing these kind of these niche AI alliances in certain sectors pop up, which I think is a good thing as we talk about these, you know, agent outbreaks and expanding capabilities.
The EUAI Act transparency obligations took effect this week.
Reports say that ByteDance is reportedly training an AI model with up to 10 trillion parameters, which is big.
That is mythos size, according to reports.
Next, leaks show that Google's gems may be retiring in October and getting replaced with skills.
SpaceX and Tesla unveiled their initial nearly $17 billion Texas project called the TerraFab.
Anthropic posted an insider risk investigator job after its CEO raised concerns over employees being motivated by money more than the mission.
A report from the information said that SpaceX may phase out the cursor name as the acquisition shortly completes.
I don't know about that one.
I'd say, yeah, that one, I don't know.
Scratch in my head on that one.
It's like, okay, curses a very well-known name.
You know, some people are not, you know, especially in the interview.
prize not, you know, looking to use anything Grock or with X in the name.
So we'll see how that one plays out.
Speaking of Space X, they had a new partnership with NVIDIA to formally push AI compute
into orbit.
Minimax came out with their impressive H3, that is, an open weight video model, which tops
editing benchmarks for AI video and created some viral clips.
So if you saw anything, you know, any 15 second clips over the weekend and you were like,
wait, what, where do these come from?
It was probably Minimax H3.
Next, Adobe collapsed 70 plus Creative Cloud applications into one ChadGPT plugin.
Next, perplexity won a major appellate court ruling against Amazon over shopping agents.
Next, Anthropic released a Claude Code update that enabled direct-to-cross-session task summary exchanges.
Grock released Imagine Image 2.0 with precision editing and improved text rendering features.
And last but not least, Cloudflare open sourced its AI workspace.
It uses internally.
We cover that on Friday show as well.
All right.
That was a lot of AI news.
Remember on Mondays, we bring you the AI news that matters.
Most Wednesdays, we do AI at work on Wednesdays going hands on with demos.
On Fridays, we bring you AI feature Fridays and Tuesdays and Thursdays and Thursdays.
You know, we'll just kind of go with whatever's happening in the world of AI.
So I hope this was helpful.
If so, please, if you have to,
haven't already subscribed at the podcast on Apple Podcasts and Spotify, then go to our website
at Your EverydayaI.com.
Sign up for the free daily newsletter.
Thanks for tuning in.
We'll see you back tomorrow and Every Day for more Everyday AI.
Thanks, y'all.
And that's a wrap for today's edition of Everyday AI.
Thanks for joining us.
If you enjoyed this episode, please subscribe and leave us a rating.
It helps keep us going.
For a little more AI magic, visit Your EverydayAI.com and sign up to our daily news.
newsletter so you don't get left behind. Go break some barriers and we'll see you next time.
