Better Offline - Monologue: Jacob Coxon and the AI Safety Grift
Episode Date: September 11, 2026In this week's Better Offline monologue, Ed Zitron runs through the nonsensical, pro-AI lab scare campaign driven by former researcher Jacob Coxon, how the media keeps falling for AI safety grifts, an...d how dangerous AI is already in the wrong hands - those of Anthropic and OpenAI.WIRED interview with Coxon: https://archive.md/2026.09.10-091328/https://www.wired.com/story/anthropic-researcher-quits-jacob-coxon-ai-fears-humanity/Thread on Hubinger and Coxon: https://bsky.app/profile/edzitron.com/post/3mv4suhyexc2g Save $10 off a year of my premium newsletter: https://edzitronswheresyouredatghostio.outpost.pub/public/promo-subscription/gzqwkv54e1 YOU CAN NOW BUY BETTER OFFLINE MERCH! Go to https://cottonbureau.com/people/better-offline and use code FREE99 for free shipping on orders of $99 or more. --- LINKS: https://www.tinyurl.com/betterofflinelinks Newsletter: https://www.wheresyoured.at/ Reddit: https://www.reddit.com/r/BetterOffline/ Discord: chat.wheresyoured.at Ed's Socials: https://twitter.com/edzitron https://www.instagram.com/edzitron https://bsky.app/profile/edzitron.com https://www.threads.net/@edzitron Email Me: ez@betteroffline.comSee omnystudio.com/listener for privacy information.
Transcript
Discussion (0)
This is an I-Heart podcast.
Guaranteed Human.
Hi, this is Kylie Breakman.
Angel Geratana.
Jeremy Colhain.
Patrick McDonald.
And we're the hosts of the Artists on Artists on Artists on Artists Podcast.
The Improvised Character Comedy Podcasts
where each week a new panel of artists discuss their craft.
Join us for a panel of Eurovision participants.
And we do a little Halloween sometimes.
Does the Pope dress up?
He winks on Halloween.
He gets right up next to the microphone.
I hear you get to.
The obel's tear, the eyelashes hit each other.
You never seen Italy so quiet
as when the Pope gets up to wink on Halloween.
Listen to Artists on Artists on Artists
every Thursday on the IHeart Radio app,
Apple Podcasts, or wherever you get your podcasts.
Didn't catch the latest Roland Martin unfiltered podcast?
Here's what you missed.
There should not be a single law enforcement agency.
It does not have body cameras.
It's real.
Black farmers have been under attack.
This is just the latest example of them just slapping DEI
on anything. It's raw. And I'm sitting here going, they are playing y'all for fools.
Catch Roland Martin's daily commentary on the Black Information Network. And download Roland Martin
unfiltered on the IHart radio app, Apple Podcasts, wherever you get your podcasts.
The new NFL season is here. And you should be listening to NFL Daily as we march along
to Super Bowl 61. It is in the name NFL Daily. You'll have fresh content in your feed every
day all season long. That game-winning drive, maybe it's Herbert, maybe it's Gino, maybe it's
Mahomes. We'll have the highlights. If Fernando Mendoza or any of those rookies are balling out,
we'll break down the tape. Join me, Greg Rosenthal, and an all-star cast of co-hosts as we preview
and recap every game. Listen to NFL Daily on the IHeart Radio app, Apple Podcasts, or wherever
you get your podcast. Whether you're a seasoned NFL fan or new to the game,
there's one place to keep up with all of it, the league's newest podcast. The NFL
Report hosted by me, Andrew Siciliano. It is your home for everything football, breaking news, expert
analysis, game picks, and hear from your favorite players too. Join me at an all-star cast of experts
for everything you need to know from around the league. Get new episodes to the NFL report Monday through
Friday all season long. Listen to the NFL Report podcast on the IHeart Radio app, Apple Podcast, or wherever
you get your podcasts.
CoeurZone Media. Hello and welcome to this week's Better OffLy
I'm monologue. I'm your host at Zittron.
Let's our online.
I've now had quite a lot of emails about the post by Evan Hubinger and Jacob Coxon of
Anthropics saying that they, and I quote, earnestly believe AI could kill all humans
with a higher than 10% chance that it happens within the next decade.
I want to lead by saying that these men, regardless of their intentions or cynicism,
are beyond loathsome. They are disgusting to me. I find them vile, and I find it equally
vile how many people are falling for their garbage. If they truly believe what they're saying,
the idea that they're continuing to work on a product that could destroy humanity makes them
want to be war criminals. Coxon, who recently quit Anthropic, did so in an extremely
public way that has been covered by much of the mainstream media, claims that he quit to quote
the Wall Street Journal because he doesn't want to participate in an industry-wide rush to build
AI systems that can improve themselves, which, as I will refer to later is referring to recursive
self-improvement, which nobody has proven, actually is possible. In an interview with Wired,
Coxon spoke at length without really explaining what it was he was scared about, outside of one
moment where Max Zeph asked for specifics about his concerns, at which point Coxon incorrectly
describes the hack as AI doing this all of its own volition, referring to the hugging face
attack, only for him to respond when pressed about whether this is companies just moving recklessly
fast, that he didn't want to focus too much on the hugging face attack, because there's plenty
of evidence that the companies don't know how to align the models properly. In other words,
the moment that Coxon was asked to get specific about what the companies are doing wrong and
where they're making mistakes, he punts to a non-answer. By the way, Coxon's solution to all
of these problems is that the AI labs should agree to not rush towards recursive self-improvement.
That's still theoretical term for an AI that can train itself that everybody is talking about like it's real.
In other words, Coxon is suggesting that Sam Orkman and Dario Amaday put their clammy hands together and agreed to delay breeding the Grinch.
We're not getting a Lorax, folks. We're not making Shrek. We will not let the fairy godmother into your house, okay?
Godzilla will not happen. Yet the most laughable part of the interview was, when asked about why he kept working at Open AI and Anthropic for years, when Coxon said that it's also not hyperbole.
that we could cure cancer, because open AI by stealing someone else's work, I'm not getting
into it, it's all over the subreddit, solve the Navia Stokes problem, adding that there really is
no reason that we, referring to the AI industry, can't transfer that to scientific domains like
biology. The kind of thing you say when you're just making shit up for attention, and you're
wrong, you're just fucking wrong. Those are two very different things. A mathematical problem is not
the same as biology. What are you talking about, Jacob? Why are the journalists just printing everything
he says without pushing back. It drives me insane. Nevertheless, people have been saying,
oh, oh, well, he's leaving because he's so concerned. He's got to tell everyone because he's so concerned.
He's so worried. No, no, the piece ends with Coxon, saying that he'd like to do some sort of
independent commentary on where things are going, which means a sub-stack, and he'd like to do something
like AI 2027 or move into some kind of accountability role. Yay! Yeah, yay, there's the Gryft.
Wee! Congrats everyone. Congratulations.
on helping Jacob Coxon get a new job.
Congrats.
Congrats, everyone.
You did it.
You helped elevate a real piece of shit.
Okay.
My dear friends in the media,
my esteemed colleagues who allegedly have been doing this for years,
Anderson Cooper, CNN and Wall Street Journal,
all of you, everyone,
how many fucking times are you going to fall for this?
Why do you believe these people,
other than that some of them are saying vague and scary stuff
and that they happen to be at the companies.
Why do they never really say what it is they're scared of?
And why when they even try and get specific, not that they do very much, do they never
blame the companies?
I say it again.
Nobody has actually achieved recursive self-improvement, nor have they shown any proof that it's
even possible.
But the fact that Coxon is bringing it up in the run-up to Anthropics IPO, and as the rest of
the AI hogs oink about it nonstop makes it impossible for me to be able to be.
believe that Coxon has anything other than the most cynical intentions. People say,
oh, he gave up his options in Anthropic, as if that matters in any way, shape or form. He was
barely there a few months. He probably didn't have that many and saw an opportunity to get a bunch
of attention. He was at open AI for years and likely has a ton of options from there too.
Why in the world does this keep happening? This is how bubbles form. This is how grifter's grift.
It starts and it finishes with the media industry giving up on their jobs.
It starts and finishes with the most important journalists in the world failing.
And I think everyone failed here.
And I am disgusted and outrage to see this happen again.
It's exactly what Matt Schumer did a few months ago with something big is happening.
It's the Citrini member about the global intelligence crisis.
It's the same shit.
It's science fiction dressed.
up as fact. And the fact it works is only because, to quote Ed Ellson, we have this cult-like
worship of the wealthy, where we believe that the people at the companies and that people
will move around from these companies in a way that's only honest, in a way that's only true,
when in fact these are some of the most cynical people in the world, which is pretty obvious
in the fact that they go, oh, I'm afraid that this stuff is going to destroy the world, but they
keep working on it. And I don't give a shit if he's not working at one of the companies. If he goes
into an alignment role or a non-profit for this shit, it's the same thing. It's a way to keep
milking the cow. And Coxon's entire argument centers around this nebulous idea that things are
moving really fast, but never really crystallizes on what it is that's moving fast,
not does it ever hold the companies in question accountable. He told Wired that it was important to ask
executives on the record to give an actual probability for extinction in the next decade.
And I want to be clear that this statement means absolutely nothing.
And anyone taking the question or the answer seriously is a goddamn mark.
What does the answer even mean?
What's the difference between a 10% and a 30% chance?
What does 10% even mean?
What does any of this mean?
Ah, who cares?
Put the AI hype in the bag.
Who gives a shit, right?
Coxon claims that this is not a marketing stunt.
By the way, the biggest sign that something is a marketing stunt.
is when someone says that.
And that executives and senior researchers
couch their phrasing in the press
to sound sensible,
which is hilarious
because I've been hearing
these kinds of threats for years.
Dario Amadeh,
Wario himself, said to Axios last year,
or maybe CNN,
that there was a 50% chance
of AI wiping out
all of white-collar labor,
or maybe it was a
AI will definitely wipe out
50% of white-collar labor.
The fact that it's not really clear
is kind of the point I'm making.
It's just saying stuff.
In any case, Huberinger, who still works at Anthropic, followed up on his post agreeing about
the dire threats of AI, but adding that the risk from the present models is low.
But don't worry, though, his worries were around superintelligence arising from recursive
self-improvement, which is happening faster than we thought, referring to something that has
not happened yet.
To be clear, there is an actual threat from LLMs.
Anthropic and Open AI using hundreds of billions of dollars worth of infrastructure provided by the largest companies in the world to brute force hack using LLMs bouncing off of each other because they can't find any other uses at scale.
These models are doing exactly what they're trained to do, which would make everybody ask why the fuck they're being trained to do it.
And women are looking for more.
More to themselves, their businesses, their elected leaders, and the world are of them.
And that's why we're thrilled to introduce the honest talk podcast.
Jennifer Stewart. And I'm Catherine Clark. And in this podcast, we interview Canada's most inspiring
women. Entrepreneurs, artists, athletes, politicians, and newsmakers, all at different stages of their
journey. So if you're looking to connect, then we hope you'll join us. Listen to the Honest Talk
podcast on IHeartRadio or wherever you listen to your podcasts. Hey Jonas Brothers with the Hey Jonas
podcast. We've been catching up with some great friends, Paul Rudd, Michael Boubley, Seth Myers,
Nile Horan, Jake Shane, and making some new friends while getting into what's
really like to have sisters, like Alex and Ashton Earl.
She likes to control everyone and tell everyone what to do.
Okay, well, control is a harsh word.
I would say that's a harsh word.
I am more outspoken.
Growing up a lot of the times because she was so quiet,
I would speak for like the both of us.
Now I'm working on like thinking a moment, reeling it in,
thinking before I speak.
And Joey King.
You know if your like mom is yelling at your sibling and even if you're annoyed
with your sibling, you just like won't let that slide.
It's like only you can yell at your sibling.
I'm like, you can't do that here.
That is not talk about her like that.
Only I talk about her like that.
Plus, we've been taking your calls,
listening to your voicemails and giving you advice.
Or at least trying to.
So if you've got catching up to do, now's the perfect time.
Listen to Hey Jonas on the Iheart radio app, Apple Podcasts,
or wherever you get your podcasts.
When I was 14 years old, I was kidnapped and held captive for nine months.
I survived and I've spent my life experience.
exploring how other people survive what should have destroyed them.
I'm Elizabeth Smart, and these are the Survivor Files.
I just remember this low, taunting voice next to my ear saying, shut up, don't say anything.
Every week, I'm with survivors who lived through the unthinkable.
I knew if he woke up, without a doubt, he was going to hurt me.
I started feeling that there was someone at the end of my bed, and I,
just started screaming.
They are abducted, stalked, controlled, and nearly silenced.
But these aren't stories about what's taken from them.
There's stories about what it takes to make it out alive.
Listen to the Survivor Files with Elizabeth Smart on the IHeart Radio app,
Apple Podcasts, or wherever you get your podcasts.
Hi, this is Kylie Breakman.
Angel Geratana.
Jeremy Colhane, Patrick McDonald.
And we're the hosts of the Artists on Artists on Artists Podcast.
The Improvised Character Comedy Podcast where each week a new panel of artists discuss their craft.
Join us for a panel of wellness influencers.
What I try to do is faux journal.
I don't know.
Have you guys heard of faux journaling?
Yeah, it's reading.
It's called reading.
So we don't like to use that word anymore.
Okay.
A panel of New Yorker cartoonists.
Have you guys done Terry Gross's Peloton class?
No!
I've been dying too.
I'm so sorry.
I'm seeing a reading of Sadako and the Thousand-T thousand people.
Paper Cranes by Audra McDonald's at the Wonderland Bookstore, 14.
That's crazy.
I went to the Wonderland bookstore yesterday, and I heard Nora Jones do a reading of Zero Dark 30.
No way.
You have to come to the last bookstore, though.
Kim Cottrell is reading the guitar tabs of yesterday.
A panel of Eurovision participants.
And we do a little Halloween sometimes.
Does the Pope dress up?
He winks on Halloween.
He gets right up next to the microphone.
I hear you can almost hear the eyelashes hit each other.
You never see.
Italy's so quiet as when the Pope gets up to wink on Halloween.
Listen to Artists on Artists on Artists on Artists on Artists on the IHard Radio app, Apple Podcasts, or wherever you get your podcasts.
As I went into with Cal Newport a few weeks ago, this is not autonomous hacking by agents that escaped sandboxes.
It's anthropic and open AI throwing unlimited compute to make LLMs that can, quote, do cybersecurity.
And having really terrible security practices and not being able to train them in a way that was consistent.
or safe, and then they still use them, they still treat them as if, oh, they'll just work it out,
or just maybe they don't give a shit.
Maybe they don't care.
I think that's probably the most likely outcome here, by the way.
To be clear, whatever anthropic and open AI models that were responsible for the hack are very, very dangerous,
but not because of any sentience or magic or thought.
These are LLMs talking to LLMs that decide what they should do next,
prompting each other again and again and again, and having unlimited resources to do so.
If we weaponise billions of dollars to sink into the automated scripts that hackers have used in the past,
we'd get probably in much the same results, and be sending people to jail.
Somehow, despite years of these dire warnings, the warnings in question never include anything
about what's happening today with really any specificity.
There are no sclergrams about how anthropic and open AI have near unlimited resources
to commit to dangerous, reckless experiments that amount to felony hacking.
There doesn't seem to be any interest in holding these people,
there doesn't actually seem to be any interest in stopping anything. It just seems that we all want to
have a jerk-off theatre around scary things that nobody actually wants to define. Even when Coxon
discussed the hugging face attack with Wired, he framed everything in terms of helplessness,
calling the LLMs an agent swarm and saying that one of the agents did this hack as part of a general
strategy for understanding more about the grader, versus a piece of software with poor security controls
and unreliable automations, pursuing a task in an unexpected way.
Coxon's language intentionally minimizes any kind of responsibility or accountability on the part of the
AI labs, choosing instead to put the blame on powerful AI that they can't understand.
There really is no economic reason to do cybersecurity stuff with LMs,
outside of the fact that these companies are hitting diminishing returns encoding,
and that there are exhaustive troves of vulnerability data online for them to train their models with
and do exactly what they've been doing.
brute force hacking of GPUs has existed for over a decade, albeit with different techniques,
and the innovation here is a direct result of the unbelievable resources handed to these two irresponsible, disgraceful companies.
But China might do this is not an answer to why are American companies doing this,
because this is clearly a situation where Open AI and Anthropica built something they were aware would mindlessly bash its head
and millions of dollars of compute against the problem until it broke through.
Make no mistake.
Dangerous AI is already in the wrong hands, those of anthropic and open AI.
If a regular person did the hugging face attack, they'd be in jail, because this seems like,
based on discussions with experts, a clear case of felony hacking.
And if this wasn't the result of powerful AI, everyone involved would be in shackles.
These companies act as if they're being forced to build these tools and run these experiments,
when they're actually acting with complete autonomy far more than they should ever have been a
to have, with resources that I've repeatedly said and will say again are virtually unlimited.
They act as if they have no responsibility or ability to stop their own experiments, or monitor
them or really do anything but continue to feed them training data and watch what happens in
awe, but also wrote long blogs about how they're scared, and also fucking hack people.
Let me be very clear about this. Anthropic and open AI have repeatedly and flagrantly engaged
in what appears to be illegal hacking of multiple different systems,
and have done so using AI GPUs sold by Nvidia
and powered by infrastructure built for them by Amazon, Google and Microsoft.
They are the ones making every one of these choices,
and they will continue to do so every time we elevate the voice
of a cynical grifter like Jacob Coxson.
The people that work at these companies are willingly engaging in acts
that, if not criminal, are morally and ethically bankrupt.
And the people who keep giving these dire warnings about the dangers of AI never seem to give a shit about what's actually happening.
Always keeping your eye on some non-specific harms in some non-specific future, all while saying that they alone are the ones who will protect you from whatever it is that might not happen,
that they're blowing the whistle on companies that have paid them hundreds of thousands of dollars.
And indeed, we don't know.
And I think there are actually very good questions to ask about what his compensation from Anthropic was, and whether it's ongoing.
I ask this because a lot of people from Anthropic are very supportive of a guy who just quit,
a guy who just quit and is basically accusing the company of creating something dangerous that will
destroy the world. Isn't that strange? Tocoxin, Hubinger and any other AI Duma, I have to ask,
what is it I am meant to do with this information? What is it any of us are meant to do?
Are we meant to be scared of you? Are we meant to sign up?
up to Claude Max? Are we meant to invest in the Anthropic IPO? Are we meant to regulate this stuff?
How? Why do you never say that? Why do you never say what it is we're meant to do to stop this?
Why do none of you fucking people ever have any kind of suggestion or call to action or anything
other than some sort of self-serving or company-serving diatribe about how powerful and scary AI is?
And if you hear any loathing in my voice, that's because I find you loathsome.
I find what you're doing, disgusting, I think you are cynical, I think you're malcontents,
and I think you're bad for society.
And the fact you're being elevated is dangerous, but not for the reasons you're saying,
but because it proves that our media ecosystem barely has object permanence.
Congratulations on exploiting it.
You're a fucking asshole.
I think the answers to these questions are actually pretty simple.
You don't actually give much of a shit and you want some attention.
If you actually feared this stuff and felt a moral obligation to tell people,
you'd be specific about what it is we should be scared of and indeed tell us what we need to do next.
Every time it's the same sorry fucking story.
Oh, we're building something so scary and powerful and we can't stop it.
nevertheless we're going to keep building it for some reason we never discuss but when it destroys
humanity in some way we can't explain you'll thank us i guess
great thank you man thank you for taking up the airwaves thank you from distracting from the real
problems you guys are horrible you guys are genuinely kind of evil if you wanted to do something
about this you'd do something about this if you cared you'd care you'd be saying we need to
regulate in this way you wouldn't be doing this mean
mealy-mouthed horse shit of, oh, maybe they'll stop doing recursive self-improvement,
a thing that doesn't exist. Weird that you don't talk about what's happening today.
Weird that you don't talk about stopping the hugging face attack of the future.
It's always the positive. It's always the positive in even when being negative.
And it's hard to see this as anything other than cynical marketing and attention seeking
from two guys who should spend more time working on building technology than posting on
social media about how scary things are getting.
I don't like them. I don't trust them. And the whole thing's a fucking grift.
Next week, I'll be back with Cal Newport and Adam Becker to talk about the people that
inspired this AI Dumerism, the rationalists, and how to push back against their noxious
bullshit. I'll catch you then. I love you all.
Hi, this is Kylie Breakman. Angel Geratana.
Jeremy Colhane, Patrick MacDonald. And we're the hosts of the Artists on Artists on Artists
podcast. The improvised character comedy
podcast where each week a new
panel of artists discuss their craft.
Join us for a panel of
Eurovision participants. And we do
a little Halloween sometimes.
Does the Pope dress up? He
winks on Halloween. He gets right
up next to the microphone. I hear
you can almost hear the eyelashes hit each other.
You never seen Italy
so quiet as when the Pope gets
up to wink on Halloween.
Listen to Artists on Artists on Artists
every Thursday on the IHeartRadio app.
Apple Podcasts or wherever you get your podcasts.
Didn't catch the latest Roland Martin Unfiltered Podcast?
Here's what you missed.
People wake up and go, oh damn, wait a hold up.
They change all of that?
Yes.
It's real.
This is a wholesale attack.
It is targeting black people in every federal agency.
It's raw.
White folks have never allowed that reckoning to last more than a decade.
Catch Roland Martin's daily commentary on the Black Information Network.
And download Roland Martin Unfiltered on the IHart Radio app
Apple Podcasts, wherever you get your podcasts.
The new NFL season is here, and you should be listening to NFL Daily as we march along
to Super Bowl 61.
It is in the name, NFL Daily.
You'll have fresh content in your feed every day all season long.
That game-winning drive, maybe it's Herbert, maybe it's Gino, maybe it's Mahomes.
We'll have the highlights.
If Fernando Mendoza or any of the rookies are balling out, we'll break down the tape.
Join me, Greg Rosenthal, and an all-star cast of co-host as we preview and recap.
every game. Listen to NFL daily on the IHeart Radio app, Apple Podcasts, or wherever you get your podcast.
Hey, it's Bobby Bones. Join me and former NFL quarterback Matt Castle every Wednesday on our
podcast, lots to say with me, Bobby Bones and Matt Castle. You're in training camp and the rookie
quarterback has one good throwing session in front of the media. Suddenly everybody on line says he should
start over you week one. How do you handle this? You just go back out to practice the next day. I wait
firm to mess up.
Listen to lots to say with Bobby Bones and Matt Castle on the IHeart Radio app, Apple Podcasts, or wherever you get your podcasts.
This is an IHeart podcast. Guaranteed human.
