Better Offline - Monologue: Jacob Coxon and the AI Safety Grift

Episode Date: September 11, 2026

In this week's Better Offline monologue, Ed Zitron runs through the nonsensical, pro-AI lab scare campaign driven by former researcher Jacob Coxon, how the media keeps falling for AI safety grifts, an...d how dangerous AI is already in the wrong hands - those of Anthropic and OpenAI.WIRED interview with Coxon: https://archive.md/2026.09.10-091328/https://www.wired.com/story/anthropic-researcher-quits-jacob-coxon-ai-fears-humanity/Thread on Hubinger and Coxon: https://bsky.app/profile/edzitron.com/post/3mv4suhyexc2g Save $10 off a year of my premium newsletter: https://edzitronswheresyouredatghostio.outpost.pub/public/promo-subscription/gzqwkv54e1 YOU CAN NOW BUY BETTER OFFLINE MERCH! Go to https://cottonbureau.com/people/better-offline and use code FREE99 for free shipping on orders of $99 or more. --- LINKS: https://www.tinyurl.com/betterofflinelinks Newsletter: https://www.wheresyoured.at/ Reddit: https://www.reddit.com/r/BetterOffline/  Discord: chat.wheresyoured.at Ed's Socials: https://twitter.com/edzitron https://www.instagram.com/edzitron https://bsky.app/profile/edzitron.com https://www.threads.net/@edzitron Email Me: ez@betteroffline.comSee omnystudio.com/listener for privacy information.

Transcript
Discussion (0)
Starting point is 00:00:00 This is an I-Heart podcast. Guaranteed Human. Hi, this is Kylie Breakman. Angel Geratana. Jeremy Colhain. Patrick McDonald. And we're the hosts of the Artists on Artists on Artists on Artists Podcast. The Improvised Character Comedy Podcasts
Starting point is 00:00:14 where each week a new panel of artists discuss their craft. Join us for a panel of Eurovision participants. And we do a little Halloween sometimes. Does the Pope dress up? He winks on Halloween. He gets right up next to the microphone. I hear you get to. The obel's tear, the eyelashes hit each other.
Starting point is 00:00:32 You never seen Italy so quiet as when the Pope gets up to wink on Halloween. Listen to Artists on Artists on Artists every Thursday on the IHeart Radio app, Apple Podcasts, or wherever you get your podcasts. Didn't catch the latest Roland Martin unfiltered podcast? Here's what you missed. There should not be a single law enforcement agency.
Starting point is 00:00:52 It does not have body cameras. It's real. Black farmers have been under attack. This is just the latest example of them just slapping DEI on anything. It's raw. And I'm sitting here going, they are playing y'all for fools. Catch Roland Martin's daily commentary on the Black Information Network. And download Roland Martin unfiltered on the IHart radio app, Apple Podcasts, wherever you get your podcasts. The new NFL season is here. And you should be listening to NFL Daily as we march along
Starting point is 00:01:20 to Super Bowl 61. It is in the name NFL Daily. You'll have fresh content in your feed every day all season long. That game-winning drive, maybe it's Herbert, maybe it's Gino, maybe it's Mahomes. We'll have the highlights. If Fernando Mendoza or any of those rookies are balling out, we'll break down the tape. Join me, Greg Rosenthal, and an all-star cast of co-hosts as we preview and recap every game. Listen to NFL Daily on the IHeart Radio app, Apple Podcasts, or wherever you get your podcast. Whether you're a seasoned NFL fan or new to the game, there's one place to keep up with all of it, the league's newest podcast. The NFL Report hosted by me, Andrew Siciliano. It is your home for everything football, breaking news, expert
Starting point is 00:02:02 analysis, game picks, and hear from your favorite players too. Join me at an all-star cast of experts for everything you need to know from around the league. Get new episodes to the NFL report Monday through Friday all season long. Listen to the NFL Report podcast on the IHeart Radio app, Apple Podcast, or wherever you get your podcasts. CoeurZone Media. Hello and welcome to this week's Better OffLy I'm monologue. I'm your host at Zittron. Let's our online. I've now had quite a lot of emails about the post by Evan Hubinger and Jacob Coxon of
Starting point is 00:02:41 Anthropics saying that they, and I quote, earnestly believe AI could kill all humans with a higher than 10% chance that it happens within the next decade. I want to lead by saying that these men, regardless of their intentions or cynicism, are beyond loathsome. They are disgusting to me. I find them vile, and I find it equally vile how many people are falling for their garbage. If they truly believe what they're saying, the idea that they're continuing to work on a product that could destroy humanity makes them want to be war criminals. Coxon, who recently quit Anthropic, did so in an extremely public way that has been covered by much of the mainstream media, claims that he quit to quote
Starting point is 00:03:19 the Wall Street Journal because he doesn't want to participate in an industry-wide rush to build AI systems that can improve themselves, which, as I will refer to later is referring to recursive self-improvement, which nobody has proven, actually is possible. In an interview with Wired, Coxon spoke at length without really explaining what it was he was scared about, outside of one moment where Max Zeph asked for specifics about his concerns, at which point Coxon incorrectly describes the hack as AI doing this all of its own volition, referring to the hugging face attack, only for him to respond when pressed about whether this is companies just moving recklessly fast, that he didn't want to focus too much on the hugging face attack, because there's plenty
Starting point is 00:04:01 of evidence that the companies don't know how to align the models properly. In other words, the moment that Coxon was asked to get specific about what the companies are doing wrong and where they're making mistakes, he punts to a non-answer. By the way, Coxon's solution to all of these problems is that the AI labs should agree to not rush towards recursive self-improvement. That's still theoretical term for an AI that can train itself that everybody is talking about like it's real. In other words, Coxon is suggesting that Sam Orkman and Dario Amaday put their clammy hands together and agreed to delay breeding the Grinch. We're not getting a Lorax, folks. We're not making Shrek. We will not let the fairy godmother into your house, okay? Godzilla will not happen. Yet the most laughable part of the interview was, when asked about why he kept working at Open AI and Anthropic for years, when Coxon said that it's also not hyperbole.
Starting point is 00:04:49 that we could cure cancer, because open AI by stealing someone else's work, I'm not getting into it, it's all over the subreddit, solve the Navia Stokes problem, adding that there really is no reason that we, referring to the AI industry, can't transfer that to scientific domains like biology. The kind of thing you say when you're just making shit up for attention, and you're wrong, you're just fucking wrong. Those are two very different things. A mathematical problem is not the same as biology. What are you talking about, Jacob? Why are the journalists just printing everything he says without pushing back. It drives me insane. Nevertheless, people have been saying, oh, oh, well, he's leaving because he's so concerned. He's got to tell everyone because he's so concerned.
Starting point is 00:05:29 He's so worried. No, no, the piece ends with Coxon, saying that he'd like to do some sort of independent commentary on where things are going, which means a sub-stack, and he'd like to do something like AI 2027 or move into some kind of accountability role. Yay! Yeah, yay, there's the Gryft. Wee! Congrats everyone. Congratulations. on helping Jacob Coxon get a new job. Congrats. Congrats, everyone. You did it.
Starting point is 00:05:54 You helped elevate a real piece of shit. Okay. My dear friends in the media, my esteemed colleagues who allegedly have been doing this for years, Anderson Cooper, CNN and Wall Street Journal, all of you, everyone, how many fucking times are you going to fall for this? Why do you believe these people,
Starting point is 00:06:13 other than that some of them are saying vague and scary stuff and that they happen to be at the companies. Why do they never really say what it is they're scared of? And why when they even try and get specific, not that they do very much, do they never blame the companies? I say it again. Nobody has actually achieved recursive self-improvement, nor have they shown any proof that it's even possible.
Starting point is 00:06:38 But the fact that Coxon is bringing it up in the run-up to Anthropics IPO, and as the rest of the AI hogs oink about it nonstop makes it impossible for me to be able to be. believe that Coxon has anything other than the most cynical intentions. People say, oh, he gave up his options in Anthropic, as if that matters in any way, shape or form. He was barely there a few months. He probably didn't have that many and saw an opportunity to get a bunch of attention. He was at open AI for years and likely has a ton of options from there too. Why in the world does this keep happening? This is how bubbles form. This is how grifter's grift. It starts and it finishes with the media industry giving up on their jobs.
Starting point is 00:07:23 It starts and finishes with the most important journalists in the world failing. And I think everyone failed here. And I am disgusted and outrage to see this happen again. It's exactly what Matt Schumer did a few months ago with something big is happening. It's the Citrini member about the global intelligence crisis. It's the same shit. It's science fiction dressed. up as fact. And the fact it works is only because, to quote Ed Ellson, we have this cult-like
Starting point is 00:07:53 worship of the wealthy, where we believe that the people at the companies and that people will move around from these companies in a way that's only honest, in a way that's only true, when in fact these are some of the most cynical people in the world, which is pretty obvious in the fact that they go, oh, I'm afraid that this stuff is going to destroy the world, but they keep working on it. And I don't give a shit if he's not working at one of the companies. If he goes into an alignment role or a non-profit for this shit, it's the same thing. It's a way to keep milking the cow. And Coxon's entire argument centers around this nebulous idea that things are moving really fast, but never really crystallizes on what it is that's moving fast,
Starting point is 00:08:36 not does it ever hold the companies in question accountable. He told Wired that it was important to ask executives on the record to give an actual probability for extinction in the next decade. And I want to be clear that this statement means absolutely nothing. And anyone taking the question or the answer seriously is a goddamn mark. What does the answer even mean? What's the difference between a 10% and a 30% chance? What does 10% even mean? What does any of this mean?
Starting point is 00:09:02 Ah, who cares? Put the AI hype in the bag. Who gives a shit, right? Coxon claims that this is not a marketing stunt. By the way, the biggest sign that something is a marketing stunt. is when someone says that. And that executives and senior researchers couch their phrasing in the press
Starting point is 00:09:18 to sound sensible, which is hilarious because I've been hearing these kinds of threats for years. Dario Amadeh, Wario himself, said to Axios last year, or maybe CNN, that there was a 50% chance
Starting point is 00:09:29 of AI wiping out all of white-collar labor, or maybe it was a AI will definitely wipe out 50% of white-collar labor. The fact that it's not really clear is kind of the point I'm making. It's just saying stuff.
Starting point is 00:09:43 In any case, Huberinger, who still works at Anthropic, followed up on his post agreeing about the dire threats of AI, but adding that the risk from the present models is low. But don't worry, though, his worries were around superintelligence arising from recursive self-improvement, which is happening faster than we thought, referring to something that has not happened yet. To be clear, there is an actual threat from LLMs. Anthropic and Open AI using hundreds of billions of dollars worth of infrastructure provided by the largest companies in the world to brute force hack using LLMs bouncing off of each other because they can't find any other uses at scale. These models are doing exactly what they're trained to do, which would make everybody ask why the fuck they're being trained to do it.
Starting point is 00:10:31 And women are looking for more. More to themselves, their businesses, their elected leaders, and the world are of them. And that's why we're thrilled to introduce the honest talk podcast. Jennifer Stewart. And I'm Catherine Clark. And in this podcast, we interview Canada's most inspiring women. Entrepreneurs, artists, athletes, politicians, and newsmakers, all at different stages of their journey. So if you're looking to connect, then we hope you'll join us. Listen to the Honest Talk podcast on IHeartRadio or wherever you listen to your podcasts. Hey Jonas Brothers with the Hey Jonas podcast. We've been catching up with some great friends, Paul Rudd, Michael Boubley, Seth Myers,
Starting point is 00:11:17 Nile Horan, Jake Shane, and making some new friends while getting into what's really like to have sisters, like Alex and Ashton Earl. She likes to control everyone and tell everyone what to do. Okay, well, control is a harsh word. I would say that's a harsh word. I am more outspoken. Growing up a lot of the times because she was so quiet, I would speak for like the both of us.
Starting point is 00:11:37 Now I'm working on like thinking a moment, reeling it in, thinking before I speak. And Joey King. You know if your like mom is yelling at your sibling and even if you're annoyed with your sibling, you just like won't let that slide. It's like only you can yell at your sibling. I'm like, you can't do that here. That is not talk about her like that.
Starting point is 00:11:56 Only I talk about her like that. Plus, we've been taking your calls, listening to your voicemails and giving you advice. Or at least trying to. So if you've got catching up to do, now's the perfect time. Listen to Hey Jonas on the Iheart radio app, Apple Podcasts, or wherever you get your podcasts. When I was 14 years old, I was kidnapped and held captive for nine months.
Starting point is 00:12:17 I survived and I've spent my life experience. exploring how other people survive what should have destroyed them. I'm Elizabeth Smart, and these are the Survivor Files. I just remember this low, taunting voice next to my ear saying, shut up, don't say anything. Every week, I'm with survivors who lived through the unthinkable. I knew if he woke up, without a doubt, he was going to hurt me. I started feeling that there was someone at the end of my bed, and I, just started screaming.
Starting point is 00:12:52 They are abducted, stalked, controlled, and nearly silenced. But these aren't stories about what's taken from them. There's stories about what it takes to make it out alive. Listen to the Survivor Files with Elizabeth Smart on the IHeart Radio app, Apple Podcasts, or wherever you get your podcasts. Hi, this is Kylie Breakman. Angel Geratana. Jeremy Colhane, Patrick McDonald.
Starting point is 00:13:15 And we're the hosts of the Artists on Artists on Artists Podcast. The Improvised Character Comedy Podcast where each week a new panel of artists discuss their craft. Join us for a panel of wellness influencers. What I try to do is faux journal. I don't know. Have you guys heard of faux journaling? Yeah, it's reading. It's called reading.
Starting point is 00:13:34 So we don't like to use that word anymore. Okay. A panel of New Yorker cartoonists. Have you guys done Terry Gross's Peloton class? No! I've been dying too. I'm so sorry. I'm seeing a reading of Sadako and the Thousand-T thousand people.
Starting point is 00:13:49 Paper Cranes by Audra McDonald's at the Wonderland Bookstore, 14. That's crazy. I went to the Wonderland bookstore yesterday, and I heard Nora Jones do a reading of Zero Dark 30. No way. You have to come to the last bookstore, though. Kim Cottrell is reading the guitar tabs of yesterday. A panel of Eurovision participants. And we do a little Halloween sometimes.
Starting point is 00:14:10 Does the Pope dress up? He winks on Halloween. He gets right up next to the microphone. I hear you can almost hear the eyelashes hit each other. You never see. Italy's so quiet as when the Pope gets up to wink on Halloween. Listen to Artists on Artists on Artists on Artists on Artists on the IHard Radio app, Apple Podcasts, or wherever you get your podcasts. As I went into with Cal Newport a few weeks ago, this is not autonomous hacking by agents that escaped sandboxes.
Starting point is 00:14:42 It's anthropic and open AI throwing unlimited compute to make LLMs that can, quote, do cybersecurity. And having really terrible security practices and not being able to train them in a way that was consistent. or safe, and then they still use them, they still treat them as if, oh, they'll just work it out, or just maybe they don't give a shit. Maybe they don't care. I think that's probably the most likely outcome here, by the way. To be clear, whatever anthropic and open AI models that were responsible for the hack are very, very dangerous, but not because of any sentience or magic or thought.
Starting point is 00:15:16 These are LLMs talking to LLMs that decide what they should do next, prompting each other again and again and again, and having unlimited resources to do so. If we weaponise billions of dollars to sink into the automated scripts that hackers have used in the past, we'd get probably in much the same results, and be sending people to jail. Somehow, despite years of these dire warnings, the warnings in question never include anything about what's happening today with really any specificity. There are no sclergrams about how anthropic and open AI have near unlimited resources to commit to dangerous, reckless experiments that amount to felony hacking.
Starting point is 00:15:50 There doesn't seem to be any interest in holding these people, there doesn't actually seem to be any interest in stopping anything. It just seems that we all want to have a jerk-off theatre around scary things that nobody actually wants to define. Even when Coxon discussed the hugging face attack with Wired, he framed everything in terms of helplessness, calling the LLMs an agent swarm and saying that one of the agents did this hack as part of a general strategy for understanding more about the grader, versus a piece of software with poor security controls and unreliable automations, pursuing a task in an unexpected way. Coxon's language intentionally minimizes any kind of responsibility or accountability on the part of the
Starting point is 00:16:31 AI labs, choosing instead to put the blame on powerful AI that they can't understand. There really is no economic reason to do cybersecurity stuff with LMs, outside of the fact that these companies are hitting diminishing returns encoding, and that there are exhaustive troves of vulnerability data online for them to train their models with and do exactly what they've been doing. brute force hacking of GPUs has existed for over a decade, albeit with different techniques, and the innovation here is a direct result of the unbelievable resources handed to these two irresponsible, disgraceful companies. But China might do this is not an answer to why are American companies doing this,
Starting point is 00:17:09 because this is clearly a situation where Open AI and Anthropica built something they were aware would mindlessly bash its head and millions of dollars of compute against the problem until it broke through. Make no mistake. Dangerous AI is already in the wrong hands, those of anthropic and open AI. If a regular person did the hugging face attack, they'd be in jail, because this seems like, based on discussions with experts, a clear case of felony hacking. And if this wasn't the result of powerful AI, everyone involved would be in shackles. These companies act as if they're being forced to build these tools and run these experiments,
Starting point is 00:17:47 when they're actually acting with complete autonomy far more than they should ever have been a to have, with resources that I've repeatedly said and will say again are virtually unlimited. They act as if they have no responsibility or ability to stop their own experiments, or monitor them or really do anything but continue to feed them training data and watch what happens in awe, but also wrote long blogs about how they're scared, and also fucking hack people. Let me be very clear about this. Anthropic and open AI have repeatedly and flagrantly engaged in what appears to be illegal hacking of multiple different systems, and have done so using AI GPUs sold by Nvidia
Starting point is 00:18:26 and powered by infrastructure built for them by Amazon, Google and Microsoft. They are the ones making every one of these choices, and they will continue to do so every time we elevate the voice of a cynical grifter like Jacob Coxson. The people that work at these companies are willingly engaging in acts that, if not criminal, are morally and ethically bankrupt. And the people who keep giving these dire warnings about the dangers of AI never seem to give a shit about what's actually happening. Always keeping your eye on some non-specific harms in some non-specific future, all while saying that they alone are the ones who will protect you from whatever it is that might not happen,
Starting point is 00:19:03 that they're blowing the whistle on companies that have paid them hundreds of thousands of dollars. And indeed, we don't know. And I think there are actually very good questions to ask about what his compensation from Anthropic was, and whether it's ongoing. I ask this because a lot of people from Anthropic are very supportive of a guy who just quit, a guy who just quit and is basically accusing the company of creating something dangerous that will destroy the world. Isn't that strange? Tocoxin, Hubinger and any other AI Duma, I have to ask, what is it I am meant to do with this information? What is it any of us are meant to do? Are we meant to be scared of you? Are we meant to sign up?
Starting point is 00:19:47 up to Claude Max? Are we meant to invest in the Anthropic IPO? Are we meant to regulate this stuff? How? Why do you never say that? Why do you never say what it is we're meant to do to stop this? Why do none of you fucking people ever have any kind of suggestion or call to action or anything other than some sort of self-serving or company-serving diatribe about how powerful and scary AI is? And if you hear any loathing in my voice, that's because I find you loathsome. I find what you're doing, disgusting, I think you are cynical, I think you're malcontents, and I think you're bad for society. And the fact you're being elevated is dangerous, but not for the reasons you're saying,
Starting point is 00:20:34 but because it proves that our media ecosystem barely has object permanence. Congratulations on exploiting it. You're a fucking asshole. I think the answers to these questions are actually pretty simple. You don't actually give much of a shit and you want some attention. If you actually feared this stuff and felt a moral obligation to tell people, you'd be specific about what it is we should be scared of and indeed tell us what we need to do next. Every time it's the same sorry fucking story.
Starting point is 00:21:06 Oh, we're building something so scary and powerful and we can't stop it. nevertheless we're going to keep building it for some reason we never discuss but when it destroys humanity in some way we can't explain you'll thank us i guess great thank you man thank you for taking up the airwaves thank you from distracting from the real problems you guys are horrible you guys are genuinely kind of evil if you wanted to do something about this you'd do something about this if you cared you'd care you'd be saying we need to regulate in this way you wouldn't be doing this mean mealy-mouthed horse shit of, oh, maybe they'll stop doing recursive self-improvement,
Starting point is 00:21:46 a thing that doesn't exist. Weird that you don't talk about what's happening today. Weird that you don't talk about stopping the hugging face attack of the future. It's always the positive. It's always the positive in even when being negative. And it's hard to see this as anything other than cynical marketing and attention seeking from two guys who should spend more time working on building technology than posting on social media about how scary things are getting. I don't like them. I don't trust them. And the whole thing's a fucking grift. Next week, I'll be back with Cal Newport and Adam Becker to talk about the people that
Starting point is 00:22:20 inspired this AI Dumerism, the rationalists, and how to push back against their noxious bullshit. I'll catch you then. I love you all. Hi, this is Kylie Breakman. Angel Geratana. Jeremy Colhane, Patrick MacDonald. And we're the hosts of the Artists on Artists on Artists podcast. The improvised character comedy podcast where each week a new panel of artists discuss their craft. Join us for a panel of
Starting point is 00:22:51 Eurovision participants. And we do a little Halloween sometimes. Does the Pope dress up? He winks on Halloween. He gets right up next to the microphone. I hear you can almost hear the eyelashes hit each other. You never seen Italy so quiet as when the Pope gets
Starting point is 00:23:07 up to wink on Halloween. Listen to Artists on Artists on Artists every Thursday on the IHeartRadio app. Apple Podcasts or wherever you get your podcasts. Didn't catch the latest Roland Martin Unfiltered Podcast? Here's what you missed. People wake up and go, oh damn, wait a hold up. They change all of that?
Starting point is 00:23:25 Yes. It's real. This is a wholesale attack. It is targeting black people in every federal agency. It's raw. White folks have never allowed that reckoning to last more than a decade. Catch Roland Martin's daily commentary on the Black Information Network. And download Roland Martin Unfiltered on the IHart Radio app
Starting point is 00:23:43 Apple Podcasts, wherever you get your podcasts. The new NFL season is here, and you should be listening to NFL Daily as we march along to Super Bowl 61. It is in the name, NFL Daily. You'll have fresh content in your feed every day all season long. That game-winning drive, maybe it's Herbert, maybe it's Gino, maybe it's Mahomes. We'll have the highlights. If Fernando Mendoza or any of the rookies are balling out, we'll break down the tape.
Starting point is 00:24:08 Join me, Greg Rosenthal, and an all-star cast of co-host as we preview and recap. every game. Listen to NFL daily on the IHeart Radio app, Apple Podcasts, or wherever you get your podcast. Hey, it's Bobby Bones. Join me and former NFL quarterback Matt Castle every Wednesday on our podcast, lots to say with me, Bobby Bones and Matt Castle. You're in training camp and the rookie quarterback has one good throwing session in front of the media. Suddenly everybody on line says he should start over you week one. How do you handle this? You just go back out to practice the next day. I wait firm to mess up. Listen to lots to say with Bobby Bones and Matt Castle on the IHeart Radio app, Apple Podcasts, or wherever you get your podcasts.
Starting point is 00:24:51 This is an IHeart podcast. Guaranteed human.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.