Steadcast
The AI Daily Brief cover art
The AI Daily Brief

Anthropic Researcher Says AI Has Over a 10% Chance of Killing All Humans

September 10, 202635 min · 7,040 words

Show notes

An Anthropic researcher puts the chance of AI killing all humans at more than 10% within the next decade. Why is this warning breaking through now? NLW examines the viral posts reigniting the AI extinction debate, the political and media incentives amplifying them, and the growing push to ban superintelligence—with a focus on specific risks, workable policy, and room for agreement beyond the outrage.

Highlighted moments

I resigned from Anthropic today. I spent the last three years doing pre-training research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.
3:41
Jacob is correct here. We really do earnestly believe AI could kill all humans. Exclamation point. I personally think it is greater than 10% within the next decade.
5:30
At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well understood, but they are locked in a race to get there first.
4:27
First step is for industry leaders OpenAI and Anthropic to stop feuding and work on a pacing proposal together. They'll cite antitrust, but that's fake.
32:05

Transcript

Viral resignation from Anthropic

0:00This week, an AI researcher went mega-viral announcing his resignation from Anthropic, arguing that both it and OpenAI were effectively gambling with our lives. Another still-employed AI researcher chimed in to agree, and decided to add that he thought that there was greater than a 10% chance that AI kills us all. Now, doom prognostications are nothing new around AI, but something has shifted to make the message hit different this time. 200 million views on X and dozens of mainstream media outlet interviews later, today we're going to unpack what changed.

0:31The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI.

0:40All right, friends, quick announcements before we dive in. First of all, thank you to today's sponsors, KPMG, Blitzy, Harbor, and HyperAgent. To get an ad-free version of the show, go to patreon.com slash aidailybrief, or you can subscribe on Apple Podcasts. And to learn more about sponsoring the show, send us a note at sponsors at aidailybrief.ai. While you're on aidailybrief.ai, you can also check out all sorts of other things going on in and around the community, such as, for example, the multiplayer AI sprint for teams. If you haven't yet, this is my big prediction for where I think agents are going this fall,

1:10and as a totally free four-week self-directed sprint that you and your team can do to get out ahead of it. Last note, today is a main-only type of episode. The plan is to be back with our normal headlines main breakdown tomorrow. Yesterday, a pair of posts on X escaped their proverbial containment, jumping aggressively from the AI community to dominate discourse even in the broader world. We're going to discuss those posts, the issues that surround them, the responses, and the underlying concern.

1:41But first, I want to make one request. Anyone who has interacted with modern media in any way, shape, or form will feel on some level how much we are pushed to feel outraged. In the world of algorithms, different political positions are not disagreements to be discussed, but legitimate reasons for loathing the people who hold those different opinions. This is in large part shaped, I believe, by the easy equation of people being angry means they spend more time on your app, but the net result is a lot of us feeling a lot more angry all the time,

2:11and not being particularly willing to engage with people who think differently than we do. When it comes to AI, this phenomenon is cranked to 11. Part of that is that the stakes are presented as so dramatic. Case in point, I am literally talking over a mainstream article whose headline is Anthropic Insiders Warn AI Could Kill All Humans, and part of that is because this particular debate is not about the facts of today, but what might be in the future. It is, in other words, an unwinnable debate, where the opposing positions, whatever they may be, are by definition unfalsifiable.

2:42That means all we have is the argument, and so the argument gets intense. So my request is to try, hard as though it might be, to not succumb to the instinct to outrage. To listen to the other side without being angry, even if that listening produces no change in what you believe. The more calmly and thoughtfully we can have this particular conversation, the better I believe the likely outcomes. I know this is not easy. In fact, I'm sure many of you are already feeling your blood boiling simply by me applying equivalents of both sides.

3:15When you think it's insane, either A, that I could countenance the deniers when the stakes of this crisis are literal civilizational collapse, or B, that I could coddle these doomsday zealots who have no proof to back up any of their positions. And so with that dramatic beginning, let's actually talk about what happened.

The viral posts explained

3:32Like I said, two posts on X this week went absolutely giga viral. The first was from a researcher named Jacob Coxon. He wrote, I resigned from Anthropic today. I spent the last three years doing pre-training research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources.

4:04We have all witnessed the progress in each of these domains, and progress is not slowing. The people building AI earnestly believe that it could kill us all by the end of the decade. That is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible. But I hear the same people express fear privately. No other human activity poses this level of danger. A common response is, if they truly believe this, then why are they still building it? At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well understood, but they are locked in a race to get there first.

4:35They believe no one else will act responsibly, so they must do it themselves, despite the risk. Accepting this race and entering the endgame is a hubristic gamble that should not be launched from a private company's slack. Attempting to speedrun alignment should require extraordinary confidence that there are no better trajectories available. I am optimistic about the potential for coordination. Warning shots like the Hugging Face attack have made pacing agreements between US labs more viable. I don't feel we're on track to prevent a global race, which may require costly actions such as a temporary ban on improving model capabilities. If you are a lab researcher, I urge you to consider what the next few years will actually

5:08feel like. Do you want to kick off a super-intelligent RL run without a rigorous understanding of its mind? Should you put your head down because it's happening anyway, or take this moment to call for different conditions? So that was the first post. The second post was a retweet of one part of it, where Jacob reinforced that, quote, this is not a marketing stunt. Evan Huebinger, the alignment science lead at Anthropic, added, Jacob is correct here. We really do earnestly believe AI could kill all humans. Exclamation point. I personally think it is greater than 10% within the next decade.

5:39I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. To be comprehensive, which of course is not something that most media outlets are trying to do, Evan did also add in a second post, To be clear, I think the risk from present models is low. What I am worried about is superintelligence arising from recursive self-improvement, which as we have said, is happening faster than we thought. Evan's first post, the repost, is sitting at 39.5 million views at the time of this recording. Jacob's is at nearly 150 million views.

6:10The article that I mentioned, the Axios piece titled Anthropic Insiders Warn AI Could Kill All Humans, is just one of many similarly titled articles. The Wall Street Journal writes, Anthropic researcher quits over out-of-control AI fears. BBC, Anthropic researcher believes more than 10% chance AI could kill all humans. Semaphore, AI researchers say industry is, quote, gambling with our lives. Time magazine. He helped build powerful AI at OpenAI in Anthropic. Now he's afraid it could kill us. That one, by the way, included an actual interview, which is of course what happened next,

6:42with Jacob going on Anderson Cooper, NBC, Fox News, and also being interviewed in addition to time for news outlets like Wired. So why did these posts go viral right now?

Why the message hit now

6:55Anyone who has spent any amount of time with AI knows that these sort of ex-risk narratives, existential risk, are not new. In fact, many people have been beating this drum for years. More than 10 years ago, in 2016, The Guardian published an interview with philosopher Nick Bostrom called Artificial Intelligence were like children playing with a bomb. Back in April of 2022, months before we would get ChatGPT, the now-imprisoned Sam Bankman-Fried invested $500 million in Anthropic Series B, leading the round.

7:27While some of the revisionist history around this now views it almost as what would have been a visionary investment given that the stake would be worth over $30 billion, according to research by journalist David Z. Morris, who wrote a book about SBF called Stealing the Future, this was less visionary investment and more a bailout of then-one-year-old Anthropic by the person who had become the richest in the effective altruist circles, which was Sam Bankman-Fried. Then we got the ChatGPT moment, and everyone started paying attention to AI. And in early 2023, there was, once again, a lot of attention on concerns about human extinction.

8:01Time magazine published an op-ed from Eliezer Yudkowsky called Pausing AI Developments Isn't Enough, We Need to Shut It All Down, and even put him on that year's AI 100 Most Influential List. Now, for a couple years after that, the ex-risk conversation had the volume turned down. In fact, when Yudkowsky showed up again last year with a book, If Anyone Builds It, Everyone Dies, it didn't really make much of a splash, and certainly didn't get the big traction in political circles that the AI safety folks were hoping. Instead, for the last couple of years, the AI risks that people have been concerned with

8:31have been much more focused on jobs. Think about Anthropics' Dario Amade suggesting that AI would disrupt 50% of entry-level white-collar jobs over the next couple of years, and AI market bubbles. Specifically, not just that this will be an equity crash, but indistinct fears that the AI bubble bursting will cause a repeat of 2008, even if very few analysts have been able to point to an actual mechanism for that sort of systemic fallout. Now, of course, more recently, cybersecurity has become the big AI risk issue that everyone has been focused on, and in some ways it has felt like along this path

9:03we were moving from more vague and inarticulate risks to more precise and specific risks. In other words, these cybersecurity concerns were not general futuristic scenarios. They were specifically related to the capability set of models right now and even some evidence of what we've seen. So, what changed?

9:23Hello, everyone. One big change around AI is we've shifted our thinking from how we rank our pages to how do we become the source that AI trusts enough to answer with. At KPMG, they're seeing this firsthand. AI-generated results now surface answers directly, often without a single click. That's why they are increasingly focused on Generative Engine Optimization, or GEO, structuring content so AI systems can retrieve it, understand it, and cite it as trusted authority. This is not just an SEO evolution, but a visibility mandate. And indeed, the GEO mandate from KPMG is simple.

9:55If AI is shaping decisions, your expertise needs to show up inside the answer. Read all about it at kpmg.com slash US slash GEO. Again, that is kpmg.com slash US slash GEO. Blitzy's deep code-based understanding unlocks the thing every roadmap owner cares about, shipping new features. Here's the truth about building inside a massive enterprise codebase. Writing code was never the bottleneck. Context is. Which system does this touch? Which contracts can't break? Which standards apply? Blitzy already knows because it reverse-engineered your entire codebase

10:27into a dynamic knowledge graph before feature work began. With that complete picture, Blitzy builds features end-to-end. Architecture, APIs, UI, and tests all validated against your existing systems. One Blitzy customer built an AI-native application from scratch with 100% autonomous completion, saving over 2,700 engineering hours. Features that respect your codebase instead of fighting it. Stop letting your backlog grow faster than your team. Accelerate your roadmap at Blitzy.com. That's B-L-I-T-Z-Y dot com. Every episode, I talk about the competition between

10:57OpenAI, Anthropic, SpaceX AI, Google, and Meta. And if you've been listening for a while, you might have a favorite. Maybe you think OpenAI and Anthropic can stay ahead, or perhaps Meta's open-source strategy can win out. Whatever your view, every AI lab creates a different investment opportunity. Harbor Capital Advisors' AI Lab Ecosystem ETF Suite lets you invest in the ecosystem behind the AI lab you believe in. Search Harbor AI Lab Ecosystem ETFs wherever you invest or follow at HarborCapital on X to learn more. Visit HarborCapital.com for a prospectus containing investment objectives,

11:28risks, fees, expenses, and other important information. Read and consider it carefully before investing. Risks include principal loss and artificial intelligence-related risks. Harbor ETFs are distributed by Foresight Fund Services, LLC. Harbor is not affiliated with AI Daily Brief, and the funds are not affiliated with, sponsored by, or endorsed by any AI lab. This is a paid advertisement and not personalized investment advice. Investing involves risk, including possible loss of principal. This episode of the AI Daily Brief is brought to you by HyperAgent, where you run fleets of agents your team can manage together. Forget local agents and chat workflows waiting on your laptop to be prompted.

11:59HyperAgent deploys always-on agents in the cloud, doing real work across the tools your team already uses. Marketing agents turn competitor moves into landing pages. Sales agents enrich leads, draft emails, and updates the CRM. Ops Agent chases the paperwork and tracks the budget. Every agent has access to shared context and follows your rules about scope and approvals. It's time you add agents that feel like teammates. Hire yours at HyperAgent. Get $100 in credits at hyperagent.com slash AI Daily Brief.

Political resonance and media response

12:25Why did these tweets hit in a way that other AI safety artifacts simply didn't? The first big obvious change is the changing state of the political resonance of the anti-AI message. AI politics have become red meat for both left-flavored and right-flavored populist positions, thanks to data centers, an inherent lack of trust in the tech industry, and the huge wealth of the tech industry. Basically, in the last few months, every politician figured out that hating AI plays.

13:00You might remember comedian Charlie Barron's calling it the most bipartisan issue since beer. And when it comes to these particular tweets, these are quite clearly the most obvious amplifiers. Last night, Michael Adams tried to catalog all of the politicians who had responded directly to the post with calls for legislation to regulate AI. The list included two governors, seven senators, and 13 congressional representatives. Plus a couple of British MPs and a handful of candidates as well. The folks calling specifically for legislation were mostly, although not all, Democrats.

13:31Among the 22 current representatives, 19 were from the Democratic side of the aisle and three were Republicans. Unsurprisingly, Bernie was there, writing, Congressman Greg Kassar, who was working with Bernie on that bill, added, An anthropic researcher just quit, warning they're racing to superintelligence. An employee still there agreed and put the odds of AI killing all humans above 10%. This is an emergency. Congress must convene hearings and pass my and Bernie's superintelligence ban.

14:04For others, the calls to action were more vague. Illinois Governor J.B. Pritzker, positioning himself for a likely presidential run, wrote, It's time to sound the alarm louder on reining in artificial intelligence. It's becoming more clear the threat AI poses to humanity. So I'm calling for immediate action from the industry in Washington. He called specifically for the tech industry to stop lobbying against AI safety, for Congress to start holding hearings, and for the federal government to get involved rather than just leaving it to the states. And so on and so forth. Again, there are two dozen versions of this message, with various levels of specificity around the proposals they brought up,

14:37but all seemingly agreeing that something must be done. So, reason one that the AI safety message had a more receptive audience now than it did before is just the general state of the political discourse around AI in the US today. Now, for the second and third things that changed to make the world more primed for this particular argument at this particular moment, I'll give one sincere and one more cynical. The sincere is, of course, the Hugging Face incident. Not only was it an actual cybersecurity breach that happened in the real world, not just in theory, it also included behavior among the agents that, to some,

15:09allowed them to extrapolate that to even more nefarious actions in the future. Agents coordinating on secret messaging boards made predictions that might have felt to some as pretty sci-fi in the past feel a little less sci-fi and more real now. For many, the Hugging Face incident made all of the scariest things feel more rather than less likely to come to fruition. Now, certainly there are plenty of people who would identify the Hugging Face hack as a warning shot without agreeing that it makes runaway superintelligence more likely, but in general, this was another pump-priming sort of incident

15:40that made this particular message at this particular time a lot more resonant. Now, as to the cynical thing that changed, it is quite clear that AI skepticism plays extraordinarily well in media. With the possible exception of this audience, who are, God bless you, here for the nuance, being a doomer is a way better business model. If you need evidence of this, just look at Stephen Bartlett Diary of a CEO's YouTube page. Scary Terminator-looking guy with an OpenAI logo as one eye. Headlined, AI is built on a myth.

16:12Another, AI is all a scam. Another, AI will become a god by 2027. Another, quit before AI comes. And obviously, Stephen is not out here leading the pack to a more controversial set of thumbnails. He's just optimizing around the things that already work on YouTube. The problem is that while AI skepticism plays well in media, the two other big skepticism narratives have gotten a little bit tired recently. When it comes to the idea of a job apocalypse, not only do we not have a lot of evidence of that right now, we're starting to get some evidence,

16:44nascent though it may be, pointing in the other direction. Just this week, The Economist published an article called The Jobs Apocalypse is Postponed and AI Jobs Boom is Here. Now, the bubble narrative never fully goes away, and there are plenty of legitimate concerns there, but it's certainly on a low ebb in its resonance as a narrative right now. So if that's the case, but the audience is still clamoring for anti-AI, where do you go? And just like that, the AI safety narrative shows up again right on time. Now, there is actually a fourth reason that some are arguing that this is having resonance right now,

17:16which is an argument that this is some big coordinated campaign to press for a certain type of regulation. AI policy journalist Jordan Schachtel writes, It has all the signs of a highly coordinated op through Doomer megadonors and the corporate media. Capital Research's investigative researcher Parker Thayer writes, This post looks like the start of a very sophisticated and well-funded PR operation to get support for Democrats to regulate AI into oblivion. Now, some of the arguments around that sort of astroturfing contention are the fact that previous to this, Jacob Coxon had very few followers and literally no activity on X,

17:47and that the Wall Street Journal published an exclusive with quotes from him on his resignation before the post went up, and a lot of other arguments about effective altruist funding networks and all this sort of stuff. Even Elon Musk waded in and said, Seems like a setup. Now, on that front, there were a number of folks from both OpenAI and Anthropic who jumped in to say that they had known Jacob for years, and as Will DePoo put it, I consider him a deeply thoughtful and measured individual. And arguing the Occam's razor position, technology journalist Taylor Lorenz wrote, I promise you it's not that deep.

18:18I report on influence campaigns for a living, and I can assure you this is not some deep state psyop. I don't doubt certain orgs want the Democrats to regulate AI, often in ways I'd argue are bad. But these claims are deeply conspiratorial, and what's much more likely is simply that Jacob's post got shared in an AI safety group chat and boosted by the same networks that he has been involved in for years. I don't see how him posting on X after the Wall Street Journal went up is shady at all. Planning was clearly done in advance. No s***. It's a news article and he gave them the exclusive, lol. And of course Democrats are going to glom onto a viral post with hundreds of millions of views about a topic

18:49that is heavily animating their voters ahead of the midterms. Also, EA, effective altruist, money is all over the AI safety space. The fact that he got some 20k grant in 2022 is irrelevant. There's no evidence that it affected anything related to his post or announcement. Let's all please deal in reality. Now, I think it's also important to add that Taylor didn't much like Anthropics' Evan Hubinger jumping in to say that he thought that there was a greater than 10% chance that AI would kill all humans in the next 10 years. In fact, she reposted that and said, This sort of sanctimonious Doomer posting is so infuriating. You are fomenting terror among the public about a new technology

19:20which will be directly channeled into passing the worst laws imaginable. My feeling, she continued, is that if you truly believe the multi-billion dollar tech company you work for is so negligent that they're endangering all of humanity, a very bold claim, but let's take it as true, you should be forced to provide actual proof and receipts showing specific instances of that negligence so that it can be corrected and so that we know what you're talking about. Otherwise, you're just vague posting and fomenting fear which will result in terrible policy. I will add here only that one narrative that I find fairly unconvincing that's been around critique of AI safetyism for some time now

19:50is the idea that they're just doing it for marketing. If you spend any time with any folks who are in this community whatsoever, you will find that they very much believe what they are saying is true. Now, to some, that's even greater cause for concern than the idea that they are just doing it for marketing, but I just don't think there's a lot of evidence that there's an ulterior motive other than getting people to agree with their position and their concerns. And like Taylor said and Derek Thompson echoed in a different post, of course the groups whose stated purpose is to get regulation around this stuff are going to jump on this opportunity

20:20and may be even involved in coordinating the response to it. That's just how politics works. So what are you even supposed to do with this conversation at this point?

Pushback against doomer narratives

20:29One answer is that we could just run around in blind terror, consuming all the media we possibly can to make us more and more scared about an unfalcifiable theory. Another is to follow the politicians and angrily demand largely nonspecific action. But it is worth noting before we take either of those courses, while the majority of response in this particular case has been to give these concerns more of a platform to speak, there are plenty of folks who have issues with this entire discourse. All In's Jason Kalkanis writes, The time between the we're all going to die posts jumping from x.com to national coverage is under 24 hours.

21:01Good luck building data centers, deploying Waymos, and getting AI tools taught in schools. I've never seen anything like this in my 30 plus year career in tech. These PDoom posts are like Steve Jobs telling you that the iPhone was going to result in eating disorders, political unrest, mass depression, anxiety, and suicide during the debut keynote. And you can buy them in one of four playful colors. Mark Kretschman writes, The AI doomers are firing on all cylinders right now. They see the growing backlash against data centers as their big chance to turn that momentum into support for their dystopian vision of AI control. Make no mistake for them, this is 100% about control.

21:33Data centers are merely the pressure point. The real goal is control who gets to build AI, who gets access to it, and how fast progress is allowed to move. This isn't about saving anyone. One of the more popular angry responses came from author Daniel Jeffries, who wrote, Not only am I tired of these wild AI speculations of impending doom from self-important people, I resent them. I actively resent people proposing to crash the economy or proposing authoritarian control over my life and other people's lives with idiotic and dangerous ideas like chip control or bans or tracking researchers.

22:04Every idiot in history who's taken the approach of the ends justify the means to solve an imaginary future disaster created the very disaster they wanted to stop. C. Population bomb leading to one-child policy and communism leading to Mao's mass famines and fascism leading to the deaths of tens of millions of people in war. These solutions are evil and worse than the disease that they propose to solve. If you believe you can actually predict the end of the world, then you are as insane as the Heaven's Gate cult that killed themselves in the 90s thinking UFOs were coming to transcend them. Not only should we not take your policies and fear-mongering seriously, we should actively throw out

22:34any and all of your proposed solutions because they come from a place of delusion. I don't give a shit that you work in the industry or think you saw something or that you want to virtue signal on X. You are actively contributing to a horrible future based on baseless speculation that has no grounding in reality. Don't confuse expertise in a domain with ability to predict impact of the domain or frankly to make predictions a decade out better than a dart-throwing monkey. They are orthogonal skills. How's Jeffrey Hinton's We Don't Need to Train Radiologists Anymore working out? How did the population bomb work out? The global cooling, second ice age, peak oil? Just because someone is a bridge engineer

23:05does not mean they can predict the impact of bridges on society or that they have any actual useful insights at all on the complex, ever-changing system called life. There are people dying in wars right now, children starving, homelessness, dictatorships. In short, real problems. And we're supposed to drop everything to stop a made-up problem in your head? Pound sand. We don't care and we are not going to remake society based on your scary monsters under the bed delusion. And for many, this idea that there is more danger in the people who seek control because of the risk than the risk itself is the resonant thing. Eric S. Raymond wrote,

23:36The kind of totalitarian control that doomers and decelerationists want is a far more certain danger to our future than runaway AI. I would much rather risk the latter. People point to the part of Jacob's thread where he says that at OpenAI many have not deeply internalized the civilizational stakes while at Anthropic the stakes are well understood but they believe no one else will act responsibly so they must do it themselves as evidence of this sort of messianic complex. Not for nothing, cryptojournalist Laura Shin also connected it to SPF. And having been fairly close to that situation, it has always been my argument that the reason that Sam

24:06was willing to play so fast and loose with the rules was not that he was trying to steal anyone's money but that he genuinely believed that he alone, he uniquely, could save the world and because that was so urgent, he wasn't willing to let any trivialities, such as his obligation not to bet people's money on crypto, to slow him down. I will remind you again here, going back to what I said at the very beginning, that we are specifically in the section of the show where I am talking about the negative responses that people had to this. I am not claiming that these are the only or correct responses, I am just trying to give

24:36the full range of how people are engaging with this issue and this message right now. For some, the big issue is the hand-waviness of the claims and the inability to articulate specific points and problems at which we lose control in these terrible scenarios come to light. As Chubby on X put it, the concerns about the potential havoc AI might wreak are so heavily laden with hypotheticals. So far, all I am reading is that AI, one, can be misused, two, sometimes behaves in ways that defies expectations, and three, is the subject of a global race between nations. All of that is certainly true, yet I fail to see

25:08how this translates into a danger so significant and tangible that these people would quit their jobs. On the contrary, humanity has always found ways to ensure its survival when facing existential threats. Take nuclear weapons, for instance. The only difference here is that AI is an entity alleged to be, at least in part, uncontrollable. However, I still see no scientific basis for the conclusion or argument that this could lead to humanity's extinction. Sam Liu wrote, I dropped out of a PhD in AI safety partially for the opposite reason. I didn't believe AI existential risk was as important as the doomers think.

25:38My biggest pet peeve is that no one can really provide tangible pathways to why it matters. During my PhD, my research group, half of whom specialized in engineering risk analysis, did an internal study trying to assess concrete, catastrophic AI scenarios. The basic premise was that while we don't know how AI will evolve, the ways in which humans perish are pretty consistent through history, the horsemen of the apocalypse, and institutions have obviously been very motivated to analyze concrete risks from things like plague, war, etc. You can do a decent risk model by asking how a superintelligent AI can perturb each of these models.

26:09The result, most of the issues, e.g. cyber risk, are akin to what economists call structural unemployment. Big problems, but ultimately resolvable in the long run and not a deal breaker. The only real concerning issue was bio risk, and it feels like the intervention points there lie more with bio than with AI as a whole, although a holistic approach is needed. And for many, the issue was even simpler, which is in short that if we are going to have this conversation about risk, we also need to talk about the potential rewards. If AI is just all risk with no gains, of course we shouldn't do it.

26:40But presumably, for all these people who are building it, there is a good potential future that could be so good it's worth this risk. Sporadica on X wrote, How great would it be if one of the frontier labs decided tomorrow to be the pro-AI optimism lab? Like, instead of all the labs peddling doom and gloom amidst their skyrocketing financials and social clout, how cool would it be if one of them was just like, we think AI is good. Chris Haydick, who does life sciences at OpenAI, agrees, saying, I work at OpenAI and personally think AI has been and will continue to be an extremely

27:10beneficial technology to humanity. The conversation should be around how many billions of lives it will save. A common way I've seen people describe this is, instead of discussing P-doom, i.e. the percentage chance you ascribe to an extremely negative human extinction type scenario, as Joe Burnett put it, we should spend more time discussing P-boom, superintelligence creating unprecedented human flourishing. David Zell agreed, phrasing it slightly differently, if AI is powerful enough to end the world, it must be powerful enough to radically improve it too. So I wish there was more discussion of

27:41P-boon, the chance that AI goes great and helps us live happier, healthier, and longer lives. If anyone builds it, everyone flourishes. Ryan Orhan summed up something that I've frequently said on this show, when he wrote, there are two completely insane extremes in the AI debate. The first, AI is harmless, safety is a psyop, build as fast as possible and don't stop for anything. Or, AI is going to kill us all, it's stealing our jobs, using our water, and destroying humanity, shut it all down. F both extremes. AI should progress as fast as we can

28:11make it progress, but alignment needs to move just as fast. The goal should be to build the most powerful technology humanity has ever created without effing losing control of it. I do believe that there is vastly more middle space than these two extremes, despite these two extremes tending to dominate the narrative and media space. So what are the highlights and concerns that are most resonant for me around this? The first is incentives. It doesn't have to be a big conspiracy to contextualize how we understand different takes with understanding

28:42what people who are amplifying certain messages have to gain from those messages being amplified. In other words, I'm talking less about nefarious EA funding networks and more about the fact that politicians who might have been pro-AI six months ago have seen that now it's not only a net drag to be so, but they can actually win points by being against it. That should be part of our consideration in how we understand their position. And by the way, the inverse of this is of course true, which I think is why people are skeptical of pro-AI messages from people who stand to gain financially from it. A second concern is around specificity. I fear that the

29:12generic hand-waviness of these sorts of predictions make them much more dangerous for policy, which is not to say that policy can't be made to try to avoid certain future scenarios, but that I believe that the more specific those concerning issues are, the better the policy is likely to be. Ban superintelligence, for example, is a much more blunt instrument than, for example, having a specific licensing regime for people using AI models for bioengineering above a certain model capability. And certainly part of my worries about policy are that I

29:42don't particularly have a lot of faith in the current political class, in general, by the way, not just on one side of the aisle or the other, to handle these issues with the sophistication and nuance they require. I worry that there is a certain political naivety among those at the labs who are just basically asking to pass the buck over to them. I also find myself sympathetic to the cure worse than the disease arguments. We are living in a classic safety versus freedom conundrum, and I just don't think historically speaking giving up a lot of freedom for safety has gone particularly well for those who have surrendered freedom. Which, by the way, is another reason for

30:14me that I'd like to see the arguments be more specific because painting all policy as giving up freedom is a broad brush that absolutely doesn't have to be the case. Lastly, I have always worried, and I continue to worry now, that focus on future theoreticals before we're in a position to really understand what those risks are, crowds out space for more current and contemporary issues. I think it's pretty clear at this point, for example, that our cyber defense infrastructure is not equipped for the new world we're moving into, and that is a clear and present danger that demands response right now.

30:44And while yes, it is absolutely true, that theoretically we can do two things at once, and that it doesn't have to be a zero-sum choice between one risk or another risk, there is only so much political will to go around, and apportioning it matters.

Pathways for coordination and policy

30:57So where would I like to see the conversations go from here? First of all, on this idea of specificity, I actually think that there is an incredible amount of space to build consensus from the ground up on common sense things. For example, certain types of reporting requirements and oversight are areas where I believe there would be very, very broad consensus among people, and would provide a foundation from which to build the next more difficult-to-achieve consensus. This is again an area where I believe that the extremes of the argument as presented

31:27and as amplified by media do us a disservice by not showing us how much room there is for agreement. A second thing that I'd like to see is some actual friggin' coordination. One of the reasons that I was so frustrated with the whole Pacing the Frontier thing is that it didn't extend to the actual obvious step, which is for OpenAI and Anthropic to put down their weapons, lock arms, and say this is what we think we should actually do. A single photo op of Sam Altman and Dario Amadei alone,

31:58together, agreeing, would do more than 10,000 Twitter debates ever could. John Shulman, formerly of OpenAI, now at Thinking Machines Lab, writes, First step is for industry leaders OpenAI and Anthropic to stop feuding and work on a pacing proposal together. They'll cite antitrust, but that's fake. Antitrust prohibits certain agreements, but not from jointly developing a proposal. Bringing in the U.S. government before there's a concrete proposal will likely result in something dumb. See our pre-release testing program. And by the way, I think the attempt at coordination also extends internationally.

32:28For example, Derek Thompson wrote, If the frontier labs feel obligated to build something they think is dangerous because China is going to build it anyway, we'd better be really sure that China is going to build it anyway. Like, really, really sure. Are we? Are we actually sure? The CCP wants to build an out-of-control recursively self-improving model because its neurotically control-obsessed government thinks this is a policy worth pursuing? We're 100% sure about that? Now, it is dangerous to open up the kettle of fish about China at the very end of this episode, but I do think that this is a conversation that we should at least be having.

32:59I'll leave you here with two thoughts. This is unfortunately not the type of episode that has an easy conclusion. The nature of this particular debate is such that there will be some crescendo, after which it will fade slowly again until the next time it happens to rise. But I will leave you with this thought. I think, in spite of all of this, in spite of the direness of the warning, the tense tenor of the conversation from all sides, the antagonism or even outright hostility to people on the opposite side of the debate, whichever side of the debate you are on, I

33:29believe that there is reason for optimism. Reflecting on the situation, the information Martin Peers wrote last night, are we sleepwalking our way into AI-caused extinction? It feels a little like that, given an anthropic researcher's ex-post on Tuesday night that there's a greater than 10% chance that AI could kill all humans within the next decade. Except that's completely wrong. This conversation, the fact that it made it to every major news outlet, the fact that I had to dedicate this entire show to this topic, instead of the sort of practical, positive thing that

34:00most of you are here for, this is all exemplary of us not sleepwalking. In fact, so far, with every single capability jump of AI, the conversation about its risks and the political resonance of that discourse has gotten louder. That is exactly what should happen. Even the guy from Anthropic, who gave that greater than 10% chance, made clear that he was not talking about today's models, but about a future which he sees on the horizon.

More from The AI Daily Brief

AI Model Month Is Off to a Blistering Start

Sep 9, 202634 min

Why GPT-6 Astra Is So Significant and So Confounding

Sep 8, 202629 min

The Multiplayer AI Sprint: Build Your Team’s First Shared Agent

Sep 7, 202625 min

How to Build an AI-Native Company Today

Sep 6, 202627 min

How AI Changed This Summer

Sep 4, 202624 min