
Whatās behind the AI industryās latest warnings of doom?
The AI industry seems to be having its loudest debate yet about whether its technology poses an existential threat to humanity. The current discussion began afterAI researcher Jacob Coxon said that heās resigned from Anthropicbecause heās worried that the leading AI companies are āgambling with our lives.ā Then Anthropicās alignment leaned chimed in witha post declaring, āWe really do earnestly believe AI could kill all humans!ā adding that he personally thinks the chance is ā>10% within the next decade.ā On the latest episode ofTechCrunchās Equity podcast, Kirsten Korosec, Sean OāKane, and I discussed the latest apocalyptic warnings. I tried to articulate why Iām skeptical of many AI doomer narratives, while Kirsten asked if this was ājust a weird way of flexing to show how far advanced their companyās AI model is,ā particularly as these companies prepare to go public. And Sean wondered how these concerns might show up in Anthropicās S-1 filing for its IPO: āAre there junior lawyers right now who are going through and having to rewrite that entire section of the S-1 filing to say, āItās a officially Anthropicās position that thereās a more than 10% chance that we could develop something that would eradicate all of humanity and that would be materially bad for our businessā?ā Keep reading for a preview of our conversation, edited for length and clarity. (Note: We recorded this episode before Anthropic CEO Dario Amodeipublished his plan for more cautious AI development.) Sean OāKane:Iām hard-pressed to think of something that blew up so fast. Not only did this warning shot come out from this young researcher who has also worked at OpenAI, but also was immediately shared on X by the alignment lead at Anthropic ā who, in what might go down as one of the best misplaced exclamation marks ever, shared Coxonās post and and thread and said, āWe really do earnestly believe AI could kill all humans!ā Exclamation mark! What a weird vibe. That was just a ton of accelerant on an already fraught post or series of posts. Coming after the Hugging Face hack from OpenAIās internal model, plus just the increased capabilities weāve seen with the latest models released by Anthropic and and now OpenAI with with Astra a few weeks ago, I think this was just perfectly timed to be a powder keg type of thing for this young researcher to say. Anthony Ha:Just to disagree with you, I do think that if you believe that AI could destroy all humanity, that does deserve an exclamation point. I would argue that that is a perfectly well used exclamation point! My issue with that tweet was more the āwe.ā Who is the āweā here? To what extent can we talk about sort of the AI community or AI research community as a monolith? And the greater than 10% chance ā thatās just a made up number, that doesnāt mean anything. Sometimes [there is] this habit in both the tech industry and other places to just throw out these percentages, theyāre not based on anything or calculated based on anything. [In retrospect, I realize the tweet was probably referencingthe concept of P(doom), but I still think itās silly.] One thing I will say about Coxonās statement and decision is ā thereās this recurring theme on Equity, when someone like Sam Altman or Dario Amodei is doing this doomer narrative, thereās always this element of: Well, then, why are you doing what youāre doing? If you actually believe that [AI could destroy humanity], you would not continue doing this. [Whereas] this is actually somebody putting his professional trajectory where his mouth is. Heās actually saying, āI believe this is really, really, really bad, and I donāt want to keep working on it.ā And so, props for having the courage to do that, if nothing else. Kirsten Korosec:Yeah, I put him in a separate camp than everyone else saying that and talking about the dangers. Iām going to put my speculative hat on, because I want to ask both of you a question, which is: Is it possible that every single time we see the increasing number of blog posts about yet another incident in which one of their AI agents breaks through unintentionally, or they talk about how humanity is at risk, is this a weird way of flexing to show how far advanced their companyās AI model is? I mean, that sounds very cynical, but it does achieve that purpose. Which is: If these AI models werenāt advanced and werenāt capable and werenāt breaking through, we wouldnāt have to worry about these things, right? Itās like a very weird way to brag about the capabilities of the models that youāve created within your own company. Anthony:Iāve definitely wondered about this. I donāt think itās completely cynical, in the sense that I donāt think itās all just a very conscious marketing ploy across the board. I think that when a lot of these people ā whether the researchers or CEOs ā talk about it, they do have real concern. But of course, it does align with [their] business interests in a lot of ways, to say, āWow, weāve built the most deadly software thatās ever been made.ā I donāt want to get too psychoanalytic here, but others have pointed out that there is this temptation on a personal level of: Of course, you want to believe that the thing youāre working on is the most important and most dangerous thing in the world. Sean:The thing that sticks out in my mind when I think about that question is, thereās certainly an element that makes it seem like, āOkay, weāre doing this thing thatās so capable, and thatās good for us in some way, even if it looks bad in a lot of different lights.ā I think whatās different about some of these most recent examples is, it really gives you the feeling that these companies donāt have a handle on this stuff in certain ways, especially with the OpenAI stuff. We keep seeing more and more reporting about otherinternal agents that have accessed different wikis on the weband are leaving messages for each other, and in a way that doesnāt seem like itās being handled in a competent way from OpenAI. I would imagine there would be just a bit more polish on the story being told, if it was wholly about getting people to believe that, āOh my gosh, theyāve made something so incredibly capable.ā The other thing that I think is really fascinating about this, in particular, [is] weāre what, a few weeks at most out from seeing Anthropicās S-1 filing for its IPO, and just a couple more weeks or month or two away from a potential IPO. And the idea that youāre going to come out and say these things in this clear language ahead of an IPO ā Iām very interested in what that means for that process. How much of this kind of stuff had they already written into the S-1 and the risk factors inside that document? Are there junior lawyers right now who are going through and having to rewrite that entire section of the S-1 filing to say, āItās a officially Anthropicās position that thereās a more than 10% chance that we could develop something that would eradicate all of humanity and that would be materially bad for our businessā? Kirsten:Youāre assuming that itās not in there already. Sean:Thatās what Iām saying, though: Is it in there already and being reworded? Or is this something thatās a true scramble? There has to have been language in there. Itās one of the reasons Iām so eager to read this document in a way that goes even further, in some ways, than the SpaceX [S-1], because Iām sure that thereās probably stuff specific to these ideas that will be interesting to see. Kirsten:Hereās the thing: In a traditional investment environment, one might believe that language like this would hurt the valuation of a company, because itās suddenly dangerous. But we donāt live in normal times. And so again, back to my point, it could end up being a weird beneficial flex for the company on the valuation side. Itās not the same as the whole rage-baiting trend that we saw last year, but itās in that same, letās say, universe, in which the strength, capability, even elements of danger of something, equals high valuation. So I guess weāll see in a few weeks. Putting that aside for a minute, what is being done about it? And can we control this? Tthe U.S. executive director of a nonprofit called ControlAI, Connor Leahy,he was on the show this week, talking about this. So what are you paying attention to in terms of how to control the dangerous aspects of AI, or are we throwing up our hands and watching it all unfold? Anthony:I myself do not necessarily have a great answer to this, but I have been thinking about some aspects of this debate and maybe why I respond the way I do. To echo one of Seanās points, I do think that part of what this speaks to is the extent to which these major AI companies are feeling like theyāre not really in control of these models anymore. Thatās definitely not great. That is something that we should all be worried about. I do think that part of the reason Iām skeptical of the doomer narrative or resistant to the doomer narrative is because it reaches this level of hysteria of, āWow, this could destroy humanity in the next 10 years.ā It is a little bit of a distraction from the more immediate harms that AI can have, whether thatās labor-related, whether thatās environment- and climate-related. Ideally, I think we should be able to discuss all of these things, and have regulatory and other kinds of safeguards against all of these things [including AIās existential threat]. But once you start using phrases like AGI and superintelligence, that just sucks up all the oxygen in the room in a way that is not very helpful.