首页 > AI前沿 > What’s behind the AI industry’s latest warnings of doom?

What’s behind the AI industry’s latest warnings of doom?

TechCrunch 2026-09-14 03:40 2 阅读 查看原文

The AI industry seems to be having its loudest debate yet about whether its technology poses an existential threat to humanity.

The current discussion began after AI researcher Jacob Coxon said that he’s resigned from Anthropic because he’s worried that the leading AI companies are “gambling with our lives.” Then Anthropic’s alignment leaned chimed in with a post declaring, “We really do earnestly believe AI could kill all humans!” adding that he personally thinks the chance is “>10% within the next decade.”

On the latest episode of TechCrunch’s Equity podcast, Kirsten Korosec, Sean O’Kane, and I discussed the latest apocalyptic warnings. I tried to articulate why I’m skeptical of many AI doomer narratives, while Kirsten asked if this was “just a weird way of flexing to show how far advanced their company’s AI model is,” particularly as these companies prepare to go public.

And Sean wondered how these concerns might show up in Anthropic’s S-1 filing for its IPO: “Are there junior lawyers right now who are going through and having to rewrite that entire section of the S-1 filing to say, ‘It’s a officially Anthropic’s position that there’s a more than 10% chance that we could develop something that would eradicate all of humanity and that would be materially bad for our business’?”

Keep reading for a preview of our conversation, edited for length and clarity. (Note: We recorded this episode before Anthropic CEO Dario Amodei published his plan for more cautious AI development.)

Sean O’Kane: I’m hard-pressed to think of something that blew up so fast. Not only did this warning shot come out from this young researcher who has also worked at OpenAI, but also was immediately shared on X by the alignment lead at Anthropic — who, in what might go down as one of the best misplaced exclamation marks ever, shared Coxon’s post and and thread and said, “We really do earnestly believe AI could kill all humans!” Exclamation mark! 

What a weird vibe. That was just a ton of accelerant on an already fraught post or series of posts. Coming after the Hugging Face hack from OpenAI’s internal model, plus just the increased capabilities we’ve seen with the latest models released by Anthropic and and now OpenAI with with Astra a few weeks ago, I think this was just perfectly timed to be a powder keg type of thing for this young researcher to say.

Anthony Ha: Just to disagree with you, I do think that if you believe that AI could destroy all humanity, that does deserve an exclamation point. I would argue that that is a perfectly well used exclamation point!

My issue with that tweet was more the “we.” Who is the “we” here? To what extent can we talk about sort of the AI community or AI research community as a monolith? And the greater than 10% chance — that’s just a made up number, that doesn’t mean anything. Sometimes [there is] this habit in both the tech industry and other places to just throw out these percentages, they’re not based on anything or calculated based on anything. [In retrospect, I realize the tweet was probably referencing the concept of P(doom), but I still think it’s silly.]

One thing I will say about Coxon’s statement and decision is — there’s this recurring theme on Equity, when someone like Sam Altman or Dario Amodei is doing this doomer narrative, there’s always this element of: Well, then, why are you doing what you’re doing? If you actually believe that [AI could destroy humanity], you would not continue doing this. 

[Whereas] this is actually somebody putting his professional trajectory where his mouth is. He’s actually saying, “I believe this is really, really, really bad, and I don’t want to keep working on it.” And so, props for having the courage to do that, if nothing else.

Kirsten Korosec: Yeah, I put him in a separate camp than everyone else saying that and talking about the dangers.

I’m going to put my speculative hat on, because I want to ask both of you a question, which is: Is it possible that every single time we see the increasing number of blog posts about yet another incident in which one of their AI agents breaks through unintentionally, or they talk about how humanity is at risk, is this a weird way of flexing to show how far advanced their company’s AI model is?

I mean, that sounds very cynical, but it does achieve that purpose. Which is: If these AI models weren’t advanced and weren’t capable and weren’t breaking through, we wouldn’t have to worry about these things, right? It’s like a very weird way to brag about the capabilities of the models that you’ve created within your own company.

Anthony: I’ve definitely wondered about this. I don’t think it’s completely cynical, in the sense that I don’t think it’s all just a very conscious marketing ploy across the board. I think that when a lot of these people — whether the researchers or CEOs — talk about it, they do have real concern.

But of course, it does align with [their] business interests in a lot of ways, to say, “Wow, we’ve built the most deadly software that’s ever been made.” I don’t want to get too psychoanalytic here, but others have pointed out that there is this temptation on a personal level of: Of course, you want to believe that the thing you’re working on is the most important and most dangerous thing in the world.

Sean: The thing that sticks out in my mind when I think about that question is, there’s certainly an element that makes it seem like, “Okay, we’re doing this thing that’s so capable, and that’s good for us in some way, even if it looks bad in a lot of different lights.”

I think what’s different about some of these most recent examples is, it really gives you the feeling that these companies don’t have a handle on this stuff in certain ways, especially with the OpenAI stuff.

We keep seeing more and more reporting about other internal agents that have accessed different wikis on the web and are leaving messages for each other, and in a way that doesn’t seem like it’s being handled in a competent way from OpenAI. I would imagine there would be just a bit more polish on the story being told, if it was wholly about getting people to believe that, “Oh my gosh, they’ve made something so incredibly capable.”

The other thing that I think is really fascinating about this, in particular, [is] we’re what, a few weeks at most out from seeing Anthropic’s S-1 filing for its IPO, and just a couple more weeks or month or two away from a potential IPO.

And the idea that you’re going to come out and say these things in this clear language ahead of an IPO — I’m very interested in what that means for that process. How much of this kind of stuff had they already written into the S-1 and the risk factors inside that document? Are there junior lawyers right now who are going through and having to rewrite that entire section of the S-1 filing to say, “It’s a officially Anthropic’s position that there’s a more than 10% chance that we could develop something that would eradicate all of humanity and that would be materially bad for our business”?

Kirsten: You’re assuming that it’s not in there already.

Sean: That’s what I’m saying, though: Is it in there already and being reworded? Or is this something that’s a true scramble? There has to have been language in there. It’s one of the reasons I’m so eager to read this document in a way that goes even further, in some ways, than the SpaceX [S-1], because I’m sure that there’s probably stuff specific to these ideas that will be interesting to see.

Kirsten: Here’s the thing: In a traditional investment environment, one might believe that language like this would hurt the valuation of a company, because it’s suddenly dangerous. But we don’t live in normal times.

And so again, back to my point, it could end up being a weird beneficial flex for the company on the valuation side. It’s not the same as the whole rage-baiting trend that we saw last year, but it’s in that same, let’s say, universe, in which the strength, capability, even elements of danger of something, equals high valuation. So I guess we’ll see in a few weeks.

Putting that aside for a minute, what is being done about it? And can we control this? Tthe U.S. executive director of a nonprofit called ControlAI, Connor Leahy, he was on the show this week, talking about this. So what are you paying attention to in terms of how to control the dangerous aspects of AI, or are we throwing up our hands and watching it all unfold?

Anthony: I myself do not necessarily have a great answer to this, but I have been thinking about some aspects of this debate and maybe why I respond the way I do. 

To echo one of Sean’s points, I do think that part of what this speaks to is the extent to which these major AI companies are feeling like they’re not really in control of these models anymore. That’s definitely not great. That is something that we should all be worried about. 

I do think that part of the reason I’m skeptical of the doomer narrative or resistant to the doomer narrative is because it reaches this level of hysteria of, “Wow, this could destroy humanity in the next 10 years.” It is a little bit of a distraction from the more immediate harms that AI can have, whether that’s labor-related, whether that’s environment- and climate-related. 

Ideally, I think we should be able to discuss all of these things, and have regulatory and other kinds of safeguards against all of these things [including AI’s existential threat]. But once you start using phrases like AGI and superintelligence, that just sucks up all the oxygen in the room in a way that is not very helpful.