跳到主内容
@wquguru
精选88Gary Marcus(RSS)行业动态

Gary Marcus:AI在2030年灭绝人类的可能性为零

No, Anderson Cooper, AI is not going to kill all humans by 2030

原文
发到 X
推荐理由

深度拆解AI安全叙事背后的商业逻辑与现实局限,对破除行业焦虑极具参考价值,建议关注AI治理与安全评估的同学细读。

Anderson Cooper looks pretty worried here, as Jacob Coxon, the ex-OpenAI/ex-Anthropic employee who is the media darling of the moment, tells us that the human species might not be here in five years.

安德森·库珀看起来相当担忧,因为雅各布·科克森(Jacob Coxon),这位前 OpenAI/Anthropic 员工、当下的媒体宠儿告诉我们,人类物种可能在五年内不复存在。

I am here to tell you that Anderson Cooper doesn’t need to worry about that particular scenario, and that you don’t either.

我来告诉你,安德森·库珀不必担心那个特定场景,你也不必担心。

§

§

Before I get there, I want to set some things straight. A bunch of people, apparently on the right, are trying to assassinate Coxon’s character. I have seen lies about how long he worked at Anthropic, possibly derived from an inaccurate AI-generated profile of him, and a weaponized thread characterizing him (without real evidence) as a Democratic operative. Someone described him as entry-level employee, which is absurd, given that he had already worked at OpenAI on a technical team for over two years.

在我展开之前,我想澄清一些事情。显然有一批右翼人士试图抹黑科克森的人设。我见过关于他在 Anthropic 工作时长多少的谎言,这些可能源自一份不准确的 AI 生成简介;还有一条被武器化的推文将他(没有确凿证据)描绘成民主党特工。有人称他是初级员工,这很荒谬,因为他此前已在 OpenAI 的技术团队工作了两年多。

In reality, we need a little nuance here; some of what Coxon says is true, some is speculative; some he is in a position to speak to, and some is out of his expertise. (The media too rarely distinguishes).

事实上,这里需要一点细微差别;科克森说的有些是真的,有些是推测性的;有些是他有能力谈论的,有些则超出了他的专业领域。(媒体也很少加以区分)。

What I liked about Coxon’s now-infamous thread is that it cast light on what people at OpenAI and Anthropic are thinking. Since he worked there for a total of about three years (mainly at OpenAI, but most recently in Anthropic), he’s presumably in a position to testify to what might be common (presumably not universal) thinking at those companies.

我喜欢科克森那条如今臭名昭著的推文的理由在于,它揭示了 OpenAI 和 Anthropic 内部人员的想法。由于他在这两家公司总共工作了大约三年(主要在 OpenAI,最近则在 Anthropic),他大概有资格证明那些公司中可能普遍存在的(未必是普遍的)思维方式。

The most important of his thread was the part that went towards state of mind:

他推文中最重要的部分是关于心态的那一段:

rightly accusing the frontier companies of hubris:

正确地指责前沿公司傲慢自大:

Corroborating evidence soon came from Evan Hubinger, a current Anthropic employee:

随后来自现任 Anthropic 员工埃文·休宾格(Evan Hubinger)的佐证证据很快出现:

And others eventually weighed in; OpenAI’s chief scientist actually said something similar. There’s good reason to believe that at least some companies really believe this kind of thing, and also (terrifyingly) that many have come up with justifications for believing that it is ok to continue working on a technology that they themselves appear to be dangerous.

其他人最终也加入了讨论;OpenAI 的首席科学家实际上说了类似的话。有充分理由相信,至少有一些公司真的相信这类事情,并且(令人恐惧地)许多公司已经为继续从事它们自身似乎具有危险性的技术找到了正当化理由。

But … there is less reason to believe that people at these companies have broad expertise about how the world works beyond their technical expertise.

但是……人们相信这些公司的员工除了技术专长外,还对世界运作方式拥有广泛的专业知识,这种理由就少得多了。

Indeed, there is good reason to doubt that so-called Doomers have any clue what they are doing in terms of enacting change in the real world, or even in thinking through the consequences of their own actions. They are fabulous at getting publicity (the Yudkowsky/Soares bestseller, Bostrom’s bestseller, Coxon’s thread, the Pause letter, and so on). And they are fabulous at getting truly massive funding, perhaps close to (if not over) a billion dollars by now.

确实,有充分的理由怀疑所谓的“末日论者”在推动现实世界变革方面毫无头绪,甚至对自己行为后果的推演也一塌糊涂。他们非常擅长获取公众关注(如尤德洛夫斯基/索雷斯的畅销书、博斯特罗姆的畅销书、科克森的帖子、暂停信等)。而且他们也非常擅长获得巨额资金,迄今为止可能已接近(甚至超过)十亿美元。

But in my view, the net effect of the last decade of “Doomer” talk (Coxon appears to be a card-carrying member of that party) has been to make the Frontier AI companies wealthier and more powerful. In so doing they have also aided and abetted a concentration of power that is in itself of highly dangerous. Everytime they scream that AI is going to kill us all, VCs toss in billions more. None of the AI safety work Doomers have sponsored has led to a demonstrated, implemented solution to the core problem. Enabling OpenAI and Anthropic may prove to be one of the worst things humanity has ever done, and they played a big role in that with their constant drama.

但在我看来,过去十年间“末日论”言论的总体影响(科克森似乎是该阵营的正式成员)是使前沿人工智能公司变得更加富有和强大。在此过程中,他们还助长了一种本身就极具危险性的权力集中。每次他们尖叫着说人工智能将杀死所有人时,风险投资家就会再投入数十亿美元。末日论者资助的所有人工智能安全工作均未带来针对核心问题的已验证、已实施的解决方案。赋能 OpenAI 和 Anthropic 可能被证明是人类有史以来犯下的最糟糕的错误之一,而他们通过不断的炒作在其中扮演了重要角色。

Furthermore, Doomers seem to constantly be selling humanity short, as I will discuss below. And (as also discussed below) they also seem to know not the slightest thing about how war works—which matters since they are implicitly or explicitly imagining a war-to-beat-all-wars between humans and machines.

此外,正如我将在下文讨论的那样,末日论者似乎一直在低估人类的潜力。而且(如下文也将讨论的那样),他们对战争如何运作似乎一无所知——这一点至关重要,因为他们隐含或明确地设想了人类与机器之间一场终极战争的爆发。

Even if you believed that machines were trying to take over the world (which I don’t, at least not at present nor anytime soon), and that they were superintelligent (also not true yet, though presumably true eventually), and even if you suspended disbelief about unplugging the machines, you still have to think through how it is that the combination of motive and superintelligence would lead to the complete annihilation of the human species.

即使你相信机器正试图接管世界(我并不这么认为,至少目前及近期不会如此),并且相信它们拥有超级智能(这也尚未实现,尽管未来可能成真),即使你暂时搁置对拔掉机器电源这一行为的质疑,你仍然需要深入思考:动机与超级智能的结合究竟如何导致人类物种的彻底灭绝。

I have never, ever seen anyone from that world address that latter question with nuance and sophistication. Certainly Coxon has (so far) not done so.

我从未见过来自那个世界的任何人以细致入微且复杂的方式回答后一个问题。当然,科克森到目前为止也没有做到。

§

§

The best attempt I have seen, such as it is, is a book that was wildly popular a year ago, Eliezer Yudkowsky and Nate Soares’ If Anyone Builds It, Everybody Dies.

我所见过的最佳尝试,姑且不论其优劣,是一本一年前广受欢迎的书:埃利泽·尤德洛夫斯基和内特·索雷斯的《如果任何人构建它,所有人都会死亡》。

The book has a lot going for it, both as a work of entertaining science fiction, and as a call to arms to get people to take AI risk seriously.

这本书有很多优点,既作为一部引人入胜的科幻小说,也作为呼吁人们认真对待人工智能风险的战鼓。

But the part in which it paints how AI might actually extinguish humanity is weak, menadering, and unconvincing. The New York Times Book Review dismissed it as being like scientology. I was kinder in the Time Literary Supplement, and much less dismissive, taking the details of the book far more seriously, but in the end the book’s argument was flaweed, for a multiple reasons.

但书中描绘 AI 可能真正灭绝人类的部分薄弱、含糊且缺乏说服力。《纽约时报书评》将其斥为类似于科学教(Scientology)。我在《时代文学增刊》中的评价更为温和,远非全盘否定,而是对书中的细节给予了更多严肃对待,但最终认为该书的论证存在缺陷,原因有多重。

Here are some excerpts of from what I wrote then, a year ago, all still quite relevant.

以下是我当时一年前所写文章的一些摘录,至今仍然相当相关。

The first sums up the Yudkowsky/Soares argument. I presume that Coxon (who hasn’t made his argument explicit as far as I know) is relying on something similar:

第一段总结了尤德科夫斯基/索雷斯的论点。我推测科克森(Coxon)(据我所知他并未明确阐述其论点)依赖的是类似的东西:

The bad news is that Premise 1 (sorry I can’t handle the British spelling) is almost certainly true, though I seriously doubt it will happen by 2030. (Lots of well-known AI researchers like Rich Sutton and Yann LeCun and probably Ilya Sutskever and Fei-Fei Li would agree with me there; we all think that the field needs more breakthroughs before we reach superintelligence.)

坏消息是前提 1(抱歉我无法处理英式拼写)几乎肯定是真的,尽管我严重怀疑它会在 2030 年之前发生。(许多知名 AI 研究人员如里奇· Sutton、杨立昆以及很可能包括伊利亚·苏茨克弗和李飞飞都会同意我的观点;我们都认为在达到超级智能之前,该领域需要更多的突破。)

That’s already reason for Anderson Cooper to breathe a sigh of relief; we have at least a bit more time before superintelligence arrives than Coxon seemed to allow. (Note that Transformers took 9 years to reach the current state; even if there a new breakthrough now it might take years to fully develop. And we probably need more than one, as I will discuss in a future essay.)

这已经是安德森·库珀可以松一口气的理由了;在超级智能到来之前,我们至少比科克森似乎允许的时间要多一点。(请注意,Transformer 花了 9 年时间才达到当前状态;即使现在出现新的突破,完全开发出来可能也需要数年。而且我们可能需要不止一个突破,正如我将在未来的文章中讨论的那样。)

The good news for humanity is that the second and third premises of the Yudkowosky/Soares argument are far, far more shaky. Quoting what I said last year in TLS:

对人类来说好消息是,尤德科夫斯基/索雷斯论点的第二和第三个前提要动摇得多得多。引用我去年在 TLS 所说的话:

[Note that the Open AI Hugging Face incident was provoked as part of a training exercise, with guard rails partly turned off, and not something that happened organically and spontaneously.]

[注意,OpenAI 与 Hugging Face 的事件是在训练演习中引发的,部分安全护栏被关闭,并非有机自发发生的情况。]

If the second premise seems unrealistic with respect to machines, the third seems unrealistic with respect to humans:

如果第二个前提相对于机器显得不切实际,那么第三个前提相对于人类则显得不切实际:

I sent the review to Yudkowsky, but never heard back. So far I know all my points about human resistance still stand.

我把评论发给了尤德科夫斯基,但从未收到回复。到目前为止,我知道我关于人类抵抗的所有观点仍然成立。

Certainly Coxon has not (thus far) added anything substantive exto the argument.

当然,科克森迄今为止没有为该论点增添任何实质性的内容。

In my view, the chance that AI will eliminate humans in the next five years is all but indistinguishable from zero. Humans are too geographically spread out, too genetically diverse, and too resourceful to simply fall apart altogether. The idea that AI will kill us all in five years is preposterous.

在我看来,AI 在未来五年内消灭人类的可能性几乎为零。人类在地理上分布过于广泛,基因多样性过高,且资源利用能力太强,不可能简单地彻底崩溃。认为 AI 会在五年内杀死所有人的想法是荒谬的。

And no, to answer a comment I got on X when I raised these issues, an AI-generated virus is not likely to kill literally all of us, either. When I asked the virologist about this, she told me “i mean, never say never but the only virus I know that has near total mortality is rabies, and considering it’s not airborne and takes 2-4 weeks to kill you, I don’t think it would take out all of humanity. Ebola doesn’t have 90% mortality like the first outbreak did. It’s closer to 50-60%. Still extremely high but not civilization ending. A pandemic could wipe us out if civilization collapsed...eventually. I’d guess anywhere north of 20-30% mortality would do that. But most pandemic viruses are not that lethal. Even the plague wasn’t that lethal.” In her view, we should be a lot more worried about how AI could cripple civilization than freaking about imaginary AI-created viruses.

而且,不。为了回答我在 X(原推特)上提出这些问题时收到的一条评论,AI 生成的病毒也不太可能真的杀死我们所有人。当我向一位病毒学家询问此事时,她告诉我:“我的意思是,永远别说‘永不’,但我知道的死亡率接近 100% 的病毒只有狂犬病。考虑到它不是通过空气传播,且需要 2-4 周才能致死,我认为它不会灭绝全人类。埃博拉病毒的死亡率不像首次爆发时那样高达 90%,更接近 50-60%。这仍然极高,但不足以终结文明。如果文明崩溃……最终大流行病可能会让我们灭绝。我猜测任何高于 20-30% 的死亡率都可能导致这种情况。但大多数大流行病毒并没有那么致命。即使是鼠疫也没那么致命。”在她看来,我们更应该担心的是 AI 如何使文明陷入瘫痪,而不是为虚构的 AI 制造病毒而惊慌失措。

§

§

Which is not at all to say we are home free.

这绝不是说我们就高枕无忧了。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

关联信息,但可能不是同一事件