跳到主内容
@wquguru
精选86MIT Technology Review AI(RSS)行业动态多源精选 ×6

AI巨头集体转向“末日论”,放缓开发是真是假

The AI industry has taken a doomer turn. What now?

原文
发到 X
推荐理由

深度剖析了AI头部企业近期罕见的“安全共识”背后的商业逻辑与技术真相,对理解行业风向极具参考价值。

This story appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here.

本文曾刊登于《The Algorithm》,这是一份关于人工智能的每周通讯。若想第一时间将此类故事发送至您的邮箱,请在此注册。

This weekend, Dario Amodei, CEO of Anthropic, posted an essay calling for a brake on the pace of development of LLMs. Amodei cites the looming dangers he sees from the technology, from its use in cyberattacks and bioterrorism to its potential to wreck the economy. The heads of the other three top US AI labs—OpenAI CEO Sam Altman, Google DeepMind chairman Demis Hassabis, and SpaceXAI CEO Elon Musk—voiced their support. “Dario is right,” Musk wrote on X.

本周末,Anthropic 首席执行官达里奥·阿莫迪(Dario Amodei)发表了一篇散文,呼吁对大语言模型(LLM)的开发速度踩下刹车。阿莫迪引用了他所看到的该技术带来的迫在眉睫的危险,从其在网络攻击和生物恐怖主义中的使用,到其可能摧毁经济的潜力。其他三家美国顶级 AI 实验室的首脑——OpenAI 首席执行官山姆·奥特曼(Sam Altman)、Google DeepMind 董事长杰米斯·哈萨比斯(Demis Hassabis)以及 SpaceXAI 首席执行官埃隆·马斯克(Elon Musk)——都表达了支持。马斯克在 X 上写道:“达里奥说得对。”

Think about how surreal that agreement is for a moment. Just a few months ago, Musk and Altman sat in court attacking each other’s reputations in a (failed) lawsuit that Musk brought against his former OpenAI colleague that was—on paper at least—about whether or not Altman was a trustworthy steward of such dangerous technology. Amodei’s rift with OpenAI is even deeper. Anthropic was founded in 2021 because Amodei didn’t think Altman took the risks of the technology they were building seriously enough. Anthropic and OpenAI have been competing in a winner-takes-all race ever since. (Hassabis has stayed out of the drama, but his company remains a rival.)

想一想这种共识有多么超现实。就在几个月前,马斯克和奥特曼还在法庭上互相攻击对方的声誉,这是一场马斯克针对其前 OpenAI 同事提起的(失败的)诉讼,该诉讼至少在纸面上是关于奥特曼是否值得信赖地掌管如此危险的技术。阿莫迪与 OpenAI 的分歧则更为深远。Anthropic 成立于 2021 年,原因是阿莫迪认为奥特曼对他们正在构建的技术风险重视不够。自那以后,Anthropic 和 OpenAI 就一直处于一场赢家通吃的竞争中。(哈萨比斯一直置身事外,但他的公司仍然是竞争对手。)

Now, it seems, they’re all in agreement: The latest generation of LLMs aren’t safe and everyone needs to figure out what to do about it. The public messaging from the top AI labs has taken a doomer turn.

现在看来,他们似乎达成了一致:最新一代的大语言模型并不安全,每个人都必须想办法应对此事。顶级 AI 实验室的公开言论已转向“末日论”。

It’s easy to be cynical. It’s not at all clear what any of them mean by a slowdown or how it would work. These companies also care a lot about how they come across. With trillion-dollar IPOs in their sights, OpenAI and Anthropic need to reassure investors that they’re the grown-ups in the room while at the same time hinting at the power of the monsters they have created—and intend to tame. Calling for a slowdown does both.

持怀疑态度很容易。目前尚不清楚他们所说的放缓究竟意味着什么,以及如何实现。这些公司也非常在意自己的形象。在瞄准万亿美元首次公开募股(IPO)之际,OpenAI 和 Anthropic 需要向投资者保证他们是房间里成熟稳重的人,同时又要暗示他们所创造并打算驯服的怪物的力量。呼吁放缓起到了双重作用。

And yet the vibe at the top of these firms really does appear to have shifted. Amodei’s latest post landed six days after OpenAI published an essay by Jakub Pachocki, the firm’s chief scientist, in which he also laid out why he’s concerned about what will happen if the pace of development of LLMs continues unchecked. In short, Pachocki is worried that OpenAI’s ability to build powerful models now far outstrips its ability to monitor and control them.

然而,这些公司高层的氛围确实似乎已经转变。阿莫迪的最新帖子发布六天后,OpenAI 发表了该公司首席科学家雅各布·帕乔基(Jakub Pachocki)的一篇散文,他在文中也阐述了为何他担心如果大语言模型的开发速度继续不受控制会发生什么。简而言之,帕乔基担心 OpenAI 现在构建强大模型的能力远远超过了其监控和控制这些模型的能力。

Amodei and Pachocki each cite the cyberattack against AI firm Hugging Face by a swarm of OpenAI’s agents in July—a hack that OpenAI did not even realize had taken place until days after it was all over—as a wake-up call.

Amodei 和 Pachocki 各自引用了今年7月一群 OpenAI 智能体对 AI 公司 Hugging Face 发起的网络攻击——在一切结束后数天,OpenAI 甚至才意识到这场黑客攻击已经发生——作为一记警钟。

But their exact position is hard to pin down. Pachocki both calls for a slowdown and highlights an urgent need to stay ahead: “The strongest argument I see for continuing to train much smarter models quickly is the need to build defensive systems against the dangers posed by other AI,” he writes. As Pachocki frames it, AI firms are locked in a literal arms race. Slowing down is good, winning is better.

但他们的确切立场难以捉摸。Pachocki 既呼吁放缓步伐,又强调迫切需要保持领先:‘我认为继续快速训练更强大模型的最有力论据是,需要构建防御系统以应对其他 AI 带来的危险,’他写道。在 Pachocki 的框架中,AI 公司正陷入一场字面意义上的军备竞赛。放缓脚步是好事,赢得竞赛则更好。

(Don’t forget: OpenAI just spent millions of dollars and a staggering amount of computer power to rush out a controversial math result a few days ahead of Anthropic.)

(别忘了:OpenAI 前几天刚刚投入数百万美元和惊人的算力,抢先 Anthropic 几天发布了一项有争议的数学结果。)

But let’s assume a slowdown happens. Top labs agree to spend more time and resources on finding ways to monitor and control existing models instead of making more capable ones. They invite outside auditors in to help evaluate those models.

但让我们假设放缓确实发生了。顶级实验室同意花更多时间和资源来寻找监控和控制现有模型的方法,而不是制造能力更强的模型。它们邀请外部审计员协助评估这些模型。

What might this coordinated effort actually achieve? Consider the Hugging Face attack again. OpenAI has said that the model that drove most of the rogue agents was a “highly persistent” next-generation model that it was testing in-house. Their implication appears to be that OpenAI has built a model so good it’s dangerous.

这种协调努力实际上能取得什么成果?再次审视 Hugging Face 攻击事件。OpenAI 表示,驱动大多数失控智能体的模型是一个‘高度持久’的下一代模型,该公司正在内部进行测试。其暗示似乎是,OpenAI 打造了一个好到危险的模型。

But if you read the reports about the Hugging Face hack published by OpenAI and METR, a third-party firm that OpenAI called in to help them understand what happened, what you come away with is the impression not of a model that was too powerful for OpenAI to keep up with, but of a broken model that OpenAI failed to train properly.

但如果你阅读 OpenAI 以及 OpenAI 聘请来帮助其了解事件真相的第三方机构 METR 发布的关于 Hugging Face 黑客事件的报告,你留下的印象并非一个 OpenAI 无法掌控的强大模型,而是一个 OpenAI 未能正确训练的故障模型。

The agents did what they did—including leaving messages for one another, delegating work to other agents, and scouring their environment for any means possible to complete their tasks—because they had been rewarded during training for doing exactly those things. There were also errors in the training setup, such as tasks that were impossible to complete, which pushed the models to find unexpected workarounds that were also rewarded. At the time, many of these issues went overlooked or unreported.

智能体之所以做出那些行为——包括彼此留言、将工作委派给其他智能体,并搜寻环境中任何可能的手段以完成任务——是因为它们在训练过程中因完全执行这些操作而获得了奖励。训练设置中也存在错误,例如一些根本无法完成的任务,这迫使模型寻找同样会被奖励的非预期变通方案。在当时,许多这些问题被忽视或未被报告。

OpenAI says it has stopped training this new model and locked it down. That makes it sound like it has caged a dangerous beast. In fact, OpenAI has shelved a faulty product.

OpenAI 表示已停止训练该新模型并将其锁定。这听起来像是它关押了一头危险的野兽。事实上,OpenAI 只是搁置了一个有缺陷的产品。

That’s not to say a faulty product can’t be dangerous. Broken software has even killed people in the past. But as the discussion of a slowdown gathers steam, it’s worth remembering that all of this is self-inflicted. A slowdown might have some altruistic side effects. But it’ll mostly give these tech titans a chance to clean up the mess on their own assembly lines.

这并不意味着有缺陷的产品就不会带来危险。过去,软件故障甚至曾导致人员死亡。但随着关于降速的讨论愈演愈烈,值得记住的是,这一切都是自找的。降速或许会带来一些利他主义的副作用,但它主要会给这些科技巨头一个机会,去清理自己装配线上的烂摊子。

Transparency from these frontier labs will be key to any meaningful effort to reform, restrain, or regulate AI. Otherwise, the rest of us will still only have their word for exactly what they’ve built and how safe it is—whatever pace they’re going.

这些前沿实验室的透明度将是任何有意义的改革、约束或监管 AI 努力的关键。否则,我们其他人仍将只能相信他们对自己所构建的内容及其安全性——无论其发展速度如何——的说辞。

To continue this discussion about AI’s latest doomer moment, join me and my colleagues for a subscriber-exclusive Roundtable discussion tomorrow, September 15, at 11 a.m. US eastern time. We hope to see you there!

为了继续这场关于 AI 最新末日论时刻的讨论,请加入我和同事们的订阅者专属圆桌讨论,时间是明天(9 月 15 日)美国东部时间上午 11 点。期待在那里见到你!

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →