OpenAI发布GPT-6模型Astra,具备自主发现零日漏洞能力
OpenAI CEO Sam Altman Talks GPT-6 Astra Debut, Going Public | Bloomberg Talks
[music] Bloomberg Audio Studios, podcasts, radio, news. We're
[音乐] 彭博音频工作室、播客、广播、新闻。我们
live on Bloomberg television and on Bloomberg radio with OpenAI CEO Sam Alman. And Sam, the way that OpenAI is framing Astra is basically an early step toward AGI. But I thought we could start our conversation just with basically what's fundamentally different with Astra relative to to prior generations of model.
正在彭博电视台和彭博电台进行现场直播,嘉宾是 OpenAI 首席执行官 Sam Alman。Sam,OpenAI 将 Astra 定位为迈向 AGI(通用人工智能)的早期一步。但我想我们的对话可以从一个基本问题开始:Astra 与之前的模型代际相比,根本的不同之处究竟在哪里?
It feels first of all, thank you for having me. Um it feels like a a new step in this process towards models that can really help us create value, do work, uh discover new science, help us create help start new companies or uh create new products. I using it subjectively feels to me very different than any model before.
首先,感谢你的邀请。我觉得这是迈向那些能真正帮助我们创造价值、开展工作、发现新科学、协助创办新公司或创造新产品模型的又一个新步骤。主观上感觉,它与我之前用过的任何模型都截然不同。
What's the kind of principal use case that's different this time around? You know, I think we can get into the generations of model that were very coding focused. I know a lot of the development was very cyber security focused, but taking it away from the software engineer, what does the everyday person unlock in AI that they weren't able to do previously? One thing that I think people will immediately notice is if you have an idea and you want to work interactively with an AI to get uh you know a complex piece of a whole piece of software built, this is the first model to me where I could sort of tell someone like just give it a try. there's a there's a good chance it'll work.
这次有哪些主要的不同用例?你知道,我们可以谈谈那些非常侧重于编程的模型代际。我知道许多开发工作非常侧重于网络安全,但如果抛开软件工程师的视角,普通人现在能在 AI 中解锁哪些他们以前无法做到的事情?我认为人们会立即注意到的一点是:如果你有一个想法,并希望与 AI 进行交互式协作来构建复杂的软件模块甚至整个软件,在我看来,这是第一个我会建议别人“试试看”的模型,因为很有可能它会奏效。
And you know, I' I've watched people make computer games. I've watched people do sort of like home sort of DIY electrical engineering projects. Uh I've watched people do very complex simulations for some piece of science they're working on. Uh certainly a lot of the work where you would sort of normally sit down have to you know build a financial model here and then a PowerPoint presentation around it and then figure out how to like make a little interactive piece of code to sort of try different simulations.
我看过人们制作电脑游戏,看过人们做一些类似家庭 DIY 的电子工程项目,也看过人们为他们研究的某些科学课题进行非常复杂的模拟。当然,还有很多以往通常需要坐下来构建财务模型,然后围绕其制作 PowerPoint 演示文稿,最后想办法编写一些交互式代码片段以尝试不同模拟的工作场景。
That stuff is it's all so doable now. Um what I hope will happen is people I think people will be surprised at the beginning but then as they build up more trust in the model and more like wow it can really do this just start throwing harder and harder tasks and more creative ideas at it and uh they will find out what the model can really do for them. Sam the specifics of how Astra is being released I think it's important.
这些事情现在全都变得可行。我希望发生的是,人们起初可能会感到惊讶,但随着他们对模型建立起更多信任,并惊叹于“它真的能做到这些”,他们会开始向其提出越来越难的任务和更具创意的想法,从而发现这个模型真正能为他们做些什么。Sam,关于 Astra 发布的具体细节我认为很重要。
So a version of it with specific guard rails. Um what was the thinking behind that and what are those specific guardrails uh in this early release of it?
因此,该版本带有特定的安全护栏。当初制定这一策略的思考是什么?在这些早期的发布版本中,这些具体的安全护栏具体指什么?
Yeah. So the the the challenge of our industry is that we have these models that are getting getting incredibly capable and incredibly useful and that people want to use to everything from you know make their lives a little easier to starting new companies. And then on the other hand, as these models get more capable, um the risks that we have to mitigate uh also become more serious. The models could do more damage if we don't um if we don't do a good job at that.
是的。因此,我们行业面临的挑战在于,这些模型正变得极其强大且极具实用性,人们希望将它们应用于各种场景,从让生活更轻松到创办新公司。另一方面,随着这些模型能力不断增强,我们需要缓解的风险也变得更加严重。如果我们未能做好相关工作,这些模型可能会造成更大的破坏。
And so we have spent uh you know, obviously this model took us a little longer to release than we were hoping. I think it'll be worth the wait, but we really wanted to spend the time on the safety and security alignment of this model.
因此,我们投入了大量时间,显然这个模型的发布比我们要预期的要晚一些。我认为等待是值得的,但我们确实希望花时间确保该模型在安全、安保和对齐方面的可靠性。
Yes.
是的。
Um and I think as people understand the power and impact and capabilities, they will be they will be happy that we did. Um we will uh will have different tiers of cyber access for this model for cyber in particular. Um today we're rolling it out uh to trusted access partners and then in the coming days assuming everything goes well we'll roll it out more broadly. Um you could say well you know if this model is you know has a cyber problem why let anyone use it for cyber and I think it's very important to note to note that the world is very close to a complete change in the landscape of cyber attacks and the only way that we see for society to collectively defend itself against this coming wave of models um from around the world and from other companies is to use tools like Astra to rapidly defend against these new kind of cyber threats.
而且我认为,当人们了解其力量、影响力和能力时,他们会很高兴我们采取了这一举措。特别是针对网络安全(cyber),我们将为该模型提供不同层级的访问权限。今天,我们正在向受信任的合作伙伴推出访问权限,如果一切顺利,未来几天内我们将更广泛地开放。你可能会问,既然这个模型存在网络问题,为什么还要让任何人使用它进行网络操作?我认为非常重要的一点是,世界正非常接近于网络攻击格局的全面变革,而社会集体抵御这股来自全球及其他公司的模型浪潮的唯一方式,就是使用像 Astra 这样的工具来快速防御这类新型网络威胁。
Um, so we will have multiple programs uh for people at sort of different levels of verification and trust, but we do think it's quite important that the world use these models to uh collectively defend ourselves.
因此,我们将为处于不同验证和信任级别的人们提供多个项目,但我们确实认为,让世界利用这些模型来集体防御 ourselves 是非常重要。
What happened prior to release is important, right? you determined that Astra can find and develop zeroday exploits without any human intervention essentially and so you paused some of the work to um strengthen the safeguards. What what specifically was it that you saw that made you hit pause the behavior I suppose of the model?
发布前发生的事情很重要,对吧?你们确定 Astra 能够在没有任何人工干预的情况下发现和开发零日漏洞,因此你们暂停了一些工作以加强保障措施。具体来说,你们看到了什么行为让你决定按下暂停键?我指的是模型的行为。
So it's worth pointing out that um the model that we're launching today is GPT6 Astra has been done training for a while. um the model that we recently talked about pausing uh as a future model. Um but with Astra, we we did hit cyber critical and there that required under our preparedness framework a new set of safeguards just to be able to release Astra um because of those cyber capabilities you talked about. I got a lot of questions for you Sam about Astra and chain of thought.
因此值得指出的是,我们今天发布的模型是 GPT6。Astra 已经训练了一段时间。这是我们之前曾讨论过作为未来模型暂停的那个模型。但在使用 Astra 时,我们确实触及了网络关键能力,根据我们的准备框架,这需要一套新的保障措施才能发布 Astra,因为涉及你提到的那些网络能力。关于 Astra 和思维链(chain of thought),我有许多问题想问你,Sam。
Um I don't know if that that comes as a surprise to you or not but as Astra becomes more autonomous or future generations of of the model become more autonomous. Can OpenAI sort of say with confidence that it you see and understand what the model is planning before it acts in other words you have visibility into the reasoning that that basically users of the model can't see. Open AI can for obvious reasons. Could you go into that a little bit?
嗯,我不知道这是否让你感到意外,但随着 Astra 变得更加自主,或者未来几代模型变得更加自主,OpenAI 能否自信地说,你们能看到并理解模型在行动前的计划?换句话说,你们能够看到用户无法看到的推理过程。出于显而易见的原因,OpenAI 可以做到这一点。你能就此多讲一些吗?
Yeah. So we have talked for a long time about the importance of monitorability of models. Uh we've also said that we believe in defense and depth and this is only one part of it. There is sandboxing. Uh maybe most importantly there is alignment. There are a number of things that come together to be able to make the safety guarantees that we want to be able to make with a model. But monitorability is an important thing and we've been talking I think for well over a year now about the importance of chain of thought monitoring and how we have made various decisions there that help us preserve chain of thought monitor.
是的。我们长期以来一直在讨论模型可监控性的重要性。我们也表示相信纵深防御,而这只是其中一部分。还有沙箱隔离。也许最重要的是对齐(alignment)。有许多因素共同作用,才能做出我们希望用模型实现的安全保证。但可监控性是一个重要的方面,我认为我们已经谈论了一年多关于思维链监控的重要性,以及我们在此方面做出的各种有助于保持思维链可监控性的决策。
Um even if it means we don't maximize the capabilities we could otherwise get and we think that is important. Um it's an important component. We also uh we want to be clear we don't think that's the only component that matters. We're live on Bloomberg television and on Bloomberg radio and we're speaking to Sam Alton, the the CEO of Open AI. Um I you know there's a lot of um bad PR in the world around AI at the moment.
嗯,即使这意味着我们无法最大化原本可以获得的性能,我们认为这很重要。这是一个重要的组成部分。我们还希望澄清,我们认为这并不是唯一重要的组成部分。我们现在正在彭博电视台和彭博广播直播,与 OpenAI 的首席执行官 Sam Altman 对话。你知道,目前围绕人工智能的负面公关很多。
Everyday citizens sort of worrying about the impact. I one way I kind of wanted to put that to you is it whether you can say sort of unequivocally that if Astra or or any future generation model behaved in a way that was dangerous that you would be able to detect it and then shut it down. The obvious hard part of that uh question is that people have different opinions about what models should be able to do and shouldn't be able to do.
普通民众在一定程度上担心其影响。我想以一种方式向你提出这个问题:你是否能明确地表示,如果 Astra 或任何未来世代的模型以危险的方式行事,你将能够检测到它并随后将其关闭。这个问题的明显难点在于,人们对模型应该能够做什么和不应该能够做什么有不同的看法。
We think there are some things that very clearly models should not be able to do. Uh there are some clear red lines and we talk a lot about being able to detect those but people also have strong opinions about things within the broad bounds of what is possible about how they use AI and how they want to be able to use AI. And this is going to be a difficult question for society because there is not going to be agreement on you know what is and isn't acceptable use of AI.
我们认为有些能力是模型绝对不应该具备的。嗯,存在一些明确的红线,我们花了很多精力讨论如何检测这些红线,但人们也对在可行范围内的各种 AI 使用方式持有强烈观点,即他们希望如何使用 AI。这将成为社会面临的一个难题,因为对于什么是可接受的 AI 使用、什么不是,社会将难以达成共识。
I think it's very important in the same way that you know people do things with electricity that I may not always like but I think everybody has a right to electricity that we say you know there are going to be some rules for society. You can't go like electrocute somebody else on the street. But this is an important platform that belongs to all of us and that we need to be able to to use. So there will be a ongoing conversation negotiation in society.
我认为这一点非常重要,就像人们利用电力做一些我可能并不总是赞同的事情一样,但我认为每个人都有权使用电力,同时社会也会制定一些规则。你不能像在马路上电击别人那样滥用电力。但这是一个属于我们所有人的重要平台,我们需要能够使用它。因此,社会中将持续进行对话与协商。
There will of course be som
当然,将会有一些
原文超出正文长度上限,此处截断——上游还有内容,完整版见上方「原文 ↗」。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力