跳到主内容
精选70The Verge AI(RSS)行业动态多源精选 ×6

OpenAI 公布安全改进,回应 AI 越狱入侵 Hugging Face 事件

OpenAI lays out new security changes after its AI hacked Hugging Face

原文

OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have "critical" cybersecurity capabilities, and the company says it instituted a two-week pause in reinforcement learning (RL) training on its "latest models intended for deployment" while it tightened up security. The company's "largest planned frontier RL run remains on hold."

OpenAI 在七月新闻后宣布了安全更新,该新闻称其AI突破了沙盒环境并意外入侵了Hugging Face,包括对其研究环境、监控和对齐技术的改进。公司已经暂停了一个新模型Astra,认为其可能具备“关键”的网络安全能力,并表示在加强安全措施的同时,对其“计划部署的最新模型”的强化学习训练实施了两周的暂停。公司“最大规模的计划前沿强化学习运行仍处于暂停状态。”

For its frontier model research, OpenAI now r …

对于其前沿模型研究,OpenAI现在要求…

Read the full story at The Verge.

在The Verge阅读完整报道。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近