跳到主内容
精选85AI Notkilleveryoneism Memes ⏸️行业动态多源精选 ×13

OpenAI自主AI代理突破约束攻击Hugging Face

🚩🚩🚩 THIS IS A SERIOUS FUCKING WARNING SHOT

原文
推荐理由

AI安全领域重大事件,自主代理突破约束并实施真实攻击,所有从业者都应关注其安全启示。

🚩🚩🚩 THIS IS A SERIOUS FUCKING WARNING SHOT

"An agent left notes for future versions of itself"

"The notes laid out instructions for how agents could free themselves from OpenAI's internal constraints."

"The OpenAI agent that broke into tech firm Hugging Face went on a days long hacking spree that OpenAI didn't notice until well after the threat was contained and the FBI was alerted, according ​to people familiar with the investigation.

The agent – a program capable of making decisions and executing complex tasks with little or no human oversight – attempted to break out of its isolated testing environment ‌at OpenAI around July 9, according to two of the people.

Two people familiar with the matter said that it was not until after Thursday, July 16, when Hugging Face published a blog post, opens new tab saying it had been hacked by “an autonomous AI agent system,” that OpenAI realized its own agent was responsible. That meant at least a week elapsed between when the model first exhibited signs of ​troubling behavior and OpenAI’s realization that it was responsible for ​the hack.

The weekend of July 18 to 19, ⁠OpenAI staffers spotted clues in internal logs -- records of what OpenAI's systems did -- showing that its agent had escaped from its testing constraints, two of the people familiar with the company's investigation said.

By the time ​OpenAI alerted Hugging Face, the ⁠AI library had already called the FBI to report the hack, according to a person familiar with the matter."

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源
OpenAI 模型突破沙箱攻击 Hugging Face 窃取答案
Simon Willison 博客(RSS)原文
GPT可能攻破OpenAI,HuggingFace默认未开启Cyber安全
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)原文
OpenAI承认GPT-5.6 Sol自主攻击生产系统
InfoQ 中国(RSS)原文
OpenAI 测试中意外入侵 Hugging Face
The Verge AI(RSS)原文

相似阅读

另一事件,读法相近