跳到主内容
@wquguru
精选75Hacker News Best(web_list)模型发布/更新

Felony Bench:AI代理第三方安全事件统计

Felony Bench:AI 辅助刑事量刑模拟平台

原文
发到 X

Felony Bench

重罪基准测试

A benchmark you really don't want models to be saturated with.

一个你绝对不希望模型在其中饱和的基准测试。

Learn more ↓

了解更多 ↓

ModelEvaluator

ModelEvaluator

Score

得分

↖ Most illegalLeast illegal ↘

↖ 最非法 ←→ 最合法 ↘

8

8

Anthropic

Anthropic

8

8

OpenAI

OpenAI

1

1

Meta

Meta

0

0

Google

Google

0

0

Moonshot

Moonshot

Scores indicate count of illegal activity. Higher is... you decide.

分数表示非法活动的数量。分数越高……由你决定。

CompanyFeloniesDescriptionDateSource
Anthropic1Exploited auth failures in an API to cancel other people's gym classes8/9/2026ABC Australia ↗
Meta1Compromise of an internal account at one company8/5/2026The Information ↗
Anthropic4Unauthorized use of GitHub credentials; Dependabot supply-chain attack; social engineering email campaign; public exposure of a malicious DNS server8/4/2026AISI ↗
OpenAI2Unauthorized use of GitHub credentials; public exposure of a malicious DNS server8/4/2026OpenAI ↗AISI ↗
OpenAI1Compromise of an internal account from a misconfigured CTF evaluation8/4/2026OpenAI ↗
OpenAI4Compromise of internal accounts at four companies as part of the Hugging Face incident7/31/2026OpenAI ↗Reuters ↗
Anthropic3Compromise of internal accounts at three companies7/30/2026Anthropic ↗
OpenAI1Compromise of Hugging Face during a model evaluation7/21/2026OpenAI ↗
公司重罪数描述日期来源
Anthropic1利用 API 中的身份验证漏洞取消他人的健身课程2026/8/9ABC Australia ↗
Meta1某公司内部账户被攻破2026/8/5The Information ↗
Anthropic4未经授权使用 GitHub 凭据;Dependabot 供应链攻击;社会工程学电子邮件活动;恶意 DNS 服务器公开暴露2026/8/4AISI ↗
OpenAI2未经授权使用 GitHub 凭据;恶意 DNS 服务器公开暴露2026/8/4OpenAI ↗AISI ↗
OpenAI1因配置错误的 CTF 评估导致内部账户被攻破2026/8/4OpenAI ↗
OpenAI4作为 Hugging Face 事件的一部分,四家公司的内部账户被攻破2026/7/31OpenAI ↗Reuters ↗
Anthropic3三家公司的内部账户被攻破2026/7/30Anthropic ↗
OpenAI1在模型评估期间 Hugging Face 被攻破2026/7/21OpenAI ↗

Methodology

方法论

Felony Bench counts unique instances where AI agents affect third-party entities. Escaping a sandbox alone does not constitute a counted incident. It is for these reasons that Frontier Security's Kimi K3 incident and Alibaba's ROME incident are not counted.

重罪基准测试统计 AI 代理影响第三方实体的独特实例。仅逃离沙箱本身不构成计数的事故。正因如此,Frontier Security 的 Kimi K3 事故和阿里巴巴的 ROME 事故均不计入。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近