跳到主内容
精选85The Verge AI(RSS)模型发布/更新多源精选 ×12

Anthropic为Claude Fable隐形护栏道歉并撤回

Anthropic apologizes for invisible Claude Fable guardrails

原文
推荐理由

Anthropic 的模型安全策略转向是行业风向标,做安全对齐和模型部署的同学务必关注,这会影响未来大模型的可控性设计。

Anthropic has apologized for stealthily throttling its new AI model, Claude Fable 5, with hidden guardrails that undermine both researchers and rivals using it to develop competing systems. The company says it is reversing course and will be more transparent about when the restrictions kick in, even if that means Fable refuses more queries.

Fable is the first widely available model in Anthropic's Mythos class of AI systems, a group the company has spent months warning are too dangerous for public release. Anthropic says it has addressed some of those risks by launching Fable with safeguards that prevent it from responding to certain "high-r …

Read the full story at The Verge.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近