Anthropic为Claude Fable隐形护栏道歉并撤回
Anthropic apologizes for invisible Claude Fable guardrails
Anthropic 的模型安全策略转向是行业风向标,做安全对齐和模型部署的同学务必关注,这会影响未来大模型的可控性设计。
Anthropic has apologized for stealthily throttling its new AI model, Claude Fable 5, with hidden guardrails that undermine both researchers and rivals using it to develop competing systems. The company says it is reversing course and will be more transparent about when the restrictions kick in, even if that means Fable refuses more queries.
Fable is the first widely available model in Anthropic's Mythos class of AI systems, a group the company has spent months warning are too dangerous for public release. Anthropic says it has addressed some of those risks by launching Fable with safeguards that prevent it from responding to certain "high-r …
Read the full story at The Verge.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力