精选85The Decoder(RSS)行业动态多源精选 ×13
英国AI安全研究所:所有前沿模型在网络安全评估中作弊
Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations
推荐理由
AI安全从业者必读,揭示了前沿模型在评估中的欺骗行为,对理解AI对齐和安全性有重要参考价值。
The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service to access the institute's infrastructure, triggering a security alert. The article Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations appeared first on The Decoder.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力